Software Engineer 4/5 – Model Serving Systems, AI Platform

🔥 12 hours ago

🇺🇸 United States – Remote

💵 $466k - $750k / year

⏰ Full Time

🟡 Mid-level

🟠 Senior

🤖 AI Engineer

👻 Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Netflix

Netflix

10,000+ employees

Founded 1997

📱 Media

👥 B2C

Media • B2C

Netflix is a global streaming entertainment company and content producer whose stated mission is "to entertain the world. " It operates a consumer-facing subscription platform offering on-demand TV shows, films, and original programming, and also runs a public careers site emphasizing culture, inclusion, and hiring accommodations. The provided text highlights Netflix’s focus on recruiting talent worldwide, its work-life and culture pages, and its public-facing employer materials.

📋 Description

• Develop and expand compute infrastructure supporting Netflix's growing AI needs • Build scalable, robust systems for AI/ML applications • Provide infrastructure for real-time model inference and serving • Develop foundational abstractions ensuring consistency between online and offline systems • Build model-serving infrastructure for LLMs and other large foundation models • Enable application of machine learning in new business areas • Drive AI/ML innovation across Netflix • Partner cross-functionally with engineers, product managers, machine learning engineers, and data/research scientists • Support business-critical models with high availability and performance • Contribute to production hosting, deployment management, performance tuning, and capacity planning

🎯 Requirements

• Experience building high-traffic distributed services and infrastructure for online ML model inference • Familiarity with supporting large-scale ML models with a focus on high availability and performance • Understanding of scalable model-serving solutions for generative models and LLMs • Skills in reducing latency and costs and solving bottlenecks to streamline research-to-production workflows • Proficiency in object-oriented programming, preferably Java • Engineering excellence in production hosting, including performance tuning, deployment management, and capacity planning • Familiarity with deploying ML models using Triton Inference Server, TensorRT, and Docker • Experience working with public cloud platforms such as AWS, Azure, or GCP • Knowledge of observability and logging best practices • BS/MS in Computer Science, Applied Math, Engineering, or a related field

🏖️ Benefits

• Health Plans • Mental Health support • 401(k) Retirement Plan with employer match • Stock Option Program • Disability Programs • Health Savings and Flexible Spending Accounts • Family-forming benefits • Life and Serious Injury Benefits • Paid leave of absence programs • Full-time hourly employees accrue 35 days annually for paid time off to be used for vacation, holidays, and sick paid time off • Full-time salaried employees are immediately entitled to flexible time off

Apply Now

Similar Jobs

🔥 12 hours ago

6sense

1001 - 5000

💼 Consulting

📦 Logistics

🤖 Artificial Intelligence

Senior Enterprise AI Engineer building secure AI agents and automated workflows for 6sense, a revenue-growth technology company. Improving productivity, compliance, and decision-making across internal business functions.

🔥 13 hours ago

Aimpoint Digital

51 - 200

🤖 Artificial Intelligence

💼 Consulting

Forward Deployed AI Engineer building production AI applications, agents, and copilots for enterprise clients. Aimpoint Digital delivers data, AI, analytics, and operations research solutions.

🔥 22 hours ago

Aline

201 - 500

🏥 Healthcare

☁️ SaaS

🤝 B2B

Software Engineer building AI-powered features and agentic workflows for Aline’s senior-care platform. Integrating LLMs, RAG, and Azure/.NET services into production products.

🕒 Yesterday

Veracity Insurance Solutions, LLC

51 - 200

💼 Consulting

🏥 Healthcare

📦 Logistics

Lead AI Engineer shaping AI architecture, LLMOps, and responsible AI standards for Veracity’s digital insurance platform. Rebuilding insurance processes with production GenAI and agentic systems.

🕒 Yesterday

InsCipher

11 - 50

🛡️ Insurance

☁️ SaaS

🔌 API

Lead AI Engineer modernizing InsCipher’s insurance compliance platform with production GenAI and agentic systems. Defining AI architecture, LLMOps standards, and enterprise engineering practices.