Staff ML Software Engineer – Platform Systems

Job not on LinkedIn

🕒 August 12

🇺🇸 United States – Remote

💵 $600k - $1.1M / year

⏰ Full Time

🔴 Lead

🏗️ Platform Engineer

👻 Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Netflix

Netflix

10,000+ employees

Founded 1997

📱 Media

👥 B2C

Media • B2C

Netflix is a global streaming entertainment company and content producer whose stated mission is "to entertain the world. " It operates a consumer-facing subscription platform offering on-demand TV shows, films, and original programming, and also runs a public careers site emphasizing culture, inclusion, and hiring accommodations. The provided text highlights Netflix’s focus on recruiting talent worldwide, its work-life and culture pages, and its public-facing employer materials.

📋 Description

• Design, build, and operate observability, evaluation, and tooling subsystems for next-generation ML architecture • Prove subsystems on current AIMS operations, including anomaly detection, root cause analysis, and operational automation • Build observability systems providing visibility into model behavior, training pipeline health, serving latency, and data quality • Drive cost optimization across AIMS training and serving infrastructure through increasingly automated frameworks and tooling • Architect reliability improvements across the AIMS AI/ML stack, reducing toil and improving on-call ergonomics • Contribute to the target architecture and migration path for the modernized AIMS AI/ML stack • Coordinate with teams driving the AIMS modernization effort • Evaluate emerging infrastructure patterns, model paradigms, and platform capabilities and translate them into a forward-looking roadmap

🎯 Requirements

• Significant experience designing, building, and operating production AI/ML systems at scale • Experience with training pipelines and familiarity with model serving and online inference under high traffic • Hands-on experience building subsystems for advanced agentic architectures, including memory, trace, eval, and replay pipelines, or orchestration and routing layers • Deep Python expertise • Working proficiency in at least one JVM language: Scala or Java • Proven track record improving AI/ML system reliability, reducing infrastructure costs, and improving operational scalability • Experience building observability and monitoring systems for AI/ML workloads across training, serving, and data pipelines • Strong distributed systems background, including batch processing at scale and real-time serving infrastructure • Experience collaborating with partner teams to drive cross-functional technical programs, manage dependencies, and build consensus without formal authority • High technical judgment and ability to identify patterns, build reusable frameworks, and make pragmatic investment decisions • Ability to operate with incomplete information, scope problems, define approaches, and adjust course • Preferred: familiarity with LLM evaluation, trace, or replay tooling • Preferred: familiarity with feature stores, model serving platforms, and experiment frameworks • Preferred: hands-on experience migrating production AI/ML systems across technology generations • Preferred: applied experience in personalization domains such as recommendation systems, search, or discovery

🏖️ Benefits

• Annual salary-only compensation structure with choice between salary and stock options • Health Plans • Mental Health support • 401(k) Retirement Plan with employer match • Stock Option Program • Disability Programs • Health Savings and Flexible Spending Accounts • Family-forming benefits • Life and Serious Injury Benefits • Paid leave of absence programs • Full-time salaried employees are immediately entitled to flexible time off

Apply Now

Similar Jobs

🕒 August 12

AlphaSense

1001 - 5000

💼 Consulting

🏥 Healthcare

📣 Marketing

Staff Platform Engineer owning scalable AWS, Kubernetes, and AI-enabled infrastructure. Supporting AlphaSense’s AI-powered market intelligence platform and globally distributed engineering teams.

🇺🇸 United States – Remote

💵 $223.1k - $305k / year

💰 Debt Financing on 2022-06

⏰ Full Time

🔴 Lead

🏗️ Platform Engineer

🦅 H1B Visa Sponsor

infoinfo

🕒 August 12

AlphaSense

1001 - 5000

💼 Consulting

🏥 Healthcare

📣 Marketing

Principal Software Engineer shaping scalable Kubernetes microservices for AlphaSense’s AI-powered market-intelligence platform. Architecting cell-based deployments, configuration management, and cloud services across regions and private clouds.

🇺🇸 United States – Remote

💵 $246k - $339k / year

💰 Debt Financing on 2022-06

⏰ Full Time

🔴 Lead

🏗️ Platform Engineer

🦅 H1B Visa Sponsor

infoinfo

🕒 August 11

SAIC

10,000+ employees

☁️ SaaS

📣 Marketing

🏢 Enterprise

Power Platform Developer building secure Power Apps, Dataverse, SharePoint, and Power BI solutions for SAIC’s DCSA defense IT program. Automating government business processes and managing end-to-end solution delivery.

🇺🇸 United States – Remote

🔥 Funding within the last year

💰 $500M Post-IPO Debt - SAIC on 2025-09

⏰ Full Time

🟠 Senior

🔴 Lead

🏗️ Platform Engineer

🕒 August 11

OnePay

501 - 1000

💳 Fintech

🏦 Banking

₿ Crypto

Platform Engineer building Kafka-based distributed infrastructure for OnePay, a consumer fintech platform. Developing AWS cloud services, developer tooling, and agentic application frameworks.

🇺🇸 United States – Remote

💵 $170k - $210k / year

💰 $300M Series unknown on 2025-01

⏰ Full Time

🟠 Senior

🔴 Lead

🏗️ Platform Engineer

🦅 H1B Visa Sponsor

infoinfo

🕒 August 10

Workiva

1001 - 5000

💼 Consulting

🏥 Healthcare

📦 Logistics

Staff Data and AI Platform Engineer building Workiva’s enterprise Snowflake platform. Defining secure, governed infrastructure and AI-agent data foundations for complex organizations.