Machine Learning Engineer – Inference, Serving

Job not on LinkedIn

🕒 November 4, 2025

🇺🇸 United States – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

🤖 Machine Learning Engineer

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Yobi

Yobi

11 - 50 employees

💼 Consulting

📣 Marketing

🤖 Artificial Intelligence

💰 $2.4M Seed Round on 2023-02

Consulting • Marketing • Artificial Intelligence

Yobi is a revolutionary AI platform that augments human capabilities to enhance business communications and productivity. The company's platform integrates seamlessly into existing operational workflows, allowing users to manage all their communications, including SMS, voice calls, and social media messages, from a unified inbox. Yobi utilizes AI-driven features such as sentiment analysis, real-time transcriptions, and automatic language translation to improve customer service and streamline operations. It caters to various industries with tailored AI teammates, offering solutions like virtual receptionists and podcast hosts. Additionally, Yobi ensures seamless integration with popular business applications like Salesforce, Slack, and HubSpot, enhancing enterprise functionality and collaboration through superior human-AI partnerships.

📋 Description

• As a Machine Learning Engineer focused on Inference and Serving at Yobi, you’ll design, optimize, and operate the systems that bring our Behavioral AI models to life in real time. • You’ll work at the core of our production environment, turning trained models into performant, reliable, and continuously improving services that power our open-web and CTV products. • This is an applied ML systems role—equal parts engineering depth, deployment craft, and model intuition. • You’ll shape how models are packaged, versioned, rolled out, and observed across environments, ensuring every prediction is fast, accurate, and accountable.

🎯 Requirements

• Deep expertise in model deployment. You’ve built or scaled production ML serving systems—handling versioning, rollouts, rollback strategies, and live experimentation. • Low-latency mindset. You understand what makes inference fast: model graph optimization, quantization, caching, batching, and efficient feature retrieval. • Systems fluency. You write robust, high-performance code in Go, Rust, C++, or Java, and are comfortable bridging to Python for model integration and analysis. • Operational maturity. You treat inference as a living system—monitoring drift, tracking model lineage, and ensuring observability from input to outcome. • Infrastructure intuition. You know how to make serving systems reproducible and portable without over-engineering them, whether that’s through custom runtime design, model registries, or lightweight orchestration. • Applied ML understanding. You can reason about model performance, interpret trade-offs, and work with researchers to make models more deployable.

🏖️ Benefits

• Competitive Base Salary • Meaningful equity & financial upside - a real % of the company • Annual bonus target based on personal and company performance • Health, Dental, Vision - most plans will pay little to 0 out of pocket • Unlimited PTO - we care about impact, not tracking days you’re out • 401k with company match %

Apply Now

Similar Jobs

🕒 October 30, 2025

Serve Robotics

51 - 200

📦 Logistics

🍽️ Food & Beverage

🚗 Transport

Software Engineer developing and maintaining data processing pipelines for ML infrastructure at Serve Robotics. Collaborating with teams to refine data attributes and classifications for a rapidly expanding fleet.

🇺🇸 United States – Remote

💵 $155k - $190k / year

💰 $30M Venture Round on 2023-08

⏰ Full Time

🟡 Mid-level

🟠 Senior

🤖 Machine Learning Engineer

🦅 H1B Visa Sponsor

info

🕒 October 22, 2025

Raspberry AI

11 - 50

💼 Consulting

📣 Marketing

🤖 Artificial Intelligence

Senior Machine Learning Engineer improving generative AI capabilities at Raspberry AI. Conducting research on diffusion models and collaborating with the team on production-ready systems.

🇺🇸 United States – Remote

⏰ Full Time

🟠 Senior

🤖 Machine Learning Engineer

🕒 October 9, 2025

Eight Sleep

11 - 50

🏥 Healthcare

🍽️ Food & Beverage

💼 Consulting

Senior MLOps Engineer responsible for enhancing ML models for sleep technology at Eight Sleep. Joining a mission-driven team that leverages high-tech solutions for better sleep experiences.

🇺🇸 United States – Remote

⏰ Full Time

🟠 Senior

🤖 Machine Learning Engineer

🕒 September 13, 2025

SeatGeek

501 - 1000

🛍️ eCommerce

⚽ Sports

Senior ML Engineer deploying scalable production ML infrastructure at SeatGeek. Focus on pricing, personalization, forecasting, and fraud detection to improve ticketing marketplace.

🇺🇸 United States – Remote

💵 $145k - $200k / year

💰 $238M Series E on 2022-08

⏰ Full Time

🟠 Senior

🤖 Machine Learning Engineer

🦅 H1B Visa Sponsor

info

Airflow

Amazon Redshift

AWS

Cloud

ElasticSearch

Postgres

Python

PyTorch

Scikit-Learn

Tensorflow

Go

.NET

🕒 August 29, 2025

Think Future Technologies

201 - 500

💼 Consulting

🏭 Manufacturing

📦 Logistics

AI/ML Engineer at Think Future Technologies focusing on remote data science. Builds and deploys ML models with Python, PyTorch, and SQL.

🇺🇸 United States – Remote

⏰ Full Time

🟢 Junior

🟡 Mid-level

🤖 Machine Learning Engineer

AWS

Azure

Cloud

Keras

NoSQL

Python

PyTorch

Scikit-Learn

SQL

Tensorflow