Product Manager – AI Inference, Model Serving

🕒 Maio 28

🤠 Texas – Remoto

infoinfo

⏰ Tempo Integral

🟠 Sênior

🔴 Especialista

✅ Gerente de Produto

🦅 Patrocina Visto H1B

infoinfo

👻 Score fantasma 31%

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Mirantis

Mirantis

501 - 1000 funcionários

💼 Consultoria

🏥 Saúde

📦 Logística

Consulting • Healthcare • Logistics

A Mirantis é uma empresa especializada em soluções de gerenciamento de contêineres e infraestrutura em nuvem. Ela oferece uma variedade de produtos, incluindo o Mirantis Kubernetes Engine (MKE), o Mirantis OpenStack para Kubernetes (MOSK) e o Mirantis Container Cloud (MCC), que fornecem plataformas de gerenciamento de Kubernetes e contêineres em nível empresarial. A Mirantis também desenvolve ferramentas para cadeias seguras de fornecimento de software, como o Mirantis Container Runtime (MCR) e o Mirantis Secure Registry (MSR). Como defensora das tecnologias de código aberto, a Mirantis apoia vários projetos e fornece recursos como o Lens Desktop, um popular IDE para Kubernetes, além de suporte técnico para empresas que adotam tecnologias nativas em nuvem. Suas soluções atendem a setores como serviços públicos, serviços financeiros e indústrias mais amplas de SaaS e serviços de tecnologia.

Descrição

• Own product strategy, roadmap, and lifecycle for inference and model serving, including serverless inference, dedicated endpoints, autoscaling, routing, KV cache management, and the related observability • Lead deep technical discovery with NeoClouds, sovereign clouds, and enterprise platform teams, and translate findings into prioritized requirements and architecture direction • Partner with engineering on system design trade-offs across runtime integration, GPU scheduling, network, storage, and serving topology, including disaggregated serving and multi-model serving • Define positioning grounded in measurable outcomes: latency distributions, throughput per GPU, utilization, tail reliability, and cost per tokens • Drive go-to-market execution: pricing and packaging, reference architectures, sizing guides, PoC playbooks, and direct engagement with customers, analysts, and ecosystem partners

🎯 Requisitos

• 7+ years in product management, technical product management, or a senior technical role owning AI/ML and inference product(s) • Strong understanding of production AI inference, including model serving, serverless execution, dedicated endpoints, autoscaling, routing, workload placement, observability, and reliability • Proven capability to reason about performance trade-offs across GPU, network, storage, orchestration, and runtime layers, and to translate low-level technical capability into business value such as TTFT, throughput per GPU, and TCO • Working knowledge of modern inference runtimes (vLLM, SGLang, TensorRT-LLM, Dynamo, Triton) and the optimization patterns that matter in production: continuous batching, KV cache management, cold starts, prefill versus decode, disaggregated serving, and multi-model serving • Credibility with engineering leaders and infrastructure operators, including comfort in production architecture reviews and technical commercial conversations with platform engineering buyers.

🏖️ Benefícios

• Work with an established Silicon Valley leader in the cloud infrastructure industry. • Work with exceptionally passionate, talented and engaging colleagues, helping Fortune 500 and Global 2000 customers implement next-generation cloud technologies. • Be a part of cutting-edge, open-source innovation. • Thrive in the high-energy environment of a young company where openness, collaboration, risk-taking, and continuous growth are valued. • Professional development and training. • Attend conferences and working groups. • Customized workstation (macOS, Windows). • A competitive compensation package with strong benefits plan and stock options.

Candidatar-se

Vagas Similares

🕒 Maio 27

GitLab

1001 - 5000

💼 Consultoria

📣 Marketing

🤖 Inteligência Artificial

Principal Product Manager shaping GitLab’s self-hosted AI and custom-model capabilities. Defining strategy, roadmaps, and agent orchestration experiences for secure, flexible DevSecOps deployments.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $203.200 - $345.600 / ano

💰 Secondary Market em 2020-11

⏰ Tempo Integral

🔴 Especialista

✅ Gerente de Produto

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Maio 27

Tiger Analytics

1001 - 5000

🏥 Saúde

📦 Logística

📣 Marketing

Senior Manager leading commercial data product development for pharma launch at Tiger Analytics. Collaborating with stakeholders in Life Sciences to deliver analytical solutions and data products.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Maio 27

hims & hers

201 - 500

🏥 Saúde

💼 Consultoria

📣 Marketing

Lead Product Manager for Hims & Hers redefining healthcare via innovative consumer experiences. Owning consumer journey from discovery through engagement with AI integration.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $160.000 - $210.000 / ano

⏰ Tempo Integral

🟠 Sênior

✅ Gerente de Produto

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Maio 27

Rightway

201 - 500

🏥 Saúde

⚕️ Seguro de Saúde

☁️ SaaS

Senior Product Manager shaping the digital pharmacy experience at Rightway. Focusing on member engagement, product decision-making, and collaboration with design, engineering, and clinical teams.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $146.000 - $200.000 / ano

💰 $100.000.000 Series C em 2021-03

⏰ Tempo Integral

🟠 Sênior

✅ Gerente de Produto

🦅 Patrocina Visto H1B

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Maio 26

NinjaHoldings

51 - 200

💸 Finanças

💳 Fintech

👥 B2C

Senior Product Manager driving key business goals in fintech startup. Collaborating with various departments to ensure alignment and support technical projects.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $120.000 - $180.000 / ano

⏰ Tempo Integral

🟠 Sênior

✅ Gerente de Produto

🗣️🇺🇸🇬🇧 Inglês obrigatório