Senior MLOps Engineer – DSX Enablement

🔥 0 minutes ago

🌐 Poland, Germany, +1 more countries – Remote

infoinfo

💵 zł292.5k - zł650k / year

⏰ Full Time

🟠 Senior

🤖 Machine Learning Engineer

👻 Ghost score 1%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of NVIDIA

NVIDIA

10,000+ employees

Founded 1993

🏥 Healthcare

🏭 Manufacturing

🤖 Artificial Intelligence

Healthcare • Manufacturing • Artificial Intelligence

NVIDIA is a leading technology company specializing in accelerated computing and artificial intelligence. NVIDIA pioneers advancements in graphical processing units (GPUs), cloud computing, data centers, and virtual reality, with a focus on gaming, automotive, healthcare, and robotics industries. The company's innovations, such as NVIDIA Omniverse, transform traditional digital processes by enabling high-fidelity simulations and rendering tasks. Their applications span various industries, from autonomous vehicles using NVIDIA DRIVE to healthcare solutions with NVIDIA Clara, and AI-driven analytics and workflows.

📋 Description

• Develop innovative solutions advancing AI infrastructure capabilities • Advise infrastructure experts on demands of ML workloads • Help practitioners diagnose and solve full-stack AI and ML system problems • Support internal and external customers’ AI and ML initiatives, including LLM performance evaluation and new hardware in open-source frameworks • Build and deploy custom AI solutions on NeoCloud platforms and NVIDIA Cloud Partners, including distributed training, inference optimization, and MLOps pipelines • Act as primary technical contact for customers and partners, guide joint engagements, ensure initiative success on DGX Cloud, and solve complex production problems • Collaborate with infrastructure software and accelerated-framework teams • Profile and tune large-scale training and inference workloads to reduce latency, cost, and operational risk • Develop open-source tools and reference architectures for machine learning and AI workloads, pipelines, and systems at scale

🎯 Requirements

• BS, MS, or Ph.D. in Computer Science, Computer/Electrical Engineering, or a related technical field, or equivalent experience • 8+ years of experience in technical roles such as data science, data engineering, or ML engineering, ideally targeting large-scale production systems • Demonstrated AI/ML experience across multiple phases of the machine learning lifecycle, from exploratory analysis to production systems • Facility with Linux, batch schedulers, Kubernetes, distributed filesystems, and advanced networking at datacenter scale • Solid scripting and programming skills in bash and Python • Solid systems programming skills in C++, Go, or Rust • Experience using machine learning or deep learning frameworks for training and inference • Excellent communication and technical presentation skills, with ability to articulate architectures, trade-offs, and recommendations to engineering and leadership audiences • Clear record of engineering discipline and execution on projects • Experience contributing to and working in open-source communities • Experience with NVIDIA ecosystem, including DGX systems, CUDA, NeMo, RAPIDS, Triton, NIM, InfiniBand, NVLink, and RoCE • Experience building machine learning systems in a security-critical environment and with distributed training and inference frameworks • Familiarity with cloud-native MLOps practices, including containerization, CI/CD, workflow automation, observability stacks, and GitOps workflows • Direct experience diagnosing and fixing cross-layer performance or correctness problems spanning hardware, networking, accelerators, hypervisors or OS, compilers or runtimes, application code, and libraries

🏖️ Benefits

• Competitive salaries • Generous benefits package

Apply Now

Similar Jobs

🕒 August 24

Spyrosoft

1001 - 5000

🚘 Automotive

🏥 Healthcare

💼 Consulting

Senior MLOps Engineer building scalable Google Cloud ML platforms at Spyrosoft, a software engineering company. Automating Vertex AI model lifecycles from experimentation through production.

BigQuery

Cloud

Google Cloud Platform

Python

🕒 August 21

Aiphoria

51 - 200

💼 Consulting

📦 Logistics

📣 Marketing

Machine Learning Engineer optimizing automatic speech recognition models for voice-assisted banking. Building audio pipelines and deploying accurate, low-latency speech systems.

🇵🇱 Poland – Remote

💰 $34M Series A - Aiphoria on 2025-07

⏰ Full Time

🟡 Mid-level

🟠 Senior

🤖 Machine Learning Engineer

Python

PyTorch

🕒 August 21

Fetcherr

51 - 200

📦 Logistics

✈️ Travel

🤖 Artificial Intelligence

MLOps Engineer building scalable AI models and data pipelines for Fetcherr’s market pricing platform. Supporting demand forecasting and real-time pricing intelligence for global aviation and other volatile markets.

Airflow

Docker

Kubernetes

Python

🕒 August 10

EcoVadis

1001 - 5000

💼 Consulting

🏥 Healthcare

📦 Logistics

Senior AI/ML Engineer building production-grade AI systems for EcoVadis, a business sustainability ratings provider. Designing MLOps pipelines and cloud solutions with Python, Azure, MLflow, and Databricks.

Azure

Cloud

Docker

Python

🕒 August 4

RedSky

11 - 50

💼 Consulting

🎖️ Defense

🏥 Healthcare

Lead ML Engineer owning deep learning, computer vision, and edge inference for a robotics automation startup. Building the AI/ML product core from scratch with the founding technical team.

🗣️🇵🇱 Polish Required

Cloud

PyTorch