Senior Engineer

🔥 14 hours ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of NVIDIA

NVIDIA

10,000+ employees

Founded 1993

🤖 Artificial Intelligence

🎮 Gaming

🚘 Automotive

Artificial Intelligence • Gaming • Automotive

NVIDIA is a leading technology company specializing in accelerated computing and artificial intelligence. NVIDIA pioneers advancements in graphical processing units (GPUs), cloud computing, data centers, and virtual reality, with a focus on gaming, automotive, healthcare, and robotics industries. The company's innovations, such as NVIDIA Omniverse, transform traditional digital processes by enabling high-fidelity simulations and rendering tasks. Their applications span various industries, from autonomous vehicles using NVIDIA DRIVE to healthcare solutions with NVIDIA Clara, and AI-driven analytics and workflows.

📋 Description

• develop innovative solutions that advance AI infrastructure capabilities. • directly influence customer success with breakthrough AI initiatives. • build and deploy custom AI solutions on NCP and Neo Cloud platforms, including distributed training, inference optimization, and MLOps pipelines constructed on NVIDIA reference architectures. • act as the main technical contact for strategic NCPs, offer remote and on-site support, troubleshoot complex production problems, and guide partner engineering teams on NVIDIA platform guidelines. • deploy and manage AI workloads across DGX Cloud, NCP data centers, and major CSP environments using Kubernetes, containers, and GPU scheduling systems aligned to NCP builds. • profile and tune large-scale training and inference workloads on NCP platforms. • implement observability and SLO/SLA monitoring. • lead detailed efforts to reduce latency, cost, and operational risk. • implement and expand NVIDIA reference architectures on partner platforms, develop integrations with partner control planes and customer environments, and ensure smooth API, data pipeline, and enterprise software connectivity. • build detailed implementation guides, runbooks, and post‑mortem documentation that codify standard methodologies for running NVIDIA AI workloads at scale on NCP platforms.

🎯 Requirements

• BS, MS, or Ph.D. in Computer Science, Computer/Electrical Engineering, or a related technical field, or equivalent experience. • 8+ years of experience in customer facing technical roles such as Solutions Engineering, DevOps, Site Reliability, or ML Infrastructure Engineering, ideally supporting large‑scale cloud or service provider environments. • Strong expertise in Linux systems, distributed computing, Kubernetes, containers, and GPU scheduling on multi-tenant or service-provider platforms. • Demonstrated AI/ML experience supporting large‑scale training and inference workloads (e.g., LLMs, generative models, recommendation systems) in production or critically important environments. • Solid programming skills in Python/Go, with hands‑on experience using frameworks such as PyTorch or TensorFlow for training and serving. • Demonstrated capability to collaborate with customer and partner engineering teams in fast-paced environments, guide intricate technical investigations, and bring issues to root cause and resolution. • Excellent communication and technical presentation skills, with the ability to clearly articulate architectures, trade‑offs, and recommendations to both engineering and leadership audiences.

🏖️ Benefits

• equity • benefits

Apply Now

Similar Jobs

🔥 14 hours ago

Helm.ai

51 - 200

🤖 Artificial Intelligence

🚗 Transport

🤝 B2B

Software Engineer designing core components for autonomous vehicles at Helm.ai. Collaborating on deep learning methodologies for advanced motion planning and decision frameworks.

Python

🔥 15 hours ago

General Dynamics Information Technology

10,000+ employees

🎖️ Defense

🔒 Cybersecurity

🤖 Artificial Intelligence

Software Developer transforming technology into opportunity at GDIT while supporting mission-critical government projects. Design, implement, and analyze cloud solutions for Navy operations.

AWS

Azure

Cloud

Google Cloud Platform

Oracle

🔥 15 hours ago

Twilio

5001 - 10000

🔌 API

🤝 B2B

Software Engineer developing innovative communication solutions remotely at Twilio. Contributing to global impact with a vibrant, diverse team.

🔥 15 hours ago

C.H. Robinson

10,000+ employees

🚗 Transport

Senior Software Engineer designing and developing software solutions for logistics challenges at C.H. Robinson. Collaborating with teams and mentoring junior engineers in a remote setting.

AWS

Azure

Cloud

Google Cloud Platform

Java

JavaScript

MongoDB

Oracle

SQL

TFS

🔥 15 hours ago

Clarium

11 - 50

🏥 Healthcare

🤖 Artificial Intelligence

☁️ SaaS

Senior Fullstack Engineer developing end-user features for Clarium's AI-powered healthcare platform. Collaborating with cross-functional teams in a remote setting to enhance hospital supply chains.

Python

React

SQL

TypeScript