
10,000+ employees
Founded 1993
š„ Healthcare
š Manufacturing
š¤ Artificial Intelligence
Healthcare ⢠Manufacturing ⢠Artificial Intelligence
NVIDIA is a leading technology company specializing in accelerated computing and artificial intelligence. NVIDIA pioneers advancements in graphical processing units (GPUs), cloud computing, data centers, and virtual reality, with a focus on gaming, automotive, healthcare, and robotics industries. The company's innovations, such as NVIDIA Omniverse, transform traditional digital processes by enabling high-fidelity simulations and rendering tasks. Their applications span various industries, from autonomous vehicles using NVIDIA DRIVE to healthcare solutions with NVIDIA Clara, and AI-driven analytics and workflows.
š„ 0 minutes ago
š California ā Remote
šµ $224k - $356.5k / year
ā° Full Time
š Senior
š» Solutions Engineer
š¦ H1B Visa Sponsor
š» Ghost score 1%
Improve your chances of getting an interview by checking your resume score before you apply.

10,000+ employees
Founded 1993
š„ Healthcare
š Manufacturing
š¤ Artificial Intelligence
Healthcare ⢠Manufacturing ⢠Artificial Intelligence
NVIDIA is a leading technology company specializing in accelerated computing and artificial intelligence. NVIDIA pioneers advancements in graphical processing units (GPUs), cloud computing, data centers, and virtual reality, with a focus on gaming, automotive, healthcare, and robotics industries. The company's innovations, such as NVIDIA Omniverse, transform traditional digital processes by enabling high-fidelity simulations and rendering tasks. Their applications span various industries, from autonomous vehicles using NVIDIA DRIVE to healthcare solutions with NVIDIA Clara, and AI-driven analytics and workflows.
⢠Solve hard Day 2 operations problems at scale alongside partner engineers ⢠Find causes, prototype approaches, validate solutions under representative load, and leave operational practices partners can run ⢠Help partners prepare operating models for new NVIDIA platforms, capacity, services, and use cases ⢠Drive adoption in live environments without degrading service ⢠Improve reliability, performance, and economics using incident frequency, recovery time, utilization, and cost-per-token measures ⢠Identify and help close Day 2 maturity gaps across people, process, tooling, telemetry, security, and incident response ⢠Convert validated work into operating procedures, reference architectures, assessments, automation, and agentic workflows ⢠Spot cross-partner patterns and provide field evidence to account teams, support, product, and engineering ⢠Improve NVIDIA's factory planning function
⢠BS, MS, or PhD in Computer Science, Electrical or Computer Engineering, Physics, Mathematics, or a related field - or equivalent experience ⢠12+ years in production infrastructure, cloud engineering, solutions architecture, site reliability engineering, HPC, or a similar technical role; alternatively, 5+ years of exceptional specialist-level work in large-scale GPU or AI infrastructure ⢠Experience building, operating, or improving distributed infrastructure under real production load ⢠Deep expertise in at least one part of the Day 2 stack, backed by hands-on work with large-scale GPU, HPC, or cloud infrastructure ⢠Working experience with Kubernetes or Slurm, GPU scheduling and multi-tenancy, Prometheus, Grafana or OpenTelemetry, and automation with Terraform, Ansible, Argo CD, or similar tooling ⢠Strong Linux knowledge and enough Python, Bash, or similar experience to automate measurement, diagnosis, validation, or remediation ⢠Detailed evidence-led troubleshooting across system boundaries ⢠Ability to lead sophisticated work with partner engineers and cross-functional teams without direct authority ⢠Strong communication, prioritization, and time-management skills across multiple partner engagements ⢠Real world experience operating a GPU cloud, HPC environment, or large-scale AI platform under customer load ⢠Experience building or maturing a 24/7 operations function, including observability, incident and problem management, coverage, and on-call design ⢠Hands-on experience with NVIDIA rack-scale platforms such as GB200 or GB300 NVL72, or NVIDIA operations technologies such as Spectrum-X, UFM, Base Command Manager, Mission Control, and GPU or Network Operators ⢠Experience improving fleet health or unit economics through benchmarking, infrastructure as code, GitOps, automated diagnosis, or agent-based remediation
⢠Competitive salaries ⢠Generous benefits package ⢠Equity ⢠Benefits
Apply Nowš„ 43 minutes ago
Enterprise Solution Architect modernizing healthcare IT systems for Ascension, a nonprofit Catholic health system. Governing architecture, integration, data, cloud, security, and AI initiatives.
šŗšø United States ā Remote
šµ $154.9k - $216k / year
š° $500M Debt on 2019-11
ā° Full Time
š Senior
š» Solutions Engineer
š„ 46 minutes ago
Senior Solutions Architect guiding defense and government customers using Virtualiticsā AI-native readiness applications. Designing architectures, supporting pre-sales, and managing technical risks.
šŗšø United States ā Remote
š„ Funding within the last year
š° $15M Debt Financing - Virtualitics on 2025-09
ā° Full Time
š Senior
š» Solutions Engineer
š„ 49 minutes ago
IAM Solutions Architect designing Okta and Auth0 identity security for Stand Together, a philanthropic community tackling Americaās biggest social problems. Building human and non-human access governance and zero-trust systems.
šŗšø United States ā Remote
šµ $175k - $205k / year
ā° Full Time
š” Mid-level
š Senior
š» Solutions Engineer
š¦ H1B Visa Sponsor
š„ 50 minutes ago
Senior actuarial solutions engineer modernizing Mercury Insuranceās P&C actuarial applications and dashboards. Improving data access, pricing analysis, monitoring, reporting, and decision support.
šŗšø United States ā Remote
šµ $110.5k - $153.4k / year
ā° Full Time
š Senior
š» Solutions Engineer
š¦ H1B Visa Sponsor
š„ 1 hour ago
Solutions Engineer guiding enterprise customers from technical discovery through deployment. Helping Traversal scale sales of its AI-powered SRE infrastructure platform.
šŗšø United States ā Remote
šµ $150k - $300k / year
š° Seed on 2025-07
ā° Full Time
š” Mid-level
š Senior
š» Solutions Engineer