Senior Deep Learning Algorithm Engineer

🔥 0 minutes ago

🏄 California – Remote

info

💵 $224k - $356.5k / year

⏰ Full Time

🟠 Senior

👷🏻‍♀️ Engineer

🦅 H1B Visa Sponsor

info
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of NVIDIA

NVIDIA

10,000+ employees

Founded 1993

🏥 Healthcare

🏭 Manufacturing

🤖 Artificial Intelligence

Healthcare • Manufacturing • Artificial Intelligence

NVIDIA is a leading technology company specializing in accelerated computing and artificial intelligence. NVIDIA pioneers advancements in graphical processing units (GPUs), cloud computing, data centers, and virtual reality, with a focus on gaming, automotive, healthcare, and robotics industries. The company's innovations, such as NVIDIA Omniverse, transform traditional digital processes by enabling high-fidelity simulations and rendering tasks. Their applications span various industries, from autonomous vehicles using NVIDIA DRIVE to healthcare solutions with NVIDIA Clara, and AI-driven analytics and workflows.

📋 Description

• Design, develop, and optimize workloads for NVIDIA’s Megatron Core and NeMo Framework teams. • Expand Megatron Core and NeMo Framework capabilities for developing, training, and optimizing LLM and multimodal foundation models. • Design and implement distributed training algorithms, model parallel paradigms, and model optimizations. • Define robust APIs and analyze and tune performance. • Expand toolkits and libraries to be more comprehensive and coherent. • Collaborate with internal partners, users, and the open-source community to analyze, design, and implement optimized solutions. • Develop algorithms for AI/deep learning, data analytics, machine learning, or scientific computing. • Contribute to and advance open-source NeMo-RL, Megatron Core, and NeMo Framework. • Solve large-scale, end-to-end AI training and inference challenges across orchestration, data preprocessing, training, tuning, and deployment. • Improve model architectures, distributed training algorithms, and model parallel paradigms. • Tune performance and optimize model training and fine-tuning with mixed-precision recipes on next-generation NVIDIA GPU architectures. • Research, prototype, and develop robust, scalable AI tools and pipelines.

🎯 Requirements

• MS, PhD, or equivalent experience in Computer Science, AI, Applied Math, or related fields. • 5+ years of industry experience. • Experience with AI frameworks such as PyTorch, JAX, or Ray, and/or inference and deployment environments such as TRTLLM, vLLM, or SGLang. • Proficiency in Python programming, software design, debugging, performance analysis, test design, and documentation. • Consistent record of working effectively across multiple engineering initiatives and improving AI libraries with new innovations. • Strong understanding of AI/deep-learning fundamentals and their practical applications. • Hands-on experience in large-scale AI training and understanding of compute system concepts including latency/throughput bottlenecks, pipelining, and multiprocessing. • Experience with reinforcement learning algorithms and compute patterns. • Expertise in distributed computing, model parallelism, and mixed-precision training. • Experience with generative AI techniques applied to LLM and multimodal learning involving text, image, and video. • Knowledge of GPU/CPU architecture and related numerical software.

🏖️ Benefits

• Equity • Benefits

Apply Now

Similar Jobs

🔥 1 hour ago

Knauf Insulation North America

-

🏗️ Construction

🚀 Aerospace

⚡ Energy

Process Development Engineer optimizing white wool insulation manufacturing processes across Knauf Insulation’s North American plants. Driving Six Sigma improvements, technical support, standardization, and measurable gains in quality, output, and cost.

🔥 2 hours ago

Falconwood, Incorporated

201 - 500

💼 Consulting

🏭 Manufacturing

🎖️ Defense

Endpoint Engineer supporting Falconwood’s DoD IT consulting and programmatic services for Navy enterprise workstations. Managing Azure, MECM, M365, hardware, software, and cybersecurity requirements.

🔥 2 hours ago

Pragmatike

11 - 50

💼 Consulting

📣 Marketing

🎯 Recruiter

Forward Deployed Engineer building AI infrastructure and production integrations for enterprise customers. Translating customer requirements into scalable systems while partnering with Product and Engineering.

🔥 2 hours ago

Pragmatike

11 - 50

💼 Consulting

📣 Marketing

🎯 Recruiter

Senior Forward Deployed Engineer building and deploying AI infrastructure solutions for enterprise customers. Translating customer requirements into scalable integrations, automation, and production systems.

🔥 2 hours ago

Pragmatike

11 - 50

💼 Consulting

📣 Marketing

🎯 Recruiter

Lead Forward Deployed Engineer building AI-powered infrastructure solutions for enterprise customers. Deploying integrations, automation, APIs, and scalable production systems across the United States.