
1 - 10 employees
Founded 2024
🤖 Artificial Intelligence
☁️ SaaS
Artificial Intelligence • Cloud Computing • SaaS
Yotta Labs is building the DeOS for AI optimization and orchestration at planet scale. The company provides a high-performance framework to aggregate geo-distributed GPUs, delivering high throughput for heterogeneous compute resources. Yotta Labs focuses on affordability and accessibility for AI training and inference across a spectrum of GPUs, including commodity to high-end options. Their platform supports large language models (LLMs) and enables users to fine-tune AI applications seamlessly, deploying secure AI agents in the cloud with optimized API endpoints.
🕒 June 29
Improve your chances of getting an interview by checking your resume score before you apply.

1 - 10 employees
Founded 2024
🤖 Artificial Intelligence
☁️ SaaS
Artificial Intelligence • Cloud Computing • SaaS
Yotta Labs is building the DeOS for AI optimization and orchestration at planet scale. The company provides a high-performance framework to aggregate geo-distributed GPUs, delivering high throughput for heterogeneous compute resources. Yotta Labs focuses on affordability and accessibility for AI training and inference across a spectrum of GPUs, including commodity to high-end options. Their platform supports large language models (LLMs) and enables users to fine-tune AI applications seamlessly, deploying secure AI agents in the cloud with optimized API endpoints.
• Design and implement high-performance kernels for Attention, MoE, GEMM, collective communication, and quantization. • Optimize kernels for NVIDIA, AMD, and AWS Trainium. • Develop custom operators and graph optimizations using Neuron SDK, PyTorch/XLA, Torch Dynamo, and Neuron Compiler. • Improve performance of vLLM, SGLang, TensorRT-LLM, and custom inference runtimes. • Design scalable distributed training and inference solutions across thousands of accelerators. • Contribute to open-source projects, publish technical findings and engage with the developer community.
• Proficiency in AI programming languages such as Python and C++ • Deep understanding of GPU architecture and performance optimization • Experience with CUDA, Triton, ROCm/HIP, or AWS Neuron • Strong understanding of AI frameworks (e.g., PyTorch, Dynamo, LMCache), model architectures and profiling tools (e.g. Nsight, ROCm Profiler, or Neuron Profiler) • Strong problem-solving skills and the ability to work in a collaborative, remote environment • A background in computer science, engineering, or a related field is preferred
• Competitive compensation with equity • Flexible, remote work environment that values innovation and autonomy
Apply Now🕒 June 26
Research Engineer pushing AI model quality and efficiency at EnCharge AI. Building fine-tuning pipelines and benchmarking frameworks while collaborating closely with hardware teams.
🇺🇸 United States – Remote
💰 $100M Series B - EnCharge AI on 2025-02
⏰ Full Time
🟡 Mid-level
🟠 Senior
🧠 AI Research Scientist
🦅 H1B Visa Sponsor
🕒 June 25
Senior AI Scientist at Atria Health developing clinical AI models leveraging advanced data for precision medicine and preventive care.
🕒 June 24
Senior AI Researcher at NVIDIA working on world foundation models for video generation. Conducting applied research and translating results into robust implementations.
🇺🇸 United States – Remote
💵 $184k - $356.5k / year
⏰ Full Time
🟠 Senior
🧠 AI Research Scientist
🦅 H1B Visa Sponsor
🕒 June 24
AI Research & Development Specialist researching and developing AI tools for creative production in TV advertising. Collaborating with editors to improve workflows and efficiency using AI technologies.
🇺🇸 United States – Remote
💵 $120k - $150k / year
⏰ Full Time
🟡 Mid-level
🟠 Senior
🧠 AI Research Scientist
🕒 June 23
Senior GenAI Scientist II delivering AI solutions to reduce healthcare costs and improve outcomes. Focused on model risk management, validation, metrics, and AI ethics.
🇺🇸 United States – Remote
💵 $145k - $170k / year
⏰ Full Time
🟠 Senior
🧠 AI Research Scientist
🦅 H1B Visa Sponsor