
1 - 10 employees
🤖 Artificial Intelligence
📱 Media
💰 $19M Series A on 2023-06
Artificial Intelligence • Media
ElevenLabs is a research lab dedicated to exploring new frontiers in voice generation. Their mission focuses on making content universally accessible in any language and voice. Through innovative research and deployment of novel methods in voice AI, ElevenLabs aims to enhance the enjoyment of content for diverse audiences and viewers.
🔥 0 minutes ago
Improve your chances of getting an interview by checking your resume score before you apply.

1 - 10 employees
🤖 Artificial Intelligence
📱 Media
💰 $19M Series A on 2023-06
Artificial Intelligence • Media
ElevenLabs is a research lab dedicated to exploring new frontiers in voice generation. Their mission focuses on making content universally accessible in any language and voice. Through innovative research and deployment of novel methods in voice AI, ElevenLabs aims to enhance the enjoyment of content for diverse audiences and viewers.
• Deploy state-of-the-art AI models to production • Own the path from research checkpoints to serving infrastructure • Optimize inference performance for latency, throughput, and cost • Apply quantization, distillation, KV-cache optimization, batching strategies, and custom kernels • Build and tune high-performance serving systems for real-time, streaming workloads • Create tooling and infrastructure enabling researchers to ship models to production quickly and safely • Measure and validate model performance characteristics
• No formal certifications or degrees required • Enthusiasm for solving impressively hard engineering problems • Ability to demonstrate work through past projects, designs, or GitHub contributions • Experience deploying and serving ML models in production, ideally for latency-sensitive or real-time applications • Strong engineering skills in GPU programming and inference optimization, including CUDA, Triton, TensorRT, vLLM, or SGLang • Ability to autonomously profile, diagnose, and eliminate bottlenecks across the serving stack • Ability to build tooling to measure serving-stack performance
• Annual discretionary professional development stipend • Annual discretionary stipend for social travel to meet colleagues • Annual company offsite • Monthly co-working stipend for employees not near a main hub • Option to work from company offices in London, New York, San Francisco, and Warsaw
Apply Now🕒 June 23
Full Stack Engineer developing privacy-first analytics capabilities for Matomo. Involves innovation-focused product development, requiring proficiency in PHP, Python, and JavaScript.