Audio Engineer, Model Efficiency

Provável vaga fantasma

🕒 Novembro 8, 2025

🇨🇦 Canadá – Remoto

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

👷🏻‍♀️ Engenheiro

👻 Score fantasma 66%

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Cohere

Cohere

11 - 50 funcionários

🤖 Inteligência Artificial

🏢 Corporativo

☁️ SaaS

Artificial Intelligence • Enterprise • SaaS

A Cohere é uma plataforma de IA líder, fornecendo às empresas modelos de linguagem avançada e um espaço de trabalho integrado projetado para eficiência e segurança. Com uma família de modelos generativos e de recuperação de alto desempenho, a Cohere permite que as organizações simplifiquem fluxos de trabalho, melhorem a segurança dos dados e descubram insights em diversas indústrias por meio de capacidades multilingues. Seu foco em soluções de IA personalizadas garante a proteção de dados críticos, facilitando a integração perfeita nos processos organizacionais existentes.

Descrição

• Work on advancing core audio model serving metrics, including latency, throughput, and quality • Dive deep into our systems, identifying bottlenecks, and delivering creative solutions for audio processing and streaming workloads. • Collaborate closely with both the training and serving infrastructure teams to ensure seamless integration between model development and deployment, with a special focus on real-time and streaming audio inference.

🎯 Requisitos

• Significant experience developing high-performance audio or machine learning inference systems. • Proficiency with programming languages such as C++ and Python. • Hands-on experience with deep learning models for audio, speech, or language applications. • A bias for action and a strong results-oriented mindset. • Considerable experience with GPU programming, low-level system optimization, model parallelization techniques over multiple GPUs • Experience with duplex real-time streaming architectures. • Internals of machine learning frameworks for audio (such as PyTorch, TensorFlow, or specialized audio libraries). • Experience with inference framework like vLLM, SGLang, Tensort-LLM, or custom distributed inference systems • Sequence modeling (e.g., transformers for audio/speech) and end-to-end audio pipeline optimization • If some of the above doesn’t line up perfectly with your experience, we still encourage you to apply!

🏖️ Benefícios

• An open and inclusive culture and work environment • Work closely with a team on the cutting edge of AI research • Weekly lunch stipend, in-office lunches & snacks • Full health and dental benefits, including a separate budget to take care of your mental health • 100% Parental Leave top-up for up to 6 months • Personal enrichment benefits towards arts and culture, fitness and well-being, quality time, and workspace improvement • Remote-flexible, offices in Toronto, New York, San Francisco, London and Paris, as well as a co-working stipend • 6 weeks of vacation (30 working days!)

Candidatar-se