
1001 - 5000 funcionários
Fundada em 2019
💼 Consultoria
📦 Logística
🏭 Manufatura
💰 Grant em 2020-12
Consulting • Logistics • Manufacturing
A Cerence Inc. é uma empresa global focada em fornecer soluções impulsionadas por IA, especialmente na indústria automotiva. Eles se especializam em tecnologias de IA conversacional e generativa que criam interações inteligentes, naturais e personalizadas entre humanos e veículos. Com inovações como seus modelos de linguagem de grande escala automotivos proprietários, a Cerence aprimora as experiências dos usuários em várias formas de transporte, incluindo carros, duas rodas e caminhões. A empresa já possui mais de 500 milhões de veículos entregues com sua tecnologia de IA, atendendo a mais de 80 OEMs e clientes Tier 1 em todo o mundo. A Cerence é dedicada a avanços contínuos em IA, com o objetivo de revolucionar as experiências do usuário no carro por meio da entrega rápida e integração perfeita de suas soluções.
🕒 Junho 29
🍂 Massachusetts – Remoto
💵 $141.400 - $226.300 / ano
⏰ Tempo Integral
🟠 Sênior
🧑💻 Engenheiro Full-stack
🗣️🇺🇸🇬🇧 Inglês obrigatório
C++
Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

1001 - 5000 funcionários
Fundada em 2019
💼 Consultoria
📦 Logística
🏭 Manufatura
💰 Grant em 2020-12
Consulting • Logistics • Manufacturing
A Cerence Inc. é uma empresa global focada em fornecer soluções impulsionadas por IA, especialmente na indústria automotiva. Eles se especializam em tecnologias de IA conversacional e generativa que criam interações inteligentes, naturais e personalizadas entre humanos e veículos. Com inovações como seus modelos de linguagem de grande escala automotivos proprietários, a Cerence aprimora as experiências dos usuários em várias formas de transporte, incluindo carros, duas rodas e caminhões. A empresa já possui mais de 500 milhões de veículos entregues com sua tecnologia de IA, atendendo a mais de 80 OEMs e clientes Tier 1 em todo o mundo. A Cerence é dedicada a avanços contínuos em IA, com o objetivo de revolucionar as experiências do usuário no carro por meio da entrega rápida e integração perfeita de suas soluções.
• Optimize and deploy high ‑ performance LLM inference pipelines • Own inference runtimes across data center, edge, and embedded platforms • Push model performance through quantization, kernel fusion, and cache optimization • Drive latency and throughput improvements that directly impact production products • Enable efficient, reliable deployment without external vendor dependency • Build deep expertise and ownership of: vLLM TensorRT‑LLM llama.cpp QAIRT • Extend and tune inference engines using custom CUDA kernels • Adapt runtimes for constrained and embedded deployment environments • Implement and evaluate quantization strategies: INT8, INT4, FP4, FP8, mixed precision AWQ GPTQ • Balance accuracy, latency, memory footprint, and throughput • Optimize key–value cache performance through: Paging Prefix caching Cache ‑ aware memory layout design • Design and tune: Batching strategies Continuous batching Speculative decoding
• Proven experience optimizing ML inference performance in production • Deep understanding of GPU architecture and memory hierarchies • Hands ‑ on experience with CUDA and low ‑ level performance tuning • Experience deploying models beyond research environments • Critical Technical Skills • Inference engines: vLLM, TensorRT ‑ LLM, llama.cpp, QAIRT • CUDA kernel development and profiling • Quantization techniques: INT8/INT4/FP4/FP8, AWQ, GPTQ • KV cache optimisation and memory layout design • Latency optimisation: batching, speculative decoding, continuous batching
• Annual bonus opportunity • Insurance coverage (medical, dental, vision, life, and disability) • Paid time off • Paid holidays • Company contribution to the RRSP (Registered Retirement Savings Plan) • Equity awards for certain positions and levels • Remote and/or hybrid work available depending on the position
Candidatar-se🕒 Junho 29
Senior Full Stack Developer integrating new features and working with clients in retail electronics. Requires 5+ years experience and knowledge of Microsoft tech stack.
🗣️🇮🇹 Italiano obrigatório
🕒 Junho 29
Senior Software Engineer developing scalable platform components and supporting cloud infrastructure at Robert Half. Leading design and implementation with a focus on CI/CD and platform reliability.
🇺🇸 Estados Unidos – Remoto (EUA)
💵 $104.000 - $153.000 / ano
⏰ Tempo Integral
🟡 Pleno
🟠 Sênior
🧑💻 Engenheiro Full-stack
🦅 Patrocina Visto H1B
🗣️🇺🇸🇬🇧 Inglês obrigatório
🕒 Junho 29
Tech Lead for Consumer Team to drive technical direction and execution for consumer web experience at Koalafi. Leading a team of engineers in modernizing systems and delivering tools for financial needs.
🇺🇸 Estados Unidos – Remoto (EUA)
💵 $167.723 - $217.053 / ano
💰 Debt Financing em 2022-08
⏰ Tempo Integral
🟠 Sênior
🧑💻 Engenheiro Full-stack
🦅 Patrocina Visto H1B
🗣️🇺🇸🇬🇧 Inglês obrigatório
🕒 Junho 29
Lead Engineer on Product Catalog Team for Stitch Fix redefining retail with technology and data. Responsible for evolving catalog systems and improving product data quality.
🇺🇸 Estados Unidos – Remoto (EUA)
💵 $111.800 - $186.000 / ano
💰 $36.900.000 Venture Round em 2017-11
⏰ Tempo Integral
🟠 Sênior
🧑💻 Engenheiro Full-stack
🦅 Patrocina Visto H1B
🗣️🇺🇸🇬🇧 Inglês obrigatório
🕒 Junho 29
Senior Software Engineer applying software engineering and machine learning for mineral exploration at KoBold Metals. Collaborating with data scientists and geologists to shape the future of energy transition metal discovery.
🇺🇸 Estados Unidos – Remoto (EUA)
💵 $170.000 - $215.000 / ano
⏰ Tempo Integral
🟠 Sênior
🧑💻 Engenheiro Full-stack
🗣️🇺🇸🇬🇧 Inglês obrigatório
Numpy
Python