
51 - 200 funcionários
Fundada em 2015
💼 Consultoria
🏥 Saúde
📦 Logística
💰 $47.000.000 Series B em 2022-11
Consulting • Healthcare • Logistics
A Deepgram é uma empresa líder em IA de voz que fornece APIs poderosas para aplicações de reconhecimento de fala, síntese de texto para fala e entendimento de linguagem. Sua plataforma permite que os desenvolvedores criem soluções avançadas de IA de voz para casos de uso como centrais de atendimento, transcrição médica, IA conversacional, entre outros. Conhecida por sua precisão inigualável, velocidade e custo-benefício, a tecnologia da Deepgram é confiada por grandes empresas e startups em todo o mundo. Oferecendo capacidades de transcrição em tempo real e altamente precisas, a Deepgram ajuda as empresas a obter insights a partir de dados de voz, tornando-se uma ferramenta essencial para transformar interações de voz.
🕒 Julho 7
🇺🇸 Estados Unidos – Remoto (EUA)
💵 $219.300 - $274.100 / ano
⏰ Tempo Integral
🟡 Pleno
🟠 Sênior
🤖 Engenheiro de IA
🦅 Patrocina Visto H1B
👻 Score fantasma 18%
🗣️🇺🇸🇬🇧 Inglês obrigatório
Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

51 - 200 funcionários
Fundada em 2015
💼 Consultoria
🏥 Saúde
📦 Logística
💰 $47.000.000 Series B em 2022-11
Consulting • Healthcare • Logistics
A Deepgram é uma empresa líder em IA de voz que fornece APIs poderosas para aplicações de reconhecimento de fala, síntese de texto para fala e entendimento de linguagem. Sua plataforma permite que os desenvolvedores criem soluções avançadas de IA de voz para casos de uso como centrais de atendimento, transcrição médica, IA conversacional, entre outros. Conhecida por sua precisão inigualável, velocidade e custo-benefício, a tecnologia da Deepgram é confiada por grandes empresas e startups em todo o mundo. Oferecendo capacidades de transcrição em tempo real e altamente precisas, a Deepgram ajuda as empresas a obter insights a partir de dados de voz, tornando-se uma ferramenta essencial para transformar interações de voz.
• Take Deepgram's Speech and Conversational models and get them running on embedded and low-power consumer hardware — defining the architecture for on-device, real-time inference across a diverse range of processors and accelerators. • Optimize models for constrained targets through quantization, pruning, distillation, operator fusion, and architecture-specific compilation to meet strict latency, memory, power, and thermal budgets. • Write and optimize performance-critical runtime code (C, C++, and/or Rust) for embedded environments, including bare-metal and real-time operating systems such as FreeRTOS and Zephyr. • Integrate with industry-standard edge inference runtimes and vendor NPU/DSP toolchains, mapping model graphs efficiently onto on-device accelerators and CPU/GPU/NPU heterogeneity. • Build the on-device runtime plumbing: model packaging, deployment pipelines, over-the-air update mechanisms, and lightweight telemetry for devices operating with limited or intermittent connectivity. • Establish repeatable benchmarking and validation across target hardware — measuring latency, accuracy, power consumption, memory footprint, and resource utilization — and catch regressions before they ship. • Partner with silicon and device vendors on SDK integration and performance tuning, getting our models to run efficiently on new chipsets and reference platforms. • Collaborate with Research and Engine teams to influence model architectures toward edge-friendly designs from the start, reducing the optimization burden at deployment time.
• Experience delivering production systems on resource-constrained hardware — embedded systems, mobile, edge AI, or small low-power devices. • Strong proficiency in C, C++, and/or Rust, with experience writing performance-critical code for constrained environments. • Hands-on experience with model optimization for on-device deployment, including quantization, pruning, knowledge distillation, or architecture-specific compilation. • Familiarity with edge inference runtimes (e.g., ONNX Runtime, TensorRT, TFLite, ExecuTorch) and/or vendor-specific NPU/DSP toolchains. • A strong understanding of hardware-software interaction — CPU/GPU/NPU/DSP architectures, memory hierarchies, fixed-point/integer arithmetic, and power management — and how they affect inference performance. • Experience working close to the metal: bare-metal or RTOS environments (e.g., FreeRTOS, Zephyr), embedded Linux, or microcontroller and edge SoC development. • Strong communication skills and a builder mindset — you can scope an ambiguous optimization problem, drive it to a measurable result, and explain the tradeoffs clearly.
• Offers Equity • Offers Bonus • 10% Annual Bonus
Candidatar-se🕒 Julho 7
Senior AI Engineer responsible for developing AI-powered solutions for legal workflows. Collaborating with legal stakeholders and employing advanced AI techniques for contract and compliance tasks.
🇺🇸 Estados Unidos – Remoto (EUA)
💵 $185.000 - $200.000 / ano
⏰ Tempo Integral
🟠 Sênior
🤖 Engenheiro de IA
🦅 Patrocina Visto H1B
🗣️🇺🇸🇬🇧 Inglês obrigatório
🕒 Julho 3
Applied AI Engineer focusing on building AI-driven features in insurance processes with global teams. Responsibilities include model evaluation, automation features, and cross-team collaboration.
🇺🇸 Estados Unidos – Remoto (EUA)
💵 $140.000 - $200.000 / ano
⏰ Tempo Integral
🟡 Pleno
🟠 Sênior
🤖 Engenheiro de IA
🗣️🇺🇸🇬🇧 Inglês obrigatório
🕒 Julho 3
Senior AI Engineer at Chemours designing and integrating Generative AI solutions in various business functions. Collaborating across teams to deliver impactful projects while mentoring junior developers.
🇺🇸 Estados Unidos – Remoto (EUA)
💵 $126.067 - $196.980 / ano
⏰ Tempo Integral
🟠 Sênior
🤖 Engenheiro de IA
🦅 Patrocina Visto H1B
🗣️🇺🇸🇬🇧 Inglês obrigatório
🕒 Julho 2
Lead AI Engineer shaping AI architecture and strategy for digital insurance platform at Veracity, ensuring integration and reliability of AI systems while collaborating across teams.
🇺🇸 Estados Unidos – Remoto (EUA)
💵 $170.000 - $215.000 / ano
⏰ Tempo Integral
🟠 Sênior
🤖 Engenheiro de IA
🗣️🇺🇸🇬🇧 Inglês obrigatório
🕒 Julho 2
Lead AI Engineer at InsCipher creating AI solutions for digital insurance platform. Shaping enterprise AI strategies and collaborating cross-functionally to enhance customer outcomes.
🇺🇸 Estados Unidos – Remoto (EUA)
💵 $170.000 - $215.000 / ano
⏰ Tempo Integral
🟠 Sênior
🤖 Engenheiro de IA
🗣️🇺🇸🇬🇧 Inglês obrigatório