Principal Performance Engineer, Lead

Vaga não está no LinkedIn

🕒 Abril 7

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Akamai Technologies

Akamai Technologies

5001 - 10000 funcionários

🔒 Cibersegurança

🏢 Corporativo

📱 Mídia

Cybersecurity • Enterprise • Media

A Akamai Technologies é uma empresa global de plataforma de borda e serviços em nuvem que oferece soluções de entrega de conteúdo, computação na borda e segurança. A empresa opera uma das maiores redes distribuídas do mundo para acelerar e proteger o tráfego da web, mídia e aplicações, oferecendo produtos para entrega de conteúdo, proteção contra DDoS, segurança de API e aplicativos, gerenciamento de bots, computação na borda (funções sem servidor/funções de borda) e inferência de IA na borda. A Akamai também fornece serviços de segurança focados em empresas (confiança zero, gerenciamento de identidade e acesso, acesso seguro à internet) e ferramentas de infraestrutura em nuvem/IA, e recentemente expandiu suas capacidades por meio de aquisições (por exemplo, LayerX) para adicionar controle de uso de IA baseado em navegador.

Descrição

• Optimize inference performance across the Akamai Inference Cloud • Collaborate closely with hardware performance engineers to deliver end-to-end optimization • Apply and evaluate quantization, distillation, and pruning techniques to optimize model performance while preserving accuracy • Design hardware-aware model placement and scheduling strategies to match models with optimal compute resources • Implement and tune speculative decoding, KV-cache optimization, and batching strategies to improve inference throughput and latency • Build benchmarking and profiling pipelines to measure model-layer performance across architectures, hardware, and serving configurations • Mentor and guide engineers on the team through code reviews, design discussions, and technical problem-solving • Collaborate with hardware performance engineers to identify and resolve end-to-end performance bottlenecks across the inference stack

🎯 Requisitos

• 12+ years of relevant experience with a Bachelor's or Master's degree in Computer Science, Machine Learning, or a related field • Possess hands-on experience optimizing LLM inference performance (quantization, speculative decoding, model compression, etc.) • Have a solid understanding of transformer architectures and how design choices impact latency, throughput, and accuracy • Possess experience with inference serving frameworks such as vLLM, TensorRT-LLM, Triton, or similar systems • Be proficient in Python and C++ with experience profiling and optimizing compute-intensive workloads • Have familiarity with hardware-aware optimization, including GPU/accelerator scheduling and memory management trade-offs.

🏖️ Benefícios

• Health insurance • 401K savings plan • Company holidays • Vacation (in the form of PTO) • Sick time • Family friendly benefits including parental leave • Employee assistance program with focus on mental and financial wellness

Candidatar-se

Vagas Similares

🕒 Abril 7

Payabli

11 - 50

💼 Consultoria

📣 Marketing

📦 Logística

Senior Software Engineer developing user interfaces for embedded payment infrastructure platform. Responsible for frontend application design, development, and integration with backend services.

🇺🇸 Estados Unidos – Remoto (EUA)

💰 $35.999.907 Series B - Payabli em 2025-06

⏰ Tempo Integral

🟠 Sênior

🧑‍💻 Engenheiro Full-stack

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Abril 7

Knowlej

1 - 10

💼 Consultoria

🏥 Saúde

📣 Marketing

Founding Product Engineer building core product for K–12 education platform. Collaborating with founder to define, build, and scale product with high ownership and influence.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $150.000 - $180.000 / ano

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

🧑‍💻 Engenheiro Full-stack

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Abril 7

GAI Consultants, Inc.

501 - 1000

💼 Consultoria

📦 Logística

🏭 Manufatura

Lead Grid Modernization Engineering efforts at GAI Consultants, Inc. Supporting microgrid and DER projects from feasibility to implementation.

🇺🇸 Estados Unidos – Remoto (EUA)

💰 Private equity em 2022-11

⏰ Tempo Integral

🟠 Sênior

🧑‍💻 Engenheiro Full-stack

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Abril 6

Vannevar Labs

11 - 50

💼 Consultoria

📦 Logística

🎖️ Defesa

Technical leader driving development and adoption of AI Agents platform at Vannevar. Innovating in the rapidly changing space of Agentic AI for defense technology.

🇺🇸 Estados Unidos – Remoto (EUA)

💰 $12.000.000 Series A em 2021-08

⏰ Tempo Integral

🟠 Sênior

🧑‍💻 Engenheiro Full-stack

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Abril 6

Mercury

201 - 500

💳 Fintech

💸 Finanças

☁️ SaaS

Senior Software Engineer developing Mercury's AI platform and enablement layer. Responsible for building and scaling technologies that enhance AI capabilities across the organization.

🗣️🇺🇸🇬🇧 Inglês obrigatório