Principal Performance Engineer, Lead

🕒 il y a 4 mois

🗣️🇺🇸🇬🇧 Anglais requis

Postuler Maintenant
Trouver des Emplois à Distance Similaires

📊 Vérifiez votre score de CV pour ce poste

Améliorez vos chances d'obtenir un entretien en vérifiant votre score de CV avant de postuler.

Logo of Akamai Technologies

Akamai Technologies

5001 - 10000 employés

🔒 Cybersecurity

🏢 Entreprise

📱 Médias

Cybersecurity • Enterprise • Media

Akamai Technologies est une entreprise mondiale de services de cloud et de plateforme à la périphérie, offrant des solutions de diffusion de contenu, de calcul en périphérie et de sécurité. L'entreprise exploite l'un des plus grands réseaux distribués au monde pour accélérer et protéger le trafic web, média et applicatif, en proposant des produits pour la diffusion de contenu, la protection contre les attaques DDoS, la sécurité des API et des applications, la gestion des bots, le calcul en périphérie (fonctions sans serveur/périphérie) et l'inférence IA en périphérie. Akamai fournit également des services de sécurité axés sur l'entreprise (confiance zéro, gestion des identités et des accès, accès internet sécurisé) et des outils d'infrastructure cloud/IA, ayant récemment élargi ses capacités grâce à des acquisitions (par exemple LayerX) pour intégrer le contrôle d'utilisation de l'IA basé sur navigateur.

Description

• Optimize inference performance across the Akamai Inference Cloud • Collaborate closely with hardware performance engineers to deliver end-to-end optimization • Apply and evaluate quantization, distillation, and pruning techniques to optimize model performance while preserving accuracy • Design hardware-aware model placement and scheduling strategies to match models with optimal compute resources • Implement and tune speculative decoding, KV-cache optimization, and batching strategies to improve inference throughput and latency • Build benchmarking and profiling pipelines to measure model-layer performance across architectures, hardware, and serving configurations • Mentor and guide engineers on the team through code reviews, design discussions, and technical problem-solving • Collaborate with hardware performance engineers to identify and resolve end-to-end performance bottlenecks across the inference stack

🎯 Exigences

• 12+ years of relevant experience with a Bachelor's or Master's degree in Computer Science, Machine Learning, or a related field • Possess hands-on experience optimizing LLM inference performance (quantization, speculative decoding, model compression, etc.) • Have a solid understanding of transformer architectures and how design choices impact latency, throughput, and accuracy • Possess experience with inference serving frameworks such as vLLM, TensorRT-LLM, Triton, or similar systems • Be proficient in Python and C++ with experience profiling and optimizing compute-intensive workloads • Have familiarity with hardware-aware optimization, including GPU/accelerator scheduling and memory management trade-offs.

🏖️ Avantages

• Health insurance • 401K savings plan • Company holidays • Vacation (in the form of PTO) • Sick time • Family friendly benefits including parental leave • Employee assistance program with focus on mental and financial wellness

Postuler Maintenant

Emplois Similaires

🕒 il y a 4 mois

Payabli

11 - 50

💼 Conseil

📣 Marketing

📦 Logistique

Senior Software Engineer developing user interfaces for embedded payment infrastructure platform. Responsible for frontend application design, development, and integration with backend services.

🇺🇸 États-Unis – Télétravail

💰 €35 999 907 Series B - Payabli en 2025-06

⏰ Temps Plein

🟠 Senior

🧑‍💻 Développeur Full-Stack

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 4 mois

Knowlej

1 - 10

💼 Conseil

🏥 Santé

📣 Marketing

Founding Product Engineer building core product for K–12 education platform. Collaborating with founder to define, build, and scale product with high ownership and influence.

🇺🇸 États-Unis – Télétravail

💵 $150 000 - $180 000 / an

⏰ Temps Plein

🟡 Intermédiaire

🟠 Senior

🧑‍💻 Développeur Full-Stack

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 4 mois

GAI Consultants, Inc.

501 - 1000

💼 Conseil

📦 Logistique

🏭 Fabrication

Lead Grid Modernization Engineering efforts at GAI Consultants, Inc. Supporting microgrid and DER projects from feasibility to implementation.

🇺🇸 États-Unis – Télétravail

💰 Private equity en 2022-11

⏰ Temps Plein

🟠 Senior

🧑‍💻 Développeur Full-Stack

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 4 mois

Vannevar Labs

11 - 50

💼 Conseil

📦 Logistique

🎖️ Défense

Technical leader driving development and adoption of AI Agents platform at Vannevar. Innovating in the rapidly changing space of Agentic AI for defense technology.

🇺🇸 États-Unis – Télétravail

💰 €12 000 000 Series A en 2021-08

⏰ Temps Plein

🟠 Senior

🧑‍💻 Développeur Full-Stack

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 4 mois

Silver.dev

1 - 10

🎯 Recrutement

👥 RH Tech

🤝 B2B

Fullstack Engineers with ambition to prove technical proficiency and compete with U.S. talent. Join Silver.dev for potential U.S. immigration sponsorship through exceptional skill validation.

🇺🇸 États-Unis – Télétravail

⏰ Temps Plein

🟡 Intermédiaire

🟠 Senior

🧑‍💻 Développeur Full-Stack

🗣️🇺🇸🇬🇧 Anglais requis