Principal Performance Engineer, Lead

đź•’ vor 6 Monaten

🍂 Massachusetts – Remote

infoinfo

đź’µ $169.300 - $304.700 / Jahr

⏰ Vollzeit

đźź  Senior

🧑‍💻 Full-Stack-Entwickler

đź‘» Geisterscore 41%

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of Akamai Technologies

Akamai Technologies

5001 - 10000 Mitarbeiter

đź”’ Cybersecurity

🏢 Unternehmen

📱 Medien

Cybersecurity • Enterprise • Media

Akamai Technologies ist eine globale Plattform für Edge- und Cloud-Dienstleistungen, die Lösungen für die Bereitstellung von Inhalten, Edge Computing und Sicherheit anbietet. Das Unternehmen betreibt eines der weltweit größten verteilten Netzwerke zur Beschleunigung und zum Schutz von Web-, Medien- und Anwendungsverkehr. Akamai bietet Produkte für die Bereitstellung von Inhalten, DDoS-Schutz, API- und App-Sicherheit, Bot-Management, Edge Computing (serverlose/Edge-Funktionen) und KI-Inferenz am Edge an. Zudem stellt Akamai unternehmensfokussierte Sicherheitsdienste bereit (Zero Trust, Identitäts- und Zugangsmanagement, sicherer Internetzugang) sowie Tools für Cloud-/KI-Infrastruktur. Kürzlich hat Akamai seine Fähigkeiten durch Akquisitionen (zum Beispiel LayerX) erweitert, um KI-Nutzungen im Browser zu steuern.

Beschreibung

• Optimize inference performance across the Akamai Inference Cloud • Collaborate closely with hardware performance engineers to deliver end-to-end optimization • Apply and evaluate quantization, distillation, and pruning techniques to optimize model performance while preserving accuracy • Design hardware-aware model placement and scheduling strategies to match models with optimal compute resources • Implement and tune speculative decoding, KV-cache optimization, and batching strategies to improve inference throughput and latency • Build benchmarking and profiling pipelines to measure model-layer performance across architectures, hardware, and serving configurations • Mentor and guide engineers on the team through code reviews, design discussions, and technical problem-solving • Collaborate with hardware performance engineers to identify and resolve end-to-end performance bottlenecks across the inference stack

🎯 Anforderungen

• 12+ years of relevant experience with a Bachelor's or Master's degree in Computer Science, Machine Learning, or a related field • Possess hands-on experience optimizing LLM inference performance (quantization, speculative decoding, model compression, etc.) • Have a solid understanding of transformer architectures and how design choices impact latency, throughput, and accuracy • Possess experience with inference serving frameworks such as vLLM, TensorRT-LLM, Triton, or similar systems • Be proficient in Python and C++ with experience profiling and optimizing compute-intensive workloads • Have familiarity with hardware-aware optimization, including GPU/accelerator scheduling and memory management trade-offs.

🏖️ Vorteile

• Health insurance • 401K savings plan • Company holidays • Vacation (in the form of PTO) • Sick time • Family friendly benefits including parental leave • Employee assistance program with focus on mental and financial wellness

Jetzt Bewerben

Ähnliche Jobs

đź•’ vor 6 Monaten

Payabli

11 - 50

đź’Ľ Beratung

📣 Marketing

📦 Logistik

Senior Software Engineer developing user interfaces for embedded payment infrastructure platform. Responsible for frontend application design, development, and integration with backend services.

🇺🇸 Vereinigte Staaten – Remote

💰 €35.999.907 Series B - Payabli im 2025-06

⏰ Vollzeit

đźź  Senior

🧑‍💻 Full-Stack-Entwickler

🗣️🇺🇸🇬🇧 Englisch erforderlich

đź•’ vor 6 Monaten

GAI Consultants, Inc.

501 - 1000

đź’Ľ Beratung

📦 Logistik

🏭 Fertigung

Lead Grid Modernization Engineering efforts at GAI Consultants, Inc. Supporting microgrid and DER projects from feasibility to implementation.

🇺🇸 Vereinigte Staaten – Remote

đź’° Private equity im 2022-11

⏰ Vollzeit

đźź  Senior

🧑‍💻 Full-Stack-Entwickler

🗣️🇺🇸🇬🇧 Englisch erforderlich

đź•’ vor 6 Monaten

Vannevar Labs

11 - 50

đź’Ľ Beratung

📦 Logistik

🎖️ Verteidigung

Technical leader driving development and adoption of AI Agents platform at Vannevar. Innovating in the rapidly changing space of Agentic AI for defense technology.

🇺🇸 Vereinigte Staaten – Remote

💰 €12.000.000 Series A im 2021-08

⏰ Vollzeit

đźź  Senior

🧑‍💻 Full-Stack-Entwickler

🗣️🇺🇸🇬🇧 Englisch erforderlich

đź•’ vor 6 Monaten

Silver.dev

1 - 10

🎯 Rekrutierung

👥 HR Tech

🤝 B2B

Fullstack Engineers with ambition to prove technical proficiency and compete with U.S. talent. Join Silver.dev for potential U.S. immigration sponsorship through exceptional skill validation.

🇺🇸 Vereinigte Staaten – Remote

⏰ Vollzeit

🟡 Mittelstufe

đźź  Senior

🧑‍💻 Full-Stack-Entwickler

🗣️🇺🇸🇬🇧 Englisch erforderlich

đź•’ vor 6 Monaten

OnePay

501 - 1000

đź’ł Fintech

🏦 Bankwesen

₿ Crypto

Software Engineer at OnePay developing backend products and features for a fintech platform. Collaborating with engineers and product managers to enhance customer financial experiences.

🇺🇸 Vereinigte Staaten – Remote

đź’µ $125.000 - $190.000 / Jahr

💰 €300.000.000 Series unknown im 2025-01

⏰ Vollzeit

🟡 Mittelstufe

đźź  Senior

🧑‍💻 Full-Stack-Entwickler

🦅 H1B-Visum-Sponsor

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich