Principal Performance Engineer, Lead

Stelle nicht auf LinkedIn

🕒 vor 3 Monaten

🗣️🇺🇸🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of Akamai Technologies

Akamai Technologies

5001 - 10000 Mitarbeiter

🔒 Cybersecurity

🏢 Unternehmen

📱 Medien

Cybersecurity • Enterprise • Media

Akamai Technologies ist eine globale Plattform für Edge- und Cloud-Dienstleistungen, die Lösungen für die Bereitstellung von Inhalten, Edge Computing und Sicherheit anbietet. Das Unternehmen betreibt eines der weltweit größten verteilten Netzwerke zur Beschleunigung und zum Schutz von Web-, Medien- und Anwendungsverkehr. Akamai bietet Produkte für die Bereitstellung von Inhalten, DDoS-Schutz, API- und App-Sicherheit, Bot-Management, Edge Computing (serverlose/Edge-Funktionen) und KI-Inferenz am Edge an. Zudem stellt Akamai unternehmensfokussierte Sicherheitsdienste bereit (Zero Trust, Identitäts- und Zugangsmanagement, sicherer Internetzugang) sowie Tools für Cloud-/KI-Infrastruktur. Kürzlich hat Akamai seine Fähigkeiten durch Akquisitionen (zum Beispiel LayerX) erweitert, um KI-Nutzungen im Browser zu steuern.

Beschreibung

• Optimize inference performance across the Akamai Inference Cloud • Collaborate closely with hardware performance engineers to deliver end-to-end optimization • Apply and evaluate quantization, distillation, and pruning techniques to optimize model performance while preserving accuracy • Design hardware-aware model placement and scheduling strategies to match models with optimal compute resources • Implement and tune speculative decoding, KV-cache optimization, and batching strategies to improve inference throughput and latency • Build benchmarking and profiling pipelines to measure model-layer performance across architectures, hardware, and serving configurations • Mentor and guide engineers on the team through code reviews, design discussions, and technical problem-solving • Collaborate with hardware performance engineers to identify and resolve end-to-end performance bottlenecks across the inference stack

🎯 Anforderungen

• 12+ years of relevant experience with a Bachelor's or Master's degree in Computer Science, Machine Learning, or a related field • Possess hands-on experience optimizing LLM inference performance (quantization, speculative decoding, model compression, etc.) • Have a solid understanding of transformer architectures and how design choices impact latency, throughput, and accuracy • Possess experience with inference serving frameworks such as vLLM, TensorRT-LLM, Triton, or similar systems • Be proficient in Python and C++ with experience profiling and optimizing compute-intensive workloads • Have familiarity with hardware-aware optimization, including GPU/accelerator scheduling and memory management trade-offs.

🏖️ Vorteile

• Health insurance • 401K savings plan • Company holidays • Vacation (in the form of PTO) • Sick time • Family friendly benefits including parental leave • Employee assistance program with focus on mental and financial wellness

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 3 Monaten

Payabli

11 - 50

💼 Beratung

📣 Marketing

📦 Logistik

Senior Software Engineer developing user interfaces for embedded payment infrastructure platform. Responsible for frontend application design, development, and integration with backend services.

🇺🇸 Vereinigte Staaten – Remote

💰 €35.999.907 Series B - Payabli im 2025-06

⏰ Vollzeit

🟠 Senior

🧑‍💻 Full-Stack-Entwickler

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 3 Monaten

Knowlej

1 - 10

💼 Beratung

🏥 Gesundheitswesen

📣 Marketing

Founding Product Engineer building core product for K–12 education platform. Collaborating with founder to define, build, and scale product with high ownership and influence.

🇺🇸 Vereinigte Staaten – Remote

💵 $150.000 - $180.000 / Jahr

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

🧑‍💻 Full-Stack-Entwickler

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 3 Monaten

GAI Consultants, Inc.

501 - 1000

💼 Beratung

📦 Logistik

🏭 Fertigung

Lead Grid Modernization Engineering efforts at GAI Consultants, Inc. Supporting microgrid and DER projects from feasibility to implementation.

🇺🇸 Vereinigte Staaten – Remote

💰 Private equity im 2022-11

⏰ Vollzeit

🟠 Senior

🧑‍💻 Full-Stack-Entwickler

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 3 Monaten

Vannevar Labs

11 - 50

💼 Beratung

📦 Logistik

🎖️ Verteidigung

Technical leader driving development and adoption of AI Agents platform at Vannevar. Innovating in the rapidly changing space of Agentic AI for defense technology.

🇺🇸 Vereinigte Staaten – Remote

💰 €12.000.000 Series A im 2021-08

⏰ Vollzeit

🟠 Senior

🧑‍💻 Full-Stack-Entwickler

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 3 Monaten

Mercury

201 - 500

💳 Fintech

💸 Finanzen

☁️ SaaS

Senior Software Engineer developing Mercury's AI platform and enablement layer. Responsible for building and scaling technologies that enhance AI capabilities across the organization.

🗣️🇺🇸🇬🇧 Englisch erforderlich