Principal Performance Engineer, Lead

Job not on LinkedIn

đź•’ April 7

🍂 Massachusetts – Remote

infoinfo

đź’µ $169.3k - $304.7k / year

⏰ Full Time

đźź  Senior

🧑‍💻 Full-stack Engineer

đź‘» Ghost score 41%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Akamai Technologies

Akamai Technologies

5001 - 10000 employees

đź”’ Cybersecurity

🏢 Enterprise

📱 Media

Cybersecurity • Enterprise • Media

Akamai Technologies is a global edge platform and cloud services company that delivers content delivery, edge computing, and security solutions. The company operates one of the world’s largest distributed networks to accelerate and protect web, media, and application traffic, offering products for content delivery, DDoS protection, API and app security, bot management, edge compute (serverless/edge functions), and AI inference at the edge. Akamai also provides enterprise-focused security services (zero trust, identity and access management, secure internet access) and cloud/AI infrastructure tools, and has recently expanded capabilities through acquisitions (for example LayerX) to add browser-based AI usage control.

đź“‹ Description

• Optimize inference performance across the Akamai Inference Cloud • Collaborate closely with hardware performance engineers to deliver end-to-end optimization • Apply and evaluate quantization, distillation, and pruning techniques to optimize model performance while preserving accuracy • Design hardware-aware model placement and scheduling strategies to match models with optimal compute resources • Implement and tune speculative decoding, KV-cache optimization, and batching strategies to improve inference throughput and latency • Build benchmarking and profiling pipelines to measure model-layer performance across architectures, hardware, and serving configurations • Mentor and guide engineers on the team through code reviews, design discussions, and technical problem-solving • Collaborate with hardware performance engineers to identify and resolve end-to-end performance bottlenecks across the inference stack

🎯 Requirements

• 12+ years of relevant experience with a Bachelor's or Master's degree in Computer Science, Machine Learning, or a related field • Possess hands-on experience optimizing LLM inference performance (quantization, speculative decoding, model compression, etc.) • Have a solid understanding of transformer architectures and how design choices impact latency, throughput, and accuracy • Possess experience with inference serving frameworks such as vLLM, TensorRT-LLM, Triton, or similar systems • Be proficient in Python and C++ with experience profiling and optimizing compute-intensive workloads • Have familiarity with hardware-aware optimization, including GPU/accelerator scheduling and memory management trade-offs.

🏖️ Benefits

• Health insurance • 401K savings plan • Company holidays • Vacation (in the form of PTO) • Sick time • Family friendly benefits including parental leave • Employee assistance program with focus on mental and financial wellness

Apply Now

Similar Jobs

đź•’ April 7

Payabli

11 - 50

đź’Ľ Consulting

📣 Marketing

📦 Logistics

Senior Software Engineer developing user interfaces for embedded payment infrastructure platform. Responsible for frontend application design, development, and integration with backend services.

🇺🇸 United States – Remote

đź’° $36M Series B - Payabli on 2025-06

⏰ Full Time

đźź  Senior

🧑‍💻 Full-stack Engineer

đź•’ April 7

GAI Consultants, Inc.

501 - 1000

đź’Ľ Consulting

📦 Logistics

🏭 Manufacturing

Lead Grid Modernization Engineering efforts at GAI Consultants, Inc. Supporting microgrid and DER projects from feasibility to implementation.

🇺🇸 United States – Remote

đź’° Private equity on 2022-11

⏰ Full Time

đźź  Senior

🧑‍💻 Full-stack Engineer

đź•’ April 6

Vannevar Labs

11 - 50

đź’Ľ Consulting

📦 Logistics

🎖️ Defense

Technical leader driving development and adoption of AI Agents platform at Vannevar. Innovating in the rapidly changing space of Agentic AI for defense technology.

🇺🇸 United States – Remote

đź’° $12M Series A on 2021-08

⏰ Full Time

đźź  Senior

🧑‍💻 Full-stack Engineer

đź•’ April 6

Silver.dev

1 - 10

🎯 Recruiter

👥 HR Tech

🤝 B2B

Fullstack Engineers with ambition to prove technical proficiency and compete with U.S. talent. Join Silver.dev for potential U.S. immigration sponsorship through exceptional skill validation.

🇺🇸 United States – Remote

⏰ Full Time

🟡 Mid-level

đźź  Senior

🧑‍💻 Full-stack Engineer

đź•’ April 5

Stride, Inc.

5001 - 10000

đź’Ľ Consulting

🏥 Healthcare

📚 Education

Senior Full Stack Developer at Stride, enhancing tutoring platforms for B2B and B2C. Collaborate in agile teams to drive innovation in online education.

🇺🇸 United States – Remote

đź’µ $66.4k - $170k / year

⏰ Full Time

đźź  Senior

🧑‍💻 Full-stack Engineer