Senior Software Engineer – Akamai Inference Cloud

🕒 April 13

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Akamai Technologies

Akamai Technologies

5001 - 10000 employees

🔒 Cybersecurity

🏢 Enterprise

📱 Media

Cybersecurity • Enterprise • Media

Akamai Technologies is a global edge platform and cloud services company that delivers content delivery, edge computing, and security solutions. The company operates one of the world’s largest distributed networks to accelerate and protect web, media, and application traffic, offering products for content delivery, DDoS protection, API and app security, bot management, edge compute (serverless/edge functions), and AI inference at the edge. Akamai also provides enterprise-focused security services (zero trust, identity and access management, secure internet access) and cloud/AI infrastructure tools, and has recently expanded capabilities through acquisitions (for example LayerX) to add browser-based AI usage control.

📋 Description

• Developing and maintaining prompt processing and tokenization pipelines that prepare inference requests for efficient model execution. • Implementing request routing, scheduling, and batching logic that optimizes throughput and latency across concurrent inference workloads. • Contributing to the integration of new model architectures and serving backends into the runtime framework. • Writing well-tested, well-documented code and participating in code reviews to maintain high engineering standards across the inference stack. • Supporting operational readiness through monitoring, debugging, and performance analysis of runtime components.

🎯 Requirements

• Have relevant experience and a Bachelor's degree or its equivalent in Computer Science or a related field. • Demonstrate proficiency in Python and at least one systems programming language such as C++, Go, or Rust. • Show understanding of natural language processing concepts including tokenization, encoding, and text processing pipelines. • Have familiarity with AI inference, model serving, or LLM deployment including inference frameworks (TensorRT, vLLM, TorchServe, Triton). • Demonstrate experience building high-throughput, low-latency data processing systems or services. • Show familiarity with Linux systems, containerized environments, and profiling or debugging tools. • Demonstrate a keen willingness to learn and grow within the AI inference and model serving field.

🏖️ Benefits

• Your health • Your finances • Your family • Your time at work • Your time pursuing other endeavors

Apply Now

Similar Jobs

🕒 March 27

Vodeno

201 - 500

💼 Consulting

🛡️ Insurance

💳 Fintech

Cloud Engineer joining Vodeno, innovators in Banking-as-a-Service, focusing on cloud-native technology and integration for financial solutions. Collaborating with teams to build fault-tolerant applications and infrastructure.

🗣️🇵🇱 Polish Required

Ansible

Apache

AWS

Azure

Cloud

ElasticSearch

Google Cloud Platform

Java

Kafka

Linux

Logstash

Puppet

Python

Redis

Terraform

VMware

Go