Search Remote Jobs

Senior Site Reliability Engineer

🕒 March 30

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Akamai Technologies

Akamai Technologies

5001 - 10000 employees

🔒 Cybersecurity

🏱 Enterprise

đŸ“± Media

Cybersecurity ‱ Enterprise ‱ Media

Akamai Technologies is a global edge platform and cloud services company that delivers content delivery, edge computing, and security solutions. The company operates one of the world’s largest distributed networks to accelerate and protect web, media, and application traffic, offering products for content delivery, DDoS protection, API and app security, bot management, edge compute (serverless/edge functions), and AI inference at the edge. Akamai also provides enterprise-focused security services (zero trust, identity and access management, secure internet access) and cloud/AI infrastructure tools, and has recently expanded capabilities through acquisitions (for example LayerX) to add browser-based AI usage control.

📋 Description

‱ Own reliability workstreams for Akamai's serverless inference platform ‱ Build automation and tooling ‱ Contribute to architecture and operational decisions ‱ Take ownership of critical reliability problems end-to-end ‱ Partner with product engineering teams ‱ Develop expertise in GPU infrastructure, Kubernetes at scale, and AI inference workloads ‱ Build and maintain observability for AI workloads, including telemetry, dashboards, alerts, SLO/SLI tracking ‱ Write automation and tooling to reduce operational toil, improve deployment safety, and accelerate incident response ‱ Integrate AI workloads into Akamai's incident management processes ‱ Build and maintain CI/CD integrations, deployment safety checks, and rollback automation ‱ Collaborate with product engineering teams to improve reliability and ensure operational readiness for product releases ‱ Contribute to capacity planning, autoscaling configuration, and workload scheduling for AI compute infrastructure

🎯 Requirements

‱ 5+ years of experience in SRE, infrastructure engineering, or platform engineering, working with large-scale distributed systems ‱ Extensive experience with Kubernetes and containerization at scale ‱ Experience defining SLOs and working with observability tools such as Prometheus, Grafana, and distributed tracing ‱ Coding ability in Python or Go for automation and tooling, with experience in CI/CD pipelines, deployment safety, and infrastructure-as-code ‱ Interest in or experience with AI/ML infrastructure, model serving, or GPU workloads ‱ Ability to take ownership of problems and drive them to resolution independently

đŸ–ïž Benefits

‱ Healthcare ‱ 401K savings plan ‱ Company holidays ‱ Vacation (in the form of PTO) ‱ Sick time ‱ Family friendly benefits including parental leave ‱ Employee assistance program focusing on mental and financial wellness ‱ Flexible working arrangements

Apply Now

Similar Jobs

🕒 March 30

Expert Executive Recruiters (EER Global)

51 - 200

đŸ’Œ Consulting

📩 Logistics

đŸ„ Healthcare

Senior DevOps/Infra Engineer needed to design, automate, and secure high-load infrastructure. Role involves kernel tuning, VPNs, monitoring, and CI/CD in a collaborative remote team.

🕒 March 28

Close

51 - 200

đŸ’Œ Consulting

📣 Marketing

☁ SaaS

Join Infrastructure Team at Close as a Site Reliability Engineer. Work on robust systems supporting a modern communication-focused CRM service for small scaling businesses.

🕒 March 27

Red River

501 - 1000

đŸ’Œ Consulting

📩 Logistics

Senior Wireless Deployment Engineer managing deployment, optimization, and lifecycle of enterprise wireless networks. Providing technical leadership and support for Aruba and Juniper Mist solutions.

🕒 March 27

SGNL

11 - 50

🔒 Cybersecurity

🔐 Security

☁ SaaS

Senior DevOps Engineer at SGNL solving authorization challenges for major companies. Collaborating and leading teams in a dynamic, scale-oriented environment.

🕒 March 27

Cority

201 - 500

đŸ„ Healthcare

📩 Logistics

đŸ’Œ Consulting

Sr. DevOps Engineer working to deploy and operate systems at Cority, the global EHS software provider. Collaborating with engineering for continuous delivery and monitoring towards operational excellence.