Senior Site Reliability Engineer, Linux

🕒 May 21

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Akamai Technologies

Akamai Technologies

5001 - 10000 employees

🔒 Cybersecurity

🏱 Enterprise

đŸ“± Media

Cybersecurity ‱ Enterprise ‱ Media

Akamai Technologies is a global edge platform and cloud services company that delivers content delivery, edge computing, and security solutions. The company operates one of the world’s largest distributed networks to accelerate and protect web, media, and application traffic, offering products for content delivery, DDoS protection, API and app security, bot management, edge compute (serverless/edge functions), and AI inference at the edge. Akamai also provides enterprise-focused security services (zero trust, identity and access management, secure internet access) and cloud/AI infrastructure tools, and has recently expanded capabilities through acquisitions (for example LayerX) to add browser-based AI usage control.

📋 Description

‱ Collaborating with our support, operations and engineering teams, investigate and troubleshoot complex problems. ‱ Developing processes, plans, and infrastructure to deploy new software components and updates safely and efficiently at scale. ‱ Participating in on-call rotations, guiding restoration and repair of service-impacting issues. ‱ Improving our system monitoring and analysis platform to speed error detection and remediation, enhancing performance and reliability.

🎯 Requirements

‱ Have 7+ years of relevant experience and a Bachelors degree in Computer Science or related field ‱ Possess expert level experience in a Systems engineering or DevOps or Software engineering role, working with large scale distributed systems. ‱ Can troubleshoot any kind of systemic issues and develop large scale automations. ‱ Demonstrate proficiency in Python or Golang and hands-on experience with SaltStack, Ansible, and Terraform for infrastructure automation. ‱ Demonstrate expertise with observability or monitoring tools like Prometheus, Grafana, ELK/OpenSearch, Datadog, and Splunk. ‱ Gain experience with any cloud platform, such as AWS, GCP, Azure, or an equivalent alternative.

đŸ–ïž Benefits

‱ We support your health, well-being, finances, and life beyond work. See our benefits. ‱ FlexBase adapts to your job's needs. Akamai's FlexBase program is yet another way we show our commitment to providing employees with an exceptional workplace experience. It's not about telling employees where to work; it's about supporting employees to do their best work.

Apply Now

Similar Jobs

🕒 May 19

Fortive

10,000+ employees

đŸ„ Healthcare

🏭 Manufacturing

📩 Logistics

Site Reliability Engineer ensuring the availability and performance of a customer-facing platform. Collaborating closely with DevOps, DBA, and Development teams for infrastructure provision and maintenance.

Ansible

ASP.NET

AWS

Azure

Cloud

Docker

Jenkins

Kubernetes

Linux

SQL

Terraform

🕒 May 13

Shuru

51 - 200

đŸ€– Artificial Intelligence

đŸ€ B2B

🏱 Enterprise

Senior DevOps Engineer at Shuru Technologies enhancing cloud platform infrastructure. Collaborating with teams for scalable solutions and operational readiness in a remote-first environment.

AWS

Azure

Cloud

Google Cloud Platform

Kubernetes

Oracle

Postgres

Redis

SQL

Terraform

🕒 May 12

Volvo Cars

10,000+ employees

🏭 Manufacturing

🚗 Transport

🚘 Automotive

Salesforce Release Engineer driving digital innovation at Volvo Cars. Managing Salesforce release lifecycle across global teams and developing cutting-edge technology solutions for the automotive industry.

🕒 May 12

Akamai Technologies

5001 - 10000

🔒 Cybersecurity

Site Reliability Engineer II managing cloud platforms and enabling advanced computing for customers at Akamai. Collaborating with teams to maintain SLO-driven reliability and secure services.

Ansible

Chef

Distributed Systems

DNS

Grafana

Jenkins

Linux

Prometheus

Python

SaltStack

Shell Scripting

Terraform

🕒 May 4

Socure

501 - 1000

đŸ’Œ Consulting

đŸ„ Healthcare

⚖ Legal

Site Reliability Engineer responsible for managing AWS infrastructure and improving Kubernetes platforms. Utilize strong observability and automation to ensure reliability and performance.

AWS

Kubernetes

Python

Terraform

Go