Senior Site Reliability Engineer – Compute Platform Services Team

🔥 0 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Akamai Technologies

Akamai Technologies

5001 - 10000 employees

🔒 Cybersecurity

💰 Post-IPO Equity on 2001-07

Cloud Computing • Cybersecurity • Content Delivery

Akamai Technologies is a leading cloud services provider that specializes in delivering security, cloud computing, and content delivery solutions. It offers a range of services such as API security, DDoS protection, and performance optimization for web applications, ensuring secure and reliable user experiences. With a robust global infrastructure, Akamai empowers businesses to streamline their digital presence while safeguarding against various cyber threats and enhancing application performance.

📋 Description

• Collaborate with support, operations, and engineering teams to investigate and troubleshoot complex problems • Develop processes, plans, and infrastructure to deploy new software components and updates safely and efficiently at scale • Participate in on-call rotations and guide restoration and repair of service-impacting issues • Improve system monitoring and analysis platforms to accelerate error detection and remediation • Enhance automation, operational excellence, and support for customer-facing applications and infrastructure

🎯 Requirements

• 7+ years of relevant experience • Bachelor's degree in Computer Science or a related field • Expert-level experience in Systems Engineering, DevOps, or Software Engineering • Experience working with large-scale distributed systems • Ability to troubleshoot systemic issues and develop large-scale automations • Proficiency in Python or Golang • Hands-on experience with SaltStack, Ansible, and Terraform • Expertise with observability or monitoring tools such as Prometheus, Grafana, ELK/OpenSearch, Datadog, and Splunk • Experience with a cloud platform such as AWS, GCP, Azure, or equivalent • Expertise in Linux systems administration, configuration management, performance optimization, and hardware engineering

🏖️ Benefits

• Health, well-being, financial, and life benefits • FlexBase flexible work arrangements: at home, in an office, or a combination of both

Apply Now

Similar Jobs

🕒 2 days ago

Ford Motor Company

10,000+ employees

📦 Logistics

💼 Consulting

📣 Marketing

Lead Site Reliability Engineer improving Ford’s automotive Marketing and Sales Tech platform from India. Enhancing observability, automation, resilience, and incident response.

Cloud

Google Cloud Platform

Java

JavaScript

Node.js

OpenShift

Python

Terraform

Go

🕒 3 days ago

Weekday

501 - 1000

👗 Fashion

🛒 Retail

🛍️ eCommerce

DevOps Engineer building scalable cloud infrastructure and CI/CD systems for a client. Managing containers, automation, monitoring, and reliability using AWS, Docker, and Kubernetes.

Ansible

AWS

Azure

Cloud

Docker

Google Cloud Platform

Grafana

Jenkins

Kafka

Kubernetes

Linux

Microservices

Prometheus

Python

Terraform

Go

🕒 3 days ago

Weekday (YC W21)

11 - 50

💼 Consulting

👥 HR Tech

☁️ SaaS

DevOps Engineer building cloud infrastructure, CI/CD pipelines, and containerized systems for a Weekday client. Automating provisioning, monitoring reliability, and deployment workflows using AWS, Docker, Kubernetes, and IaC tools.

Ansible

AWS

Azure

Cloud

Docker

Google Cloud Platform

Grafana

Jenkins

Kafka

Kubernetes

Linux

Microservices

Prometheus

Python

Terraform

Go

🕒 4 days ago

Empower

10,000+ employees

💸 Finance

💳 Fintech

👥 B2C

Senior DevOps Engineer automating AWS infrastructure and CI/CD for Empower’s financial SaaS products. Improving reliability, security, and developer productivity through tooling, cloud services, and automation.

Angular

Ansible

Apache

AWS

Azure

Chef

Cloud

DNS

Docker

DynamoDB

EC2

Google Cloud Platform

Java

JavaScript

Jenkins

jQuery

Kubernetes

Linux

Maven

MySQL

NGINX

Node.js

Python

React

Ruby

Splunk

Terraform

Go

🕒 4 days ago

Signalmash

51 - 200

💼 Consulting

📦 Logistics

🏥 Healthcare

DevOps Engineer owning Kubernetes, CI/CD, PostgreSQL, observability, and security for Signalmash’s cloud communications platform. Improving reliability, deployment speed, recovery, and infrastructure costs from India.

AWS

Azure

Cloud

Docker

Flux

Google Cloud Platform

Grafana

JavaScript

Kubernetes

Linux

Node.js

Postgres

Prometheus

Python

Shell Scripting