Site Reliability Engineer

đŸ”„ 0 minutes ago

🇼🇳 India – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

đŸ‘» Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Akamai Technologies

Akamai Technologies

5001 - 10000 employees

🔒 Cybersecurity

💰 Post-IPO Equity on 2001-07

Cloud Computing ‱ Cybersecurity ‱ Content Delivery

Akamai Technologies is a leading cloud services provider that specializes in delivering security, cloud computing, and content delivery solutions. It offers a range of services such as API security, DDoS protection, and performance optimization for web applications, ensuring secure and reliable user experiences. With a robust global infrastructure, Akamai empowers businesses to streamline their digital presence while safeguarding against various cyber threats and enhancing application performance.

📋 Description

‱ Own reliability and performance investigations across Akamai's global edge platform, media delivery, and web delivery systems ‱ Diagnose problems across application, platform, network, and operating-system layers using logs, metrics, traces, and diagnostic tools ‱ Collaborate with Product, Engineering, Support, Network, and senior SRE teams to determine root causes and apply lasting solutions ‱ Use telemetry, SLIs, SLOs, KPIs, dashboards, and alerts to evaluate system health and customer impact ‱ Analyze platform behavior, traffic patterns, and system bottlenecks to improve performance, scalability, and resilience ‱ Create scripts, automation, internal tools, and workflows to reduce operational effort and improve diagnostics and incident response ‱ Apply AI-powered analysis and create workflows with precision, safety, and human evaluation ‱ Act as a technical escalation resource for global support and resolve critical issues across delivery products and platform technologies

🎯 Requirements

‱ Expertise in Computer Science, Engineering, or related fields, or substantial industry experience in large-scale SRE roles ‱ Ability to evaluate technical issues, assess system behavior, and derive insights from evidence ‱ Understanding of caching, proxies, TLS, TCP/IP, DNS, and HTTP/HTTPS architectures ‱ Proficiency with Linux or Unix systems, command-line utilities, and foundational diagnostic methods ‱ Proficiency with structured data and telemetry or foundational SQL capabilities ‱ Ability to read and analyze C++ code ‱ Experience with large-scale distributed systems and SRE environments

đŸ–ïž Benefits

‱ Health, well-being, and financial benefits ‱ FlexBase flexible workplace program ‱ Ability to work at home, in an office, or a combination of both ‱ Exceptional workplace experience support

Apply Now

Similar Jobs

🕒 4 days ago

Study Now

51 - 200

đŸ’Œ Consulting

📣 Marketing

📚 Education

DevOps Engineer making Study Now’s student-recruitment platform reliable, secure, and recoverable. Owning cloud infrastructure, backups, CI/CD, monitoring, and UK GDPR alignment.

Angular

Ansible

AWS

Cloud

DNS

Docker

JavaScript

Linux

MongoDB

Node.js

Terraform

🕒 5 days ago

Granicus

501 - 1000

đŸ›ïž Government

☁ SaaS

📋 Compliance

Site Reliability Engineer modernizing Granicus’s government technology platforms through AIOps, observability, and automation. Improving reliability, incident response, and resilient cloud operations.

Ansible

AWS

Azure

Cloud

Distributed Systems

ElasticSearch

Google Cloud Platform

ITSM

Kubernetes

Linux

Logstash

Terraform

Unix

🕒 5 days ago

4Pharma Ltd

11 - 50

đŸ’Œ Consulting

đŸœïž Food & Beverage

📩 Logistics

Senior DevOps Engineer building AWS, Kubernetes and CI/CD infrastructure for BC Platforms’ global healthcare data and analytics platform. Driving reliability, security and automation.

AWS

Azure

Cloud

Grafana

Kubernetes

Linux

Prometheus

🕒 6 days ago

MFSG

11 - 50

🏭 Manufacturing

🔧 Hardware

🚗 Transport

Site Reliability Engineer automating reliable, compliant digital banking platforms for MFSG Technologies. Managing CI/CD, observability, incident response, and resilient production deployments.

Ansible

AWS

Azure

Cloud

Docker

Kubernetes

Python

Terraform

🕒 6 days ago

iCert Global

51 - 200

đŸ’Œ Consulting

📣 Marketing

📚 Education

Lead SRE managing Azure infrastructure, AKS, observability, and major incidents for Icertis’s AI-powered contract intelligence platform. Driving automation, reliability, and cloud-native operations.

AWS

Azure

Cloud

Distributed Systems

Docker

Kubernetes

Python

ServiceNow

Terraform