Search Remote Jobs

Senior Site Reliability Engineer

🕒 July 16

🇼🇳 India – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

đŸ‘» Ghost score 16%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Akamai Technologies

Akamai Technologies

5001 - 10000 employees

🔒 Cybersecurity

💰 Post-IPO Equity on 2001-07

Cloud Computing ‱ Cybersecurity ‱ Content Delivery

Akamai Technologies is a leading cloud services provider that specializes in delivering security, cloud computing, and content delivery solutions. It offers a range of services such as API security, DDoS protection, and performance optimization for web applications, ensuring secure and reliable user experiences. With a robust global infrastructure, Akamai empowers businesses to streamline their digital presence while safeguarding against various cyber threats and enhancing application performance.

📋 Description

‱ Leading complex reliability and performance investigations across Akamai's global edge, media delivery, and web delivery platforms. ‱ Troubleshooting critical distributed systems issues spanning application, platform, network, and operating system layers, serving as the highest technical escalation point. ‱ Partnering with Engineering, Product, Support, and Network teams to identify root causes and deliver scalable, long-term solutions that improve platform reliability. ‱ Designing and improving observability through SLIs, SLOs, KPIs, telemetry, dashboards, and alerts to identify and address customer-impacting issues. ‱ Analyzing platform performance, traffic patterns, and system bottlenecks to improve scalability, resilience, and overall service reliability. ‱ Developing automation, internal tools, AI-assisted diagnostics, and self-service workflows to streamline operations, reduce manual effort, and accelerate incident response. ‱ Enhancing operational excellence through reliability-centered architecture reviews, post-incident analysis, continuous improvements, and offering off-hours support during critical incidents as needed.

🎯 Requirements

‱ Possess Bachelors in CS/Engineering or a related field with 6 years of industry experience in large-scale SRE/Systems Infrastructure roles. ‱ Have logical reasoning skills diagnosing complex performance bottlenecks, data integrity anomalies, and system failure modes in distributed environments. ‱ Have understanding of internet technologies and foundational networking concepts, including caching, proxies, TLS, TCP/IP, DNS, and HTTP/HTTPS architectures. ‱ Have foundation in Linux/Unix administration, diagnostic tools, and low-level environment troubleshooting. ‱ Be able to retrieve data, analyze telemetry streams, and troubleshoot platform data integrity issues through SQL queries. ‱ Have experience developing automation tools using languages like Python, Bash, or Go. ‱ Demonstrate expertise in AI models and focus on implementing agentic workflows to reduce operational inefficiencies effectively.

đŸ–ïž Benefits

‱ We support your health, well-being, finances, and life beyond work. See our benefits. ‱ FlexBase adapts to your job's needs ‱ Akamai's FlexBase program is yet another way we show our commitment to providing employees with an exceptional workplace experience. It's not about telling employees where to work; it's about supporting employees to do their best work.

Apply Now

Similar Jobs

🕒 July 9

DBSync

51 - 200

đŸ’Œ Consulting

đŸ„ Healthcare

📩 Logistics

Forward Deployment Engineer at DBSync solving tasks for Cloud technology users and ensuring customer success through technical expertise.

🇼🇳 India – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 8

MariaDB

201 - 500

🏱 Enterprise

QA and release engineer testing MariaDB MaxScale, a database proxy for MariaDB clusters. Managing releases, Linux packages, CI/CD automation, and build infrastructure.

🇼🇳 India – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 8

Pythian

201 - 500

đŸ’Œ Consulting

đŸ„ Healthcare

📩 Logistics

Site Reliability Engineer at Pythian focusing on operating large-scale distributed systems. Responsible for designing, deploying, and operating infrastructure with strong collaboration across teams.

🇼🇳 India – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 7

Resilinc

201 - 500

đŸ’Œ Consulting

📩 Logistics

đŸ„ Healthcare

Site Reliability Engineer responsible for platform availability and automation in cloud environments at Resilinc. Focused on leveraging agentic AI for impactful supply chain solutions.

🇼🇳 India – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 1

Empower

10,000+ employees

💾 Finance

💳 Fintech

đŸ‘„ B2C

Senior SRE architecting reliable AWS and Kubernetes infrastructure for Empower’s financial services platform. Leading incident response, automation, observability, security, and engineer mentorship.

🇼🇳 India – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)