Site Reliability Engineer II

🔥 8 minutes ago

🇮🇳 India – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of MRSOOL | مرسول

MRSOOL | مرسول

201 - 500 employees

Founded 2015

🍽️ Food & Beverage

✈️ Travel

💼 Consulting

Food & Beverage • Travel • Consulting

MRSOOL is one of the largest delivery platforms in the region, offering an on-demand experience with high user ratings in both Apple's App Store and Google's Play store. MRSOOL provides a "order anything from anywhere" service backed by a large fleet of registered couriers. It enables businesses to access a vast user base, facilitating transformation to on-demand eCommerce and offering a flexible bidding system for service prices. MRSOOL also offers a personalized delivery experience with real-time tracking and communication with couriers, making it a leading option for ordering from local shops, groceries, and restaurants directly to your door.

📋 Description

• Collaborate with development teams to design and implement scalable infrastructure • Design and implement automated deployment and testing pipelines • Develop and maintain monitoring and alerting systems • Troubleshoot and escalate production incidents to minimize downtime and improve reliability • Continuously improve infrastructure and processes for scalability and efficiency • Participate in and own on-call rotations to provide 24/7 application support • Perform routine maintenance and upgrades • Improve security posture and compliance with industry standards • Communicate technical concepts to technical and non-technical stakeholders • Mentor and coach junior engineers • Stay current with site reliability engineering advancements and share knowledge • Identify organizational enhancements and propose alternatives to optimize team structures and execution

🎯 Requirements

• Bachelor’s degree in Computer Engineering, Computer Science, or related field • 5+ years of experience in a similar role, preferably in a high-traffic, high-availability environment • Proficiency in at least one programming language, such as Python, Ruby, Java, or Go • Strong understanding of cloud infrastructure and related technologies, including AWS, GCP, Azure, Kubernetes, and Docker • Excellent troubleshooting and problem-solving skills • Experience with automation and configuration management tools, such as Chef, Ansible, Puppet, or Terraform • Familiarity with monitoring and alerting tools, such as Prometheus, Grafana, or Nagios • Strong communication and interpersonal skills • Ability to navigate ambiguity, set clear expectations, and thrive in a fast-paced, dynamic environment • Strong grasp of computer science fundamentals involving distributed systems and networks • Experience or familiarity with running, tuning, and optimizing databases and queries • Prior experience in a leadership or senior-level site reliability engineering role • Familiarity with backend technologies and frameworks • Experience driving technical decisions and implementing organizational changes

🏖️ Benefits

• Inclusive and diverse workplace • Remote work environment • Competitive compensation • Potential share options for certain roles • Regular training • Annual learning stipend • High degree of autonomy • Mentorship • Ambitious goals supporting personal and company growth

Apply Now

Similar Jobs

🔥 7 hours ago

Endava

10,000+ employees

🏥 Healthcare

📣 Marketing

📦 Logistics

Senior DevOps Engineer building Terraform-based cloud infrastructure, CI/CD pipelines, and Kubernetes environments for Endava’s technology consulting clients. Supporting secure, automated, AI-enabled cloud operations.

Ansible

AWS

Azure

Chef

Cloud

Docker

Google Cloud Platform

Jenkins

Kubernetes

Puppet

Python

Terraform

🔥 19 hours ago

Akamai Technologies

5001 - 10000

🔒 Cybersecurity

Site Reliability Engineer II improving reliability, scalability, and performance across Akamai’s distributed cloud and edge content delivery platform. Automating operations, monitoring systems, and resolving complex infrastructure issues.

Cloud

Grafana

JavaScript

Linux

Oracle

Prometheus

Python

SQL

Unix

🕒 Yesterday

Akamai Technologies

5001 - 10000

🔒 Cybersecurity

Site Reliability Engineer II automating Akamai's global backbone and network operations. Building observability, cloud-native systems, and infrastructure-as-code solutions for Akamai's distributed platform.

Apache

Cassandra

Cloud

JavaScript

Kubernetes

Microservices

Postgres

Python

SDLC

🕒 Yesterday

Level AI

51 - 200

🤖 Artificial Intelligence

☁️ SaaS

🏢 Enterprise

Senior SRE optimizing Kubernetes costs, GPU throughput, and reliability for Level AI’s AI-powered customer-intelligence platform. Building backend enablement and security tooling across hybrid infrastructure.

Google Cloud Platform

Kubernetes

Node.js

Python

Rust

Terraform

Go

🕒 2 days ago

Miratech

501 - 1000

🤝 B2B

💼 Consulting

☁️ SaaS

DevOps Engineer designing highly available Kubernetes infrastructure for Miratech’s global IT services and consulting clients. Managing event-driven systems, GitOps, cloud platforms, databases, and observability.

AWS

Cloud

Consul

Distributed Systems

Flux

Grafana

Kafka

Kubernetes

Linux

MongoDB

MySQL

NGINX

Postgres

Prometheus

RabbitMQ

Redis

Terraform

Go