Site Reliability Engineer

🕒 July 8

🇮🇳 India – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 45%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Pythian

Pythian

201 - 500 employees

Founded 1997

💼 Consulting

🏥 Healthcare

📦 Logistics

Consulting • Healthcare • Logistics

<Pythian> Pythian is a global consulting firm specializing in data, analytics, cloud, and AI solutions. With decades of experience and a large team of domain experts, they provide strategy, custom AI development, advanced analytics, database and cloud consulting, migrations, and managed services (including AIOps and DBA services) to enterprise customers across industries. Pythian focuses on helping organizations modernize their data platforms, deploy generative AI use cases, and optimize cloud and database operations through partnerships with major cloud and data technology providers.

📋 Description

• Operate and optimize Kubernetes clusters, Istio service mesh, and Linux-based systems • Automate workflows using Go, Python, and Shell scripting • Build monitoring and observability solutions with Prometheus, Grafana, and Loki • Troubleshoot complex networking, storage, and system performance issues • Participate in on-call rotations and postmortem reviews to improve system resilience • Partner with AI/ML teams to ensure infrastructure readiness for model training and data pipelines

🎯 Requirements

• Experience with Google Cloud, plus IaC tools (Terraform) • Strong knowledge of microservices, containers (Kubernetes, Docker), and networking • SRE mindset with a focus on automation, scalability, and reliability • Hands-on experience with PKI, service mesh, and Linux systems administration

🏖️ Benefits

• Competitive total rewards package • Blog during work hours; take a day off and volunteer for your favorite charity • Flexibly work remotely from your home, there’s no daily travel requirement to an office! • All you need is a stable internet connection • Collaborate with some of the best and brightest in the industry! • Hone your skills or learn new ones with our substantial training allowance; participate in professional development days, attend training, become certified, whatever you like! • We give you all the equipment you need to work from home including a laptop with your choice of OS, and an annual budget to personalize your work environment! • Pythian cares about the health and well-being of our team. You will have an annual wellness budget to make yourself a priority (use it on gym memberships, massages, fitness and more) • Additionally, you will receive a generous amount of paid vacation and sick days, as well as a day off to volunteer for your favorite charity.

Apply Now

Similar Jobs

🕒 July 7

Resilinc

201 - 500

💼 Consulting

📦 Logistics

🏥 Healthcare

Site Reliability Engineer responsible for platform availability and automation in cloud environments at Resilinc. Focused on leveraging agentic AI for impactful supply chain solutions.

Azure

Cloud

Distributed Systems

DNS

Docker

Grafana

Hadoop

HDFS

Kafka

Kubernetes

Linux

Postgres

Redis

🕒 July 3

Moniepoint Inc. (Formerly TeamApt Inc.)

1001 - 5000

💳 Fintech

🏦 Banking

Site Reliability Engineer engineering reliability for Moniepoint’s distributed African financial platform. Building automation, observability, and self-healing systems for hyper-growth.

AWS

Azure

Cloud

Distributed Systems

Google Cloud Platform

Java

Kafka

Kubernetes

Microservices

MySQL

Postgres

Prometheus

Python

RabbitMQ

Rust

Go

🕒 July 1

Empower

10,000+ employees

💸 Finance

💳 Fintech

👥 B2C

Senior SRE architecting reliable AWS and Kubernetes infrastructure for Empower’s financial services platform. Leading incident response, automation, observability, security, and engineer mentorship.

AWS

Cloud

Consul

Flux

Kubernetes

Python

Splunk

Terraform

Go

🕒 June 23

SigNoz

11 - 50

☁️ SaaS

🏢 Enterprise

SRE responsible for the reliability and operability of SigNoz cloud platform while scaling observability systems and ingest pipelines. Work in a fast-paced, remote-first environment with a high-caliber team.

Cloud

Distributed Systems

Kubernetes

Open Source

Go

🕒 June 19

BETSOL

501 - 1000

💼 Consulting

🏥 Healthcare

📦 Logistics

Senior Cloud Engineer at BETSOL building and operating cloud portal workloads across Azure and GCP. Focused on DevOps and DevSecOps with AI-first development practices.

Ansible

Azure

Cloud

Google Cloud Platform

Grafana

JavaScript

Jenkins

Kubernetes

Prometheus

Python

Terraform

TypeScript

Vault