Site Reliability Engineer

🕒 July 3

🇮🇳 India – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 36%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Moniepoint Inc. (Formerly TeamApt Inc.)

Moniepoint Inc. (Formerly TeamApt Inc.)

1001 - 5000 employees

💳 Fintech

🏦 Banking

Fintech • Banking • Payments

Moniepoint Inc. is Africa's all-in-one financial ecosystem that provides seamless solutions in payments, banking, credit, and business management for over 10 million businesses and individuals. Operating as Nigeria's largest merchant acquirer, Moniepoint powers the majority of Point of Sale (POS) transactions in the country. The company processes $17 billion monthly while ensuring profitable operations. With operations starting in 2019, Moniepoint continues to support businesses through its comprehensive financial services platform, making significant strides in financial inclusion across emerging markets.

📋 Description

• Participate in on-call rotations as the primary technical lead • Act as Incident Commander during major severity incidents • Initiate war rooms, coordinate cross-functional teams, and provide status updates • Instrument code with high-cardinality metrics and distributed traces • Define, measure, and defend Service Level Objectives and Error Budgets with product owners • Write production-ready code for internal tooling, automation platforms, and self-healing mechanisms • Partner with Product Engineering teams to build reliability, scalability, and observability into new services • Apply circuit breakers, rate limiting, backpressure, and fallback strategies • Analyze system performance and traffic patterns to model future capacity needs • Conduct load testing and chaos engineering experiments

🎯 Requirements

• 3–6 years of experience in SRE or Backend Engineering • Strong ability to write clean, performant, and tested code in Java, Go, Rust, or Python • Deep understanding of distributed systems architecture and design patterns • Strong command of microservices fundamentals and event-driven architectures • Experience with Google Cloud Platform (GCP) or similar cloud providers such as AWS/Azure • Proficiency running production workloads on Kubernetes (GKE/EKS) • Ability to troubleshoot cluster and infrastructure issues • Experience designing observability strategies using OpenTelemetry, Prometheus, New Relic, Datadog, or SigNoz • Familiarity with operating and tuning PostgreSQL or MySQL • Familiarity with Kafka or RabbitMQ in a high-throughput environment

🏖️ Benefits

• Learning and development-focused environment • Knowledge sharing, training, and regular internal technical talks • Attractive salary • Pension • Health insurance • Annual bonus • Other benefits • Culture prioritizing employee well-being, respect, and inclusion

Apply Now

Similar Jobs

🕒 July 2

PlexTrac

51 - 200

🔒 Cybersecurity

☁️ SaaS

🤖 Artificial Intelligence

Senior DevSecOps Engineer leading security and reliability for PlexTrac's cloud-based platform. Involves architecture and implementation of secure, resilient systems with a focus on infrastructure.

Ansible

Cloud

Google Cloud Platform

Kubernetes

Linux

Python

SDLC

Terraform

Go

🕒 July 1

Synmatch AI

1 - 10

👥 HR Tech

🤖 Artificial Intelligence

🎯 Recruiter

Mobile DevOps & Release Engineer on the founding team for power-quality analysis software. Responsible for CI/CD, observability, and release lifecycle of mobile applications.

Android

AWS

Azure

Cloud

Dart

Flutter

iOS

Vault

🕒 June 24

Smart Working

51 - 200

💼 Consulting

🏥 Healthcare

📣 Marketing

DevSecOps Engineer securing AWS infrastructure for a technology platform. Maintaining compliant, reliable, and scalable cloud infrastructure with CI/CD automation.

Ansible

AWS

Docker

Grafana

Kubernetes

Prometheus

Terraform

🕒 June 23

SigNoz

11 - 50

☁️ SaaS

🏢 Enterprise

SRE responsible for the reliability and operability of SigNoz cloud platform while scaling observability systems and ingest pipelines. Work in a fast-paced, remote-first environment with a high-caliber team.

Cloud

Distributed Systems

Kubernetes

Open Source

Go

🕒 June 19

BETSOL

501 - 1000

💼 Consulting

🏥 Healthcare

📦 Logistics

Senior Cloud Engineer at BETSOL building and operating cloud portal workloads across Azure and GCP. Focused on DevOps and DevSecOps with AI-first development practices.

Ansible

Azure

Cloud

Google Cloud Platform

Grafana

JavaScript

Jenkins

Kubernetes

Prometheus

Python

Terraform

TypeScript

Vault