Senior DevOps Engineer

Job not on LinkedIn

🔥 0 minutes ago

🇮🇳 India – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Karat

Karat

201 - 500 employees

👥 HR Tech

🏢 Enterprise

☁️ SaaS

💰 Funding Round on 2022-04

HR Tech • Enterprise • SaaS

Karat is a company that provides an innovative, end-to-end solution for technical hiring. Its platform improves the quality, efficiency, and equity of the hiring process for organizations by offering 24/7 interviews and unlimited capacity with expert interviewers. Karat aims to delight candidates with a fair, friendly, and engaging experience while reclaiming engineering bandwidth through real interviews. The company also focuses on equitable and inclusive hiring, internal mobility, and global hiring solutions, ensuring a diverse and well-evaluated pool of candidates. With its human and technology approach, Karat integrates seamlessly into existing processes, offering a reliable and consistent interview experience.

📋 Description

• Evolve infrastructure, delivery systems, and operational practices enabling engineering teams to build and run reliable software • Establish scalable approaches to cloud infrastructure, CI/CD, observability, alerting, operational readiness, and maintenance • Partner with software engineers to improve developer experience and strengthen reliability, security, performance, and cost efficiency of Karat’s hosted SaaS platform • Own and evolve Karat’s AWS SaaS infrastructure • Design, improve, and operate CI/CD pipelines using CircleCI and related tooling • Build and maintain observability capabilities including metrics, logs, traces, dashboards, and actionable alerting using Datadog and related tools • Apply SRE principles to service reliability, availability, performance, capacity planning, incident response, root-cause analysis, and operational learning • Influence engineering-wide technical decisions and delivery practices through partnership and standards

🎯 Requirements

• 5+ years of experience in DevOps, Site Reliability Engineering, infrastructure engineering, platform engineering, or a closely related discipline • Significant hands-on production experience with AWS; required qualification • Experience designing, operating, and improving CI/CD systems using CircleCI, GitHub Actions, Jenkins, GitLab CI, or another major CI/CD platform • Strong experience with Docker and containerized application environments • Strong Linux, networking, security, and cloud-infrastructure fundamentals • Practical experience applying SRE principles to production systems, including observability, alerting, incident response, root-cause analysis, capacity planning, and reliability improvement • Hands-on experience with a leading telemetry and observability platform, such as Datadog, New Relic, Dynatrace, Grafana Cloud, or Splunk • Experience designing monitoring and alerting systems • Experience with cloud cost management and optimization • Experience partnering with application-engineering teams • Clear written and verbal English communication skills • Comfort working with globally distributed teams and regularly collaborating with colleagues in the United States • Must reside in Bengaluru (formerly known as Bangalore), India • Schedule must overlap with U.S. business hours • Application submissions must be 100% in English

🏖️ Benefits

• 100% remote work • Inclusive workplace and non-discrimination commitment • Accommodation support for disabilities or special needs

Apply Now

Similar Jobs

🕒 2 days ago

Atlan Stormwater

51 - 200

🏭 Manufacturing

📦 Logistics

💼 Consulting

Senior Software Engineer improving Atlan's multi-agent AI SRE platform for enterprise data and AI context. Building trusted investigations, safe auto-remediation, evaluation harnesses, and production guardrails.

🕒 3 days ago

Akamai Technologies

5001 - 10000

🔒 Cybersecurity

Senior Site Reliability Engineer automating Linux infrastructure, monitoring, and deployments. Supporting Akamai’s globally distributed cloud and edge platform for reliable, secure digital experiences.

🕒 5 days ago

Ford Motor Company

10,000+ employees

📦 Logistics

💼 Consulting

📣 Marketing

Senior DevOps Engineer automating secure, reliable GCP cloud platforms. Building Terraform infrastructure, Python tools, CI/CD pipelines, and GitOps solutions for enterprise teams.

🕒 October 2

JumpCloud

201 - 500

☁️ SaaS

🔐 Security

🏢 Enterprise

Senior SRE architecting multi-region cloud infrastructure, Kubernetes, disaster recovery, and observability for JumpCloud’s AI-powered IT management platform. Driving reliability, FinOps, automation, and incident management.

🕒 October 2

JumpCloud

201 - 500

☁️ SaaS

🔐 Security

🏢 Enterprise

Site Reliability Engineer strengthening AWS/GCP reliability, observability, Kubernetes, and disaster recovery. Supporting JumpCloud’s AI-powered unified IT management platform through automation and incident response.