Staff Engineer – Site Reliability Engineering

🕒 Yesterday

🇮🇳 India – Remote

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Nagarro

Nagarro

10,000+ employees

Founded 1996

💼 Consulting

📣 Marketing

🏥 Healthcare

Consulting • Marketing • Healthcare

Nagarro is a global leader in digital engineering and technology consulting. The company helps clients become innovative, digital-first businesses by leveraging technology to drive business breakthroughs. Known for its entrepreneurial agility and CARING mindset, Nagarro offers a wide range of services, including digital engineering, intelligent enterprise solutions, and experience and design services. With over 17,900 employees across 37 countries, Nagarro collaborates with industry leaders to accelerate digitalization and technology-led innovation.

📋 Description

• Design and develop a scalable SRE ecosystem following SRE and DevSecOps best practices • Develop reusable TypeScript scaffolding libraries for cloud-native components • Build and enhance solutions using AWS, EKS, Kubernetes and Infrastructure as Code • Drive automation, reliability, scalability and operational excellence across microservices • Define and implement SRE and DevOps best practices across applications • Collaborate with technology, product, operations and functional teams on SRE initiatives • Analyse business requirements and their impact across applications and cloud systems • Establish standardized and automated onboarding paths for applications onto the SRE platform • Evaluate emerging technologies and define strategies for cloud and SRE adoption • Implement CI/CD, observability, monitoring and service mesh capabilities • Identify and address reliability, scalability and operational challenges • Provide technical guidance and support to distributed engineering teams

🎯 Requirements

• Total experience 5.5+ years • Strong experience in TypeScript development, coding and design patterns • Hands-on experience with AWS cloud and cloud-native technologies • Experience with AWS CDK, Terraform and Infrastructure as Code (IaC) • Experience with Kubernetes, Docker and Amazon EKS • Experience in SRE, DevOps, scalability, reliability and cloud automation • Experience with CI/CD tools such as Jenkins and Git • Knowledge of observability and monitoring tools such as CloudWatch, Splunk and Dynatrace • Knowledge of service mesh technologies such as Istio • Experience developing reusable scaffolding libraries and cloud-native components • Ability to analyse application and infrastructure dependencies across microservices environments • Experience working with distributed teams across multiple time zones • Excellent communication, presentation and stakeholder collaboration skills • Bachelor’s or master’s degree in computer science, Information Technology, or a related field

🏖️ Benefits

• Employees can work remotely

Apply Now

Similar Jobs

🕒 2 days ago

PeopleCert

501 - 1000

📚 Education

☁️ SaaS

🤝 B2B

Commercial Head driving PeopleCert’s global DevOps certification growth. Expanding partnerships, revenue, and market adoption across international technology and education markets.

🕒 June 1

OpenAI

201 - 500

🤖 Artificial Intelligence

☁️ SaaS

🏢 Enterprise

Partner AI Deployment Engineer responsible for AWS deployment strategies and technical leadership in OpenAI. Guiding enterprise customers from ideation to production while influencing joint account strategy.

AWS

🕒 May 18

Akamai Technologies

5001 - 10000

🔒 Cybersecurity

🏢 Enterprise

📱 Media

Site Reliability Engineering Manager leading APJ-based site reliability engineers. Collaborating to define and improve Compute products operation and customer supportability.

Ansible

Chef

Distributed Systems

Puppet

React

SaltStack

🕒 May 13

AlphaSense

1001 - 5000

💼 Consulting

🏥 Healthcare

📣 Marketing

Staff Site Reliability Engineer at AlphaSense enhancing reliability, performance, and scalability of systems. Leading SRE practices and mentoring engineers in a global team.

AWS

Azure

Cloud

DNS

Google Cloud Platform

Grafana

Kubernetes

Prometheus

Python

TCP/IP

Go

🕒 May 13

AlphaSense

1001 - 5000

💼 Consulting

🏥 Healthcare

📣 Marketing

Staff Site Reliability Engineer shaping reliability and performance standards at AlphaSense, driving cultural adoption of SRE best practices across the engineering organization.

AWS

Azure

Cloud

DNS

Google Cloud Platform

Grafana

Kubernetes

Prometheus

Python

TCP/IP

Go