Lead Site Reliability Developer

🕒 May 1

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Ticketmaster

Ticketmaster

10,000+ employees

Founded 1976

🛍️ eCommerce

⚽ Sports

eCommerce • Entertainment • Sports

Ticketmaster is a leading ticketing platform that facilitates the sale of tickets for concerts, sports events, theater performances, and other live entertainment. The platform offers a user-friendly experience for purchasing tickets, as well as managing events and finding popular shows and games. Ticketmaster serves as the official ticket marketplace for many major sports leagues and artist events, making it a key player in the live entertainment industry.

📋 Description

• Lead consulting work from discovery through delivery • Establish working cadence and facilitate decision forums • Align stakeholders on reliability targets and trade-offs • Identify systemic risks and coordinate remediation • Drive change adoption by embedding reliability mechanisms • Design and implement reusable reliability mechanisms • Lead complex incident investigations and ensure learnings translate into durable fixes

🎯 Requirements

• Deep practical understanding of SRE principles • Proven ability to lead cross-team technical work • Strong experience designing and troubleshooting distributed systems • Strong Kubernetes and AWS experience • Ability to design reliability automation and tooling • Excellent communication skills

🏖️ Benefits

• Inclusive work environment • Professional development opportunities • Opportunities to work with talented people

Apply Now

Similar Jobs

🕒 May 1

Live Nation Entertainment

10,000+ employees

📱 Media

Lead Site Reliability Engineer leading consulting work at Ticketmaster for reliability improvements across multiple teams. Aligning stakeholders and driving adoption of SRE principles.

AWS

Distributed Systems

Kubernetes

🕒 April 24

GitLab

1001 - 5000

🤖 Artificial Intelligence

🏢 Enterprise

☁️ SaaS

Cloud Cost Utilization SRE responsible for making cloud spending actionable. Collaborating with Finance and Engineering at GitLab to optimize resource usage.

Ansible

AWS

Cloud

Google Cloud Platform

Grafana

Prometheus

Terraform

🕒 April 22

NICE

5001 - 10000

☁️ SaaS

🤖 Artificial Intelligence

📡 Telecommunications

SRE - NOC role focuses on service reliability, incident response, and operational automation. Precision in dealing with operational toil through engineering practices for global operations at NICE.

Ansible

AWS

Cloud

DNS

Docker

Grafana

Kubernetes

Linux

Prometheus

Python

Splunk

TCP/IP

Terraform

Go

🕒 April 21

Ripjar

51 - 200

💸 Finance

📋 Compliance

🤖 Artificial Intelligence

DevOps Engineer ensuring reliability and security of infrastructure for software combating financial crime at Ripjar. Focus on continuous improvement and automation within a remote-first team.

Ansible

AWS

Azure

Cloud

Docker

JavaScript

Kubernetes

Linux

Prometheus

Python

Terraform

🕒 April 17

Recruiting.com

11 - 50

🎯 Recruiter

☁️ SaaS

🤝 B2B

Lead DevOps Engineer overseeing Azure infrastructure and CI/CD pipelines improvements at Cencora. Mentor engineers and align initiatives with business goals in the pharmaceutical consulting sector.

Azure

Cloud

Kubernetes

Python

Terraform

Go