Team Leader, SRE

🔥 17 minutes ago

🌏 Anywhere in the World

💵 $75.5k - $169.7k / year

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Remote

Remote

501 - 1000 employees

💼 Consulting

📦 Logistics

🏥 Healthcare

Consulting • Logistics • Healthcare

Remote is a global HR platform that simplifies the process of hiring, onboarding, managing, and paying employees and contractors worldwide. It offers comprehensive solutions for recruitment, payroll management, contractor management, and compliance. The platform supports businesses in handling HR tasks seamlessly and efficiently, ensuring fast and compliant payouts, providing employer of record services, and facilitating employee benefits and equity offerings. Additionally, Remote integrates with various HR systems, allowing for a flexible, scalable, and reliable solution for businesses looking to expand globally.

📋 Description

• Lead the Site Reliability Engineering team as a 60% individual contributor and 40% leadership role • Own direct reports' onboarding, feedback, performance assessment, progression and hiring • Steer the team's focus against company goals and represent the team across engineering and senior leadership • Set technical direction and review the team's work • Own SRE goals, prioritization, support rotation and on-call model • Oversee Kubernetes, AWS, PostgreSQL, DNS and TLS, and CI infrastructure • Develop the reliability practice, including SLOs, error budgets, incident response and observability • Partner with Security on threats, patching, infrastructure controls, audit and compliance obligations • Manage platform vendor relationships, renewals and commercial conversations • Improve operational load versus project delivery balance and extend the SLO framework across teams

🎯 Requirements

• Experience leading an SRE, infrastructure or platform engineering team • Ownership of reports' growth, performance and career progression • Experience coaching technical craft and soft skills • Experience handling underperformance directly and early • Experience hiring engineers and assessing engineering quality • Hands-on background in site reliability, DevOps or cloud infrastructure engineering • Kubernetes in production • AWS at meaningful scale • Hands-on AI building, enablement, and scaling AI infrastructure • Solid observability practices and principles • Infrastructure as code with Terraform • CI/CD systems such as GitLab CI, GitHub Actions or Jenkins • Docker and shell scripting • Experience running a reliability practice: incident response, on-call, SLOs and error budgets • Understanding and history of working in regulated environments • Ability to prioritize operational load and project work • Clear written communication for asynchronous work • Ability to build relationships across teams • English application materials required

🏖️ Benefits

• work from anywhere • flexible paid time off • flexible working hours (we are async) • 16 weeks paid parental leave • budget towards co-working spaces, learning and wellness (including gym memberships) • mental health support services • stock options • home office budget & IT equipment • life-work balance and schedule flexibility • employee resource groups (Women, Disability, Queer, Minorities in Tech) • accommodation support during interviews and beyond

Apply Now

Similar Jobs

🕒 September 15

Stellartech Research Corporation

51 - 200

🏥 Healthcare

💼 Consulting

⚕️ Healthcare Insurance

DevOps Engineer managing AWS, Kubernetes, CI/CD, and observability for StellarTech’s scalable global EdTech products. Improving reliability, security, and production operations across multiple products.

🌏 Anywhere in the World

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 August 21

Social Discovery Group

1001 - 5000

🌍 Social Impact

📱 Media

Site Reliability Engineer improving reliable, scalable infrastructure for Social Discovery Group’s social discovery platforms. Automating deployments, Kubernetes operations, observability, and CI/CD worldwide.

🌏 Anywhere in the World

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🗣️🇷🇺 Russian Required

🕒 August 10

Yuno

11 - 50

💳 Fintech

🏢 Enterprise

☁️ SaaS

Staff SRE defining reliability for Yuno’s AI-powered payments infrastructure. Owning AWS architecture, messaging, observability, incident response, and resilience at scale.

🌏 Anywhere in the World

⏰ Full Time

🟠 Senior

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 27

Empowers Staffing Inc

11 - 50

💼 Consulting

🎯 Recruiter

🤖 Artificial Intelligence

Infrastructure Automation Engineer at LAK Technology Inc managing cloud infrastructure with Terraform and Ansible. Focused on automation and CI/CD processes across AWS, Azure, or GCP environments.

🌏 Anywhere in the World

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 27

Supabase

51 - 200

☁️ SaaS

🔌 API

🤖 Artificial Intelligence

Release Engineer at Supabase, ensuring safe and observable deployments and operational reliability across systems. Engage in incident management, monitoring, and process documentation for improved deployment efficiency.

🌏 Anywhere in the World

💰 $80M Series B on 2022-05

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)