Senior Manager, Site Reliability Engineering

🕒 June 23

🇺🇸 United States – Remote

💵 $187k - $243k / year

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

info
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Clover Health

Clover Health

501 - 1000 employees

🏥 Healthcare

🛡️ Insurance

🤖 Artificial Intelligence

Healthcare • Insurance • Artificial Intelligence

Clover Health is a healthcare technology company helping members live their healthiest lives with our Medicare Advantage plans. Focused on seniors who have historically lacked access to affordable and high-quality healthcare, Clover aims to provide care sustainably by improving medical outcomes while simultaneously lowering avoidable costs. They leverage a proprietary software platform, the Clover Assistant, to aggregate patient data and offer real-time recommendations to healthcare providers. Their services include affordable Medicare Advantage plans, a home care program, and they manage care for Medicare Advantage members across several U. S. states.

📋 Description

• Lead and grow our SRE team of ~10 engineers, including hiring, retention, career development, and performance management across multiple time zones (US, HK, NZ). • Build strategic partnerships with product engineering pillars — shifting SRE from reactive, ticket-based support to proactive co-ownership of reliability outcomes. • Scale our multi-tenant infrastructure to support new customer onboarding and growing patient populations. • Own cloud cost management and FinOps practices, building frameworks that balance cost control with reliability and performance. • Champion developer self-service and platform engineering. Build self-service capabilities so product teams can manage routine operations without filing SRE tickets. Establish SLOs/SLIs for critical services and improve alert quality so every page is meaningful. • Ensure the SRE team is fully leveraging AI tooling in their workflows — using tools like Claude Code for IaC generation, log analysis, root cause investigation, and automating repetitive work — at the same level as the rest of engineering.

🎯 Requirements

• You have 6+ years managing an SRE team and 10+ years of hands-on SRE or infrastructure engineering experience. • You're deeply comfortable with our core stack: Kubernetes, GCP (GKE, Cloud SQL, Pub/Sub, GCS), Terraform, Helm, ArgoCD, PostgreSQL, and Prometheus/Grafana. • You have strong programming skills in Python and/or Go, and you're comfortable writing and reviewing infrastructure tooling code — including using AI coding tools to do so. • You have experience with CI/CD pipelines (GitHub Actions) and a track record of building or improving developer tooling and automation. • You have sound build vs. buy judgment — you default to the right answer, not the easiest one, and you're comfortable building internal tooling when existing solutions don't fit. • You have experience leading teams across multiple time zones and a track record of developing engineers into strong technical contributors.

🏖️ Benefits

• Financial Well-Being: Our commitment to attracting and retaining top talent begins with a competitive base salary and equity opportunities. Additionally, we offer a performance-based bonus program, 401k matching, and regular compensation reviews to recognize and reward exceptional contributions. • Physical Well-Being: We prioritize the health and well-being of our employees and their families by providing comprehensive medical, dental, and vision coverage. Your health matters to us, and we invest in ensuring you have access to quality healthcare. • Mental Well-Being: We understand the importance of mental health in fostering productivity and maintaining work-life balance. To support this, we offer initiatives such as No-Meeting Fridays, monthly company holidays, access to mental health resources, and a generous flexible time-off policy. Additionally, we embrace a remote-first culture that supports collaboration and flexibility, allowing our team members to thrive from any location. • Professional Development: Developing internal talent is a priority for Clover. We offer learning programs, mentorship, professional development funding, and regular performance feedback and reviews. • Additional Perks: Employee Stock Purchase Plan (ESPP) offering discounted equity opportunities • Reimbursement for office setup expenses • Monthly cell phone & internet stipend • Remote-first culture, enabling collaboration with global teams • Paid parental leave for all new parents • And much more!

Apply Now

Similar Jobs

🕒 June 23

General Dynamics Information Technology

10,000+ employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Lead DevSecOps Systems Engineer at GDIT focusing on secure platform experiences for development teams and evolving cloud footprint with advanced technologies.

🕒 June 22

Stack AV

51 - 200

📦 Logistics

🏭 Manufacturing

💼 Consulting

Stack AV Site Reliability Engineer managing large-scale autonomous systems development and infrastructure performance. Collaborating across teams to enhance reliability, scalability, and automation of compute platforms.

🇺🇸 United States – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 June 22

nDeavour Consulting

1 - 10

💼 Consulting

📦 Logistics

📣 Marketing

Site Reliability Engineer ensuring health, performance, and delivery of infrastructure systems at Mobile Wave Solutions. Working collaboratively with engineers to automate processes and improve operational reliability.

🇺🇸 United States – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 June 21

Tkxel

501 - 1000

💼 Consulting

📣 Marketing

🏥 Healthcare

Senior Azure DevOps Engineer at Tkxel designing, implementing, and maintaining Azure DevOps infrastructure while mentoring junior team members in a dynamic environment.

🕒 June 20

Gorilla Logic

501 - 1000

💼 Consulting

📣 Marketing

📦 Logistics

Technical Engineering Manager leading high-performing cloud and DevOps teams. Guiding architecture and delivery of scalable, reliable, and secure cloud solutions for clients.

🇺🇸 United States – Remote

⏰ Full Time

🟠 Senior

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)