Senior Site Reliability Engineer

Job not on LinkedIn

🔥 2 hours ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Counterpart Health

Counterpart Health

51 - 200 employees

Founded 2024

🏥 Healthcare

🤖 Artificial Intelligence

☁️ SaaS

Healthcare • Artificial Intelligence • SaaS

Counterpart Health is an AI-powered physician enablement platform that delivers data-driven, point-of-care insights to support value-based care. Incubated at Clover Health as Clover Assistant, its proprietary machine-learning models ingest data from 100+ sources, integrate with EHRs, and surface prioritized clinical actions to improve post‑hospitalization follow-up, quality (HEDIS) performance, and cost/risk management for payors, ACOs, and primary care practices. The company offers web-based SaaS integrations, flexible partnership models for value‑based care transformation, and patented ML technologies for diagnosis and medication management.

📋 Description

• Build systems for declarative application and infrastructure lifecycle management, including continuous deployment, continuous integration, Kubernetes cluster management, and service/workload inventory. • Prioritize and troubleshoot infrastructure issues, minimizing downtime and responding to alerts efficiently. • Contribute to setting the direction of the Site Reliability Engineering (SRE) team, ensuring goals align with Counterpart Health’s company-wide objectives. • Foster a collaborative, high-performance culture that promotes motivation, innovation, and cross-disciplinary teamwork. • Streamline and automate infrastructure processes, including delivery pipelines and database changes.

🎯 Requirements

• You have 5+ years of programming experience and are proficient in at least one of the following languages: Python, Go, or Shell Scripting. • You have in-depth knowledge of containerization technologies and orchestration, such as Docker, Containerd, and Kubernetes, along with experience with CNCF-based technologies like Helm, gRPC, and Prometheus. • You have experience with public cloud platforms such as GCP, Azure, or AWS. • You are knowledgeable in networking fundamentals, including TCP/IP, UDP, firewalls, routing, DNS, and load balancing. • You have experience with Linux system administration and a solid understanding of Linux design principles. • You understand key SRE concepts, such as monitoring, performance tuning, and automation. • You can work autonomously with limited guidance, proactively identifying and solving problems. • You have excellent communication and collaboration skills, with the ability to work effectively with cross-functional teams and adapt to new challenges and evolving technologies.

🏖️ Benefits

• Financial Well-Being: Our commitment to attracting and retaining top talent begins with a competitive base salary and equity opportunities. Additionally, we offer a performance-based bonus program, 401k matching, and regular compensation reviews to recognize and reward exceptional contributions. • Physical Well-Being: We prioritize the health and well-being of our employees and their families by providing comprehensive medical, dental, and vision coverage. Your health matters to us, and we invest in ensuring you have access to quality healthcare. • Mental Well-Being: We understand the importance of mental health in fostering productivity and maintaining work-life balance. To support this, we offer initiatives such as No-Meeting Fridays, monthly company holidays, access to mental health resources, and a generous flexible time-off policy. Additionally, we embrace a remote-first culture that supports collaboration and flexibility, allowing our team members to thrive from any location. • Professional Development: Developing internal talent is a priority for Clover. We offer learning programs, mentorship, professional development funding, and regular performance feedback and reviews. • Additional Perks: Employee Stock Purchase Plan (ESPP) offering discounted equity opportunities • Reimbursement for office setup expenses • Monthly cell phone & internet stipend • Remote-first culture, enabling collaboration with global teams • Paid parental leave for all new parents • And much more!

Apply Now

Similar Jobs

🔥 3 hours ago

Cognativ

11 - 50

💼 Consulting

🥽 AR/VR

🤖 Artificial Intelligence

Senior Site Reliability Engineer ensuring reliability of a distributed AI video monitoring platform. Leading incident response and managing service reliability and operational quality.

Apache

AWS

Cloud

Grafana

IoT

Java

Kafka

Linux

Postgres

Prometheus

Python

Terraform

Go

🔥 4 hours ago

National Trust

10,000+ employees

🤲 Charity

🏨 Hospitality

🛒 Retail

DevSecOps Engineer III at National Digital Trust Company securing digital asset infrastructure. Lead modernization in CI/CD, security controls, and cloud-native systems.

Cloud

ITSM

Kubernetes

SDLC

Terraform

🔥 4 hours ago

PayNearMe

201 - 500

💳 Fintech

☁️ SaaS

🤝 B2B

Site Reliability Engineer at PayNearMe, Inc. responsible for infrastructure management and application reliability. Collaborating with cross-functional teams to enhance system performance and incident response.

🇺🇸 United States – Remote

💵 $180k - $200k / year

🔥 Funding within the last year

💰 $50M Series E - PayNearMe on 2025-09

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

Ansible

AWS

Azure

Chef

Cloud

Docker

EC2

Google Cloud Platform

Grafana

Kubernetes

Prometheus

Puppet

Python

Ruby

Ruby on Rails

Splunk

Terraform

Go

🔥 4 hours ago

Made4net

51 - 200

📦 Logistics

☁️ SaaS

🏢 Enterprise

Cloud Operations Engineer supporting AWS infrastructure for supply chain software solutions. Monitoring systems and ensuring reliability in a global operations team.

Ansible

AWS

Cloud

DNS

EC2

Grafana

ITSM

Linux

Microservices

Oracle

Postgres

Python

Terraform

🔥 5 hours ago

Mercadona

10,000+ employees

🛒 Retail

🍽️ Food & Beverage

DevOps Prime overseeing GCH’s cloud infrastructure while directing vendor resources and ensuring security. Responsible for CI/CD strategies, observability, and incident response across platforms.

AWS

Cloud

Terraform