Senior Site Reliability Engineer

🔥 12 hours ago

🇺🇸 United States – Remote

💵 $191k - $226k / year

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 15%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Garner Health

Garner Health

51 - 200 employees

💼 Consulting

📦 Logistics

🏥 Healthcare

Consulting • Logistics • Healthcare

Garner Health is focused on improving the way employees find high-quality doctors. With a belief in transparency and data-driven decision-making, they create solutions that facilitate better medical choices for employees. The team comprises healthcare operators, clinicians, engineers, and benefits experts, allowing for a multidisciplinary approach in developing these healthcare solutions.

📋 Description

• Own the end-to-end reliability, performance, and resilience of Garner’s AWS and Kubernetes cloud environments, including AI/ML workloads • Define, measure, and uphold SLOs across critical services • Serve in the on-call rotation and lead incident response • Drive root cause analysis, corrective actions, and rigorous infrastructure-change reviews • Build and maintain monitoring, alerting, and observability systems • Translate scaling requirements into automated, composable Terraform infrastructure-as-code deliverables • Implement cloud cost-efficiency and performance improvements across the stack • Reduce operational toil and technical debt using AI tools and automation • Build and maintain deployment and observability standards for engineering teams • Communicate cloud and reliability concepts to technical and non-technical stakeholders • Ensure infrastructure and operations meet security and HIPAA compliance obligations

🎯 Requirements

• 4+ years of hands-on experience operating production cloud infrastructure at scale in an SRE, DevOps, or platform engineering role • Deep expertise with Kubernetes and Terraform in a cloud-first environment • AWS preferred • Strong production observability experience, including defining SLOs, building monitoring and alerting, leading incident response, and conducting blameless post-incident reviews • Strong software engineering fundamentals in Python or Go, applied to infrastructure automation • Experience driving cloud cost-efficiency and performance optimization across compute, storage, and networking • Fluency with AI tools such as Claude applied to engineering and operations workflows, or strong motivation to build it quickly • Must be unable to require employer sponsorship or transfer of an employment visa • Experience supporting AI/ML or data-intensive workloads in production is a plus • Experience operating in a security-conscious or regulated environment such as HIPAA or SOC 2 is a plus • Experience with Kubernetes APIs is a plus

🏖️ Benefits

• Equity incentive plan • Flexible PTO • Medical plan options • Dental plan options • Vision plan options • 401(k) with company match • Flexible spending accounts • Teladoc Health

Apply Now

Similar Jobs

🔥 12 hours ago

Skimmer

11 - 50

☁️ SaaS

🤝 B2B

⚡ Productivity

Senior DevOps Engineer building Azure infrastructure, CI/CD, and observability for Skimmer’s pool-service platform. Owning deployment reliability for a new product line.

🔥 13 hours ago

Empower AI

501 - 1000

🎖️ Defense

🏥 Healthcare

📦 Logistics

Senior DevOps Engineer automating AWS infrastructure and supporting USCIS systems for Empower AI, an AI platform provider for government. Building secure, scalable solutions across 14 development teams.

🔥 13 hours ago

ICU Medical

5001 - 10000

🏥 Healthcare

🍽️ Food & Beverage

🏭 Manufacturing

Senior SRE optimizing AWS infrastructure for ICU Medical, a healthcare technology and IV therapy products company. Supporting HIPAA-compliant production systems, incident response, and Kubernetes modernization.

🔥 15 hours ago

DriveTime

1001 - 5000

🚘 Automotive

📦 Logistics

📣 Marketing

ITSM engineer building Freshservice ITOM, Asset Management, CMDB, Change, and Release practices. Supporting DriveTime’s used-car sales, financing, and servicing operations.

🔥 22 hours ago

Sprezzatura

51 - 200

🏛️ Government

💼 Consulting

🏥 Healthcare

DevOps Engineer building secure AWS and Kubernetes infrastructure for VA.gov. Automating Terraform-based deployments, CI/CD pipelines, observability, and developer tooling for government digital services.