Staff Site Reliability Engineer

Job not on LinkedIn

🔥 1 minute ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Datavant

Datavant

201 - 500 employees

Founded 2017

🏥 Healthcare

💼 Consulting

⚕️ Healthcare Insurance

💰 $40M Series B on 2020-10

Healthcare • Consulting • Healthcare Insurance

Datavant is a company that provides a platform and network focused on making health data secure, accessible, and usable across the healthcare ecosystem. With a focus on data connectivity and interoperability, Datavant facilitates the movement of healthcare records across a vast network of organizations, including hospitals, clinics, health systems, and data partners. Their suite of products and solutions covers areas such as health data exchange, data transformation, and privacy compliance, serving various clients including health plans, healthcare providers, life sciences, and government organizations. Datavant's mission is to advance human health through improved data exchange and analytics.

📋 Description

• Own the technical direction of SRE functions for designated teams from architecture definition and test planning through implementation and ongoing operations • Lead design and development of tooling and processes that improve service delivery at scale: reliability, resiliency, efficiency, visibility, and quality • Drive adoption of researched, verified, and standardized engineering practices across the teams you support • Enhance system designs and infrastructure implementations to make them more efficient, standardized, and maintainable • Take end-to-end ownership of major projects defining architecture, test plans, and implementation and empower others to do the same • Push for and advocate superior architectural solutions; provide a compelling, data-backed case for change • Drive transformation and standardization of foundational infrastructure services as the environment evolves and the business grows • Architect and lead the integration of newly acquired cloud environments into Datavant's infrastructure, ensuring security, reliability, and consistency across an expanding multi-cloud footprint • Design and implement hybrid-cloud and cross-cloud connectivity strategies to ensure interoperability as M&A activity adds new environments • Collaborate with security, platform engineering, and development teams to ensure newly integrated environments adhere to Zero Trust principles, governance frameworks, and automation-first practices.

🎯 Requirements

• 8+ years of experience in site reliability engineering, platform engineering, or infrastructure architecture • Strong expertise in cloud infrastructure, including networking, security, compute, storage, and IAM • Experience migrating workloads and integrating new cloud environments into existing architectures including M&A-driven consolidation scenarios • Deep understanding of multi-cloud networking (AWS VPCs, Azure VNets, Transit Gateway, ExpressRoute, PrivateLink, DNS, hybrid connectivity) • Hands-on experience with Infrastructure as Code and automation (Terraform and Ansible) • Strong security knowledge, including IAM, encryption, network security, PKI, certificate management, and compliance frameworks (SOC2, HITRUST, NIST, FedRAMP) • Experience managing multi-account cloud governance and security policies (AWS Organizations, Azure Policy, SCPs) • Proven track record of technical leadership owning large, complex projects end-to-end and driving them to completion • Demonstrated ability to mentor engineers, lead design reviews, and elevate the technical quality of a team • Excellent problem-solving skills, with the drive to bring clarity to vague situations and a desire to constantly expand your skills • Strong communication skills, including the ability to describe complex problems to non-technical staff, paired with the curiosity, adaptability, and flexibility needed to navigate a growing enterprise • Experience leveraging AI agents to accelerate daily workload.

🏖️ Benefits

• Health insurance • 401(k) plans • Flexible work arrangements • Professional development opportunities

Apply Now

Similar Jobs

🔥 22 minutes ago

Skylo

51 - 200

📡 Telecommunications

🤝 B2B

Staff Network Reliability Engineer managing RAN operations for satellite connectivity at Skylo. Overseeing RAN health and incident resolution in a production NTN environment.

🇺🇸 United States – Remote

💵 $150k - $162k / year

💰 $30M Venture Round - Skylo on 2025-02

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

Grafana

IoT

Kubernetes

Node.js

Prometheus

TypeScript

🔥 2 hours ago

CDW

10,000+ employees

💼 Consulting

🏥 Healthcare

📚 Education

Principal Consulting Engineer optimizing private cloud infrastructure built on VCF 9.0 for CDW. Driving standardization and automation across compute, storage, and networking management layers.

Cloud

🔥 4 hours ago

RTX

10,000+ employees

🚀 Aerospace

🎖️ Defense

🏭 Manufacturing

Site Reliability Engineer automating and improving reliability at Collins Aerospace. Collaborating with teams to tackle complex technical problems in flight operations.

🇺🇸 United States – Remote

💵 $107.5k - $204.5k / year

💰 $200k Grant - RTX on 2024-11

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

Ansible

Docker

Kubernetes

Linux

SaltStack

Terraform

Unix

🔥 12 hours ago

Ad Hoc LLC

501 - 1000

💼 Consulting

🏥 Healthcare

📦 Logistics

Staff DevOps Engineer at Ad Hoc shaping the long-term technical strategy and mentoring team members. Leading critical projects and ensuring compliance in delivering software for Veterans Affairs.

AWS

Cloud

HAProxy

Kubernetes

Terraform

🕒 Yesterday

Cisco

10,000+ employees

🔧 Hardware

🔐 Security

Customer Reliability Engineer for Cisco managing escalations for Hypershield on Nexus switches. Responsibilities include diagnosing complex issues and improving system reliability.

Linux