Senior Site Reliability Engineer

đź•’ August 13

🇺🇸 United States – Remote

đź’µ $168k - $200k / year

⏰ Full Time

đźź  Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

infoinfo

đź‘» Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Datavant

Datavant

201 - 500 employees

Founded 2017

🏥 Healthcare

đź’Ľ Consulting

⚕️ Healthcare Insurance

đź’° $40M Series B on 2020-10

Healthcare • Consulting • Healthcare Insurance

Datavant is a company that provides a platform and network focused on making health data secure, accessible, and usable across the healthcare ecosystem. With a focus on data connectivity and interoperability, Datavant facilitates the movement of healthcare records across a vast network of organizations, including hospitals, clinics, health systems, and data partners. Their suite of products and solutions covers areas such as health data exchange, data transformation, and privacy compliance, serving various clients including health plans, healthcare providers, life sciences, and government organizations. Datavant's mission is to advance human health through improved data exchange and analytics.

đź“‹ Description

• Own the Databricks and Snowflake platform lifecycle, including automation, workspace governance, job orchestration, and cost optimization • Architect resilient, scalable, and secure infrastructure across cloud environments • Drive failover, autoscaling, chaos testing, and capacity planning initiatives • Build and maintain monitoring, alerting, and logging infrastructure using Datadog and other open tooling • Define and enforce SLOs and SLAs for critical services • Automate deployments of data pipelines, ML workflows, and infrastructure components using GitHub Actions, Terraform, and related IaC tooling • Build patterns and tooling for inter- and intra-cloud data movement across Snowflake, S3, Delta Lake, and Kafka • Use EventBridge, SNS/SQS, and Lambda to build loosely coupled, scalable data systems • Partner with analytics, data science, product, and engineering teams • Influence decisions on data platform architecture, ML enablement, and data product strategy

🎯 Requirements

• 6+ years of experience in SRE, platform engineering, or DevOps supporting data-intensive or ML-powered applications • Daily use of Claude Code, Cursor, Copilot, or equivalent AI coding tools • Hands-on Databricks experience, including workspace setup, cluster/job management, and CI/CD and data orchestration integration • Experience with Snowflake • Deep understanding of AWS or similar cloud-native infrastructure, including VPCs, IAM, event-driven patterns, and serverless compute • Expertise with observability tools, especially Datadog • Strong command of CI/CD tooling, especially GitHub Actions • Experience with infrastructure-as-code, especially Terraform • Working knowledge of shell scripting and Python • Experience building and supporting highly available, fault-tolerant systems • Excellent communication and collaboration skills • No employment sponsorship available

🏖️ Benefits

• Total rewards strategy • Post-offer health screenings and vaccinations as required by clients • Reasonable accommodations for individuals with physical and mental disabilities • Equal Employment Opportunity protections

Apply Now

Similar Jobs

đź•’ August 13

Sumsub

501 - 1000

đź’Ľ Consulting

⚖️ Legal

📦 Logistics

DevSecOps Engineer building automated vulnerability and fleet security capabilities for Sumsub’s AI-powered trust infrastructure. Creating self-service APIs, dashboards, and compliance controls for global engineering teams.

🇺🇸 United States – Remote

đź’° $30M Series B - Sumsub on 2022-12

⏰ Full Time

🟡 Mid-level

đźź  Senior

⛑ DevOps & Site Reliability Engineer (SRE)

đź•’ August 13

Union Home Mortgage Corp.

1001 - 5000

🏦 Banking

🏠 Real Estate

Infrastructure DevOps Engineer modernizing Union Home Mortgage’s AWS and Azure cloud platforms. Building automation, CI/CD pipelines, infrastructure as code, and operational processes for enterprise systems.

🇺🇸 United States – Remote

⏰ Full Time

🟡 Mid-level

đźź  Senior

⛑ DevOps & Site Reliability Engineer (SRE)

đź•’ August 13

Bitdeer Group

201 - 500

đź’Ľ Consulting

📦 Logistics

🏗️ Construction

SRE Expert automating EVPN/BGP infrastructure for Bitdeer’s AI cloud and Bitcoin mining operations. Owning multi-site data center fabrics, WAN, internet edge, and intent-driven network workflows.

🇺🇸 United States – Remote

đź’° Post-IPO Equity on 2023-05

⏰ Full Time

🟡 Mid-level

đźź  Senior

⛑ DevOps & Site Reliability Engineer (SRE)

đź•’ August 12

Trimble Inc.

10,000+ employees

📦 Logistics

🏭 Manufacturing

đź’Ľ Consulting

Enterprise AI DevOps SRE scaling cloud automation across AWS, Azure, and GCP for Trimble’s connected-industry technology. Building AI workflows, observability, and cost-optimized infrastructure across enterprise platforms.

🇺🇸 United States – Remote

đź’µ $115.6k - $158.9k / year

đź’° Post-IPO Debt on 2022-12

⏰ Full Time

🟡 Mid-level

đźź  Senior

⛑ DevOps & Site Reliability Engineer (SRE)

đź•’ August 12

Mirantis

501 - 1000

đź’Ľ Consulting

🏥 Healthcare

📦 Logistics

Senior DevOps Engineer operating high-performance Kubernetes storage for Mirantis, a Kubernetes-native AI infrastructure company. Automating NFS, Linux, and air-gapped storage platforms for GPU workloads at scale.