Director of SRE

Vaga não está no LinkedIn

🕒 6 dias atrás

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $175.000 - $200.000 / ano

⏰ Tempo Integral

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Intus Care

Intus Care

11 - 50 funcionários

💼 Consultoria

📣 Marketing

📦 Logística

💰 $13.100.000 Venture Round em 2023-01

Consulting • Marketing • Logistics

A IntusCare é uma empresa de tecnologia que se especializa em soluções de gestão de cuidados para programas PACE (Programa de Cuidados Abrangentes para Idosos) e Medicare. A empresa oferece diversas ferramentas e serviços, incluindo serviços de cuidados integrados, ajuste de risco e análises de saúde populacional, destinados a aprimorar a entrega de cuidados baseados em valor. Com mais de 50 parcerias em 18 estados, a IntusCare oferece tecnologia inovadora para simplificar o cuidado complexo e melhorar os resultados de saúde ao abordar o risco clínico e facilitar o planejamento de cuidados proativos.

Descrição

• Own and execute the SRE strategy and multi-quarter roadmap across reliability, observability, incident management, QA maturity, and release engineering • Define, measure, and improve SLAs, SLOs, error budgets, uptime, performance, and operational health metrics • Lead production reliability, including monitoring, alerting, on-call operations, incident response, root cause analysis, and MTTR reduction • Establish release readiness standards, deployment safety controls, and quality gates • Manage external SRE vendors and partners, including service delivery, SLA governance, escalations, performance reviews, and compliance expectations • Lead QA engineering strategy focused on automation, regression prevention, test coverage, and reducing escaped production defects • Partner with Security and Engineering leaders on cloud infrastructure, CI/CD pipelines, operational tooling, HIPAA, SOC2, and internal security standards • Oversee Azure AKS, Kubernetes, GitOps workflows, CI/CD pipelines, GitHub Actions, secrets management, access controls, and audit readiness • Drive observability maturity using Grafana, Prometheus, logging platforms, tracing tools, and automated alerting frameworks • Collaborate with Product, Platform, and Engineering teams to embed reliability and quality practices throughout the software development lifecycle • Build, mentor, and scale SRE and QA teams • Drive AI-enabled automation and intelligent tooling to reduce manual toil and improve operational excellence

🎯 Requisitos

• 12+ years of SRE, infrastructure, or platform engineering experience • 5+ years of engineering leadership experience • Proven ownership of site reliability for complex, multi-tenant SaaS platforms with demanding availability requirements • Experience defining SLA and SLO frameworks, error budgets, and incident management processes at scale • Experience managing managed infrastructure or SRE service vendors, including SLA governance and performance management • Experience leading QA or quality engineering functions, test automation maturity, and release gate ownership • Strong hands-on experience with Microsoft Azure, preferably including AKS, networking, storage, IAM, and security services • Deep expertise in Kubernetes, containerized workloads, and production-scale distributed systems • Experience with CI/CD pipelines using GitHub Actions, ArgoCD, Terraform, or similar DevOps tooling • Strong background in monitoring, logging, tracing, and observability platforms such as Grafana, Prometheus, Datadog, or Splunk • Experience with scripting and automation using Python, Bash, PowerShell, or similar languages • Strong understanding of release engineering, automated testing frameworks, QA tooling, and shift-left quality practices • Experience supporting SaaS applications with uptime, scalability, and security requirements in regulated industries • Knowledge of HIPAA, SOC2, vulnerability management, access controls, and infrastructure security best practices • Familiarity with databases, APIs, networking, and troubleshooting across modern web application stacks • Strong communication and cross-functional influence skills • Preferred: healthcare technology or regulated SaaS experience; FHIR-native or EMR/EHR architectures; AI-assisted SRE automation; Playwright or equivalent; building internal SRE capability alongside managed services • Must be based in the United States • Position is not eligible for sponsorship

🏖️ Benefícios

• Variable compensation component • Stock options • Fully remote, collaborative engineering environment • Direct access to executive leadership • Opportunity to build the SRE function from the ground up • Opportunity to lead a blended team model • Opportunity to work on systems impacting clinical care • AI-assisted software development environment with Claude Code

Candidatar-se

Vagas Similares

🕒 6 dias atrás

Veeam Software

1001 - 5000

💼 Consultoria

📦 Logística

☁️ SaaS

Site Reliability Engineer building reliability practices for Veeam’s Government and Sovereign Cloud SaaS platform. Designing Azure infrastructure, observability, automation, and incident-response systems in regulated environments.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $138.900 - $231.400 / ano

💰 $500.000.000 Private Equity Round em 2019-01

⏰ Tempo Integral

🟠 Sênior

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 1

Filevine

201 - 500

☁️ SaaS

⚖️ Jurídico

🤖 Inteligência Artificial

Staff Site Reliability Engineer at Filevine shaping engineering culture and driving reliability practices. Leading technical standards and mentorship within a remote engineering team focused on legal AI technology.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $235.000 - $275.000 / ano

💰 $108.000.000 Series D em 2022-04

⏰ Tempo Integral

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 31

Aya Healthcare

5001 - 10000

🏥 Saúde

💼 Consultoria

📦 Logística

Manager of Site Reliability Engineering leading a team for Aya Healthcare's workforce platform. Ensuring product reliability and outstanding user experience through innovative solutions.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $230.000 - $255.000 / ano

⏰ Tempo Integral

🟠 Sênior

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 30

TalentWerx

11 - 50

🎯 Recrutamento

👥 RH Tech

🤝 B2B

DevOps Engineer IV designing and optimizing deployment solutions for Aether Aerospace. Collaborating with developers to enhance software development processes and ensure system security.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $123.601 - $159.000 / ano

⏰ Tempo Integral

🟠 Sênior

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 30

Toast

1001 - 5000

🍽️ Alimentos e Bebidas

💼 Consultoria

📦 Logística

Technical leader for release lifecycle and architecting CI/CD framework for Salesforce deployments. Driving automation and technical support in an enterprise ecosystem.

🗣️🇺🇸🇬🇧 Inglês obrigatório