Site Reliability Engineering Manager

🕒 Julho 27

🐊 Florida – Remoto

infoinfo

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

infoinfo

👻 Score fantasma 10%

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of NationsBenefits

NationsBenefits

1001 - 5000 funcionários

Fundada em 2013

🏥 Saúde

💼 Consultoria

📦 Logística

💰 Private Equity Round em 2022-04

Healthcare • Consulting • Logistics

A NationsBenefits é uma empresa líder em tecnologia de saúde que se especializa em fornecer soluções de tecnologia financeira e gestão de benefícios suplementares. A empresa oferece uma variedade de serviços, incluindo cuidados auditivos através do NationsHearing, produtos de saúde e bem-estar via NationsOTC, e entrega de refeições através do NationsMarket. A NationsBenefits também oferece serviços especializados como assistência de emergência, transporte para cuidados de saúde e orientação personalizada de saúde utilizando inteligência artificial. Suas plataformas proprietárias, incluindo Benefits Pro™, facilitam o suporte aos membros, configuração de benefícios e transações de e-commerce. A NationsBenefits foca em melhorar os resultados dos membros, fechar lacunas no atendimento e aumentar a satisfação através de análises avançadas e programas personalizados.

Descrição

• Lead, mentor, and develop a US-based team of Site Reliability Engineers • Conduct regular 1:1s, performance reviews, and career development discussions • Own hiring, onboarding, and retention efforts as the team scales • Foster a culture of ownership, blameless postmortems, and continuous improvement • Lead day-to-day production operations and ensure timely incident triage, resolution, and escalation • Serve as an escalation point and incident commander for major production incidents • Drive problem management and root cause analysis processes • Carry PagerDuty on-call escalation responsibilities for critical issues • Track and report operational KPIs, SLAs, and SLOs, including availability, MTTR, and incident trends • Improve system reliability, observability, and resilience using Datadog and related tooling • Drive automation, self-healing capabilities, and runbook maturity • Partner with Development, DevOps, DevSecOps, and Engineering teams to embed reliability into the SDLC • Contribute hands-on to tooling, automation, and technical reviews as needed • Coordinate closely with SRE leadership in India to ensure seamless follow-the-sun coverage • Represent the US SRE organization in cross-functional planning and operational reviews • Communicate effectively with both technical and non-technical stakeholders • Maintain high-quality documentation for incidents, postmortems, runbooks, and operational procedures • Ensure adherence to healthcare and fintech compliance standards, including HIPAA, PCI DSS, SOC 2, ISO 27001, and HITRUST

🎯 Requisitos

• 5–8 years of experience in Site Reliability Engineering, DevOps, Production Support, or Platform Engineering • 1–2+ years of experience leading, mentoring, or managing engineers • Demonstrated success operating in a player-coach leadership model • Strong hands-on experience with production incident management and escalation processes • Proficiency with Datadog or similar observability platforms • Hands-on experience with Kubernetes and Docker in production environments • Strong scripting or programming skills in PowerShell, Bash, Python, Java, or C# • Experience with Helm, CI/CD pipelines, and deployment automation • Working knowledge of ITIL processes and Agile methodologies • Experience working with SQL, MySQL, or NoSQL databases • Excellent communication and stakeholder management skills • Willingness to participate in PagerDuty on-call escalation and work within a global follow-the-sun operating model

🏖️ Benefícios

• Competitive compensation and comprehensive benefits • Unlimited PTO • Fully remote work environment (US-based) • Opportunity to lead and grow a high-impact SRE organization • Exposure to modern cloud-native technologies and large-scale reliability challenges • Collaborative culture focused on innovation, learning, and continuous improvement • Meaningful work that directly impacts healthcare technology and millions of members

Candidatar-se

Vagas Similares

🕒 Julho 27

Smithfield Foods

10.000+ funcionários

🏭 Manufatura

🌾 Agricultura

🍽️ Alimentos e Bebidas

Sr. Utilities Engineer optimizing and managing utility systems at Smithfield Foods. Focusing on industrial refrigeration, boiler, and compressed air systems for manufacturing processes.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 27

Dropzone AI

51 - 200

🤖 Inteligência Artificial

Senior DevOps Engineer at Dropzone AI enhancing infrastructure for our AI cybersecurity platform. Collaborating with engineering teams on scalable, resilient, and secure systems.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $170.000 - $185.000 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 27

Counterpart Health

51 - 200

🏥 Saúde

🤖 Inteligência Artificial

☁️ SaaS

Senior Site Reliability Engineer scaling Kubernetes and cloud infrastructure for Counterpart Health’s AI-enabled primary care platform. Automating deployments, reducing toil, and improving reliability for healthcare workloads.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $160.000 - $208.000 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 27

Otoe Missouria Group

11 - 50

🏛️ Governo

🔒 Cibersegurança

💼 Consultoria

DevSecOps Engineer supporting federal agency application modernization efforts in Washington, DC. Building secure, efficient pipelines for mission-critical systems.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 27

Imagineeer

11 - 50

🏛️ Governo

🔒 Cibersegurança

💼 Consultoria

DevOps Engineer specializing in SharePoint and Microsoft 365 platform deployment automation. Collaborating with teams to enhance cybersecurity, IT modernization in a federal context.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $120.000 - $130.000 / ano

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório