Senior Site Reliability Engineer

🕒 3 dias atrás

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $160.000 - $185.000 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Talkiatry

Talkiatry

501 - 1000 funcionários

Fundada em 2019

🏥 Saúde

👥 B2C

Healthcare • B2C

A Talkiatry é um fornecedor de psiquiatria virtual que oferece atendimento psiquiátrico 100% online e gestão de medicamentos. A plataforma conecta pacientes a psiquiatras licenciados e outros profissionais, oferece acompanhamento com o mesmo provedor e trata condições como TDAH, ansiedade, depressão, transtorno bipolar, TOC, TEPT, insônia e necessidades peripartum/pós-parto. A Talkiatry opera dentro da rede com grandes seguradoras, fornece estimativas de copagamento, um portal do paciente e recursos como questionários e informações sobre medicamentos. O site destaca que seus profissionais têm em média 10 anos de experiência, representam múltiplas subespecialidades e falam vários idiomas.

Descrição

• Define and roll out an SRE practice for a six-team organization: SLOs/SLIs, error budgets, and reliability standards that teams genuinely adopt. • Build and improve observability—metrics, logging, distributed tracing, dashboards, and alerting—so that more incidents are detected by monitoring before anyone outside engineering notices. • Drive down outage frequency by surfacing systemic reliability risks and partnering with teams to remediate them at the root. • Reduce toil through automation, infrastructure-as-code, and self-service tooling that teams can own and extend themselves. • Own the health and usability of our observability tooling, providing documentation and training where necessary. • Run production readiness reviews for new services and partner with engineering leadership on reliability priorities and capacity planning.

🎯 Requisitos

• 7+ years in software or infrastructure engineering, with substantial hands-on SRE or production reliability experience. • A track record of reducing incidents and improving detection—the outcomes this role is judged on. • Hands-on experience defining SLOs/SLIs and using error budgets to guide engineering decisions. • Deep observability expertise across metrics, logging, tracing, and alerting (e.g., Datadog, Prometheus, Grafana, or similar). • Strong experience operating production systems on AWS. • Proficiency with infrastructure-as-code (e.g., Terraform) and comfort building automation and tooling (Python, TypeScript, or similar).

🏖️ Benefícios

• medical, dental, vision, effective day 1 of employment • 401K with match • generous PTO plus paid holidays • paid parental leave • it all comes back to care: we’re a mental health company, and we put our team’s well-being first • grow your career with us: hone your skills and build new ones with our Learning team as Talkiatry expands

Candidatar-se

Vagas Similares

🕒 4 dias atrás

Multi Media, LLC

51 - 200

💼 Consultoria

📣 Marketing

📱 Mídia

Site Reliability Engineer optimizing infrastructure resilience and performance for a leading live streaming platform. Driving enhancement and automation of cloud-based infrastructure with a global network.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $169.000 - $215.000 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 4 dias atrás

Thumbtack

1001 - 5000

🏪 Marketplace

☁️ SaaS

Senior Software Engineer designing and maintaining scalable systems to improve reliability and efficiency at Thumbtack. Collaborating with cross-functional teams to optimize platform services.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $179.400 - $232.100 / ano

💰 $75.000.000 Debt Financing - Thumbtack em 2024-07

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 4 dias atrás

Manulife

10.000+ funcionários

🛡️ Seguros

💸 Finanças

Lead Power Platform Reliability Engineer enhancing enterprise-level solutions through collaboration and mentorship. Shape future data-driven applications and drive cloud integration.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 4 dias atrás

ICF

5001 - 10000

💼 Consultoria

🏛️ Governo

🏥 Saúde

DevOps Engineer building healthcare reporting services for ICF. Implementing cloud-based solutions using AWS and fostering collaboration on CI/CD pipeline improvements.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $108.476 - $184.409 / ano

💰 $29.000.000 Grant em 2023-03

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 4 dias atrás

ICF

5001 - 10000

💼 Consultoria

🏛️ Governo

🏥 Saúde

Senior DevOps Engineer delivering best in class healthcare reporting services for ICF. Working collaboratively to implement cloud solutions and establish CI/CD pipelines using AWS.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $108.476 - $184.409 / ano

💰 $29.000.000 Grant em 2023-03

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório