Senior Site Reliability Engineer

🕒 Julho 28

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $137.900 - $221.400 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

infoinfo

👻 Score fantasma 23%

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of ServiceTitan

ServiceTitan

1001 - 5000 funcionários

Fundada em 2012

💼 Consultoria

📦 Logística

📣 Marketing

💰 $200.000.000 Series G em 2021-06

Consulting • Logistics • Marketing

ServiceTitan é uma plataforma de software abrangente projetada para a indústria de serviços, oferecendo soluções para melhorar a produtividade e a rentabilidade das empresas. Apresenta uma variedade de recursos, incluindo despacho, agendamento, marketing, relatórios e ferramentas de experiência do cliente, adaptadas para áreas como encanamento, HVAC, serviços elétricos e mais. A ServiceTitan busca capacitar as empresas ao otimizar operações, melhorar o fluxo de caixa e oferecer experiências superiores aos clientes através de uma plataforma tudo-em-um. O software inclui análises de dados em tempo real, opções de financiamento e capacidades móveis para apoiar as necessidades operacionais dos empreiteiros e aumentar suas fontes de receita. Ao consolidar múltiplas funções de negócios em uma única plataforma, a ServiceTitan visa ajudar os empreiteiros a crescer de forma rentável e eficiente.

Descrição

• Participate in an on-call rotation and diagnose and resolve production issues using runbooks and playbooks • Design, build, and maintain observability dashboards and alerting based on SLIs and SLOs • Operate and improve the Kubernetes-based compute platform • Work across Azure/AWS cloud networking and infrastructure • Investigate and resolve production incidents, including root-cause analysis and remediation • Partner with product engineering teams to review architecture and infrastructure decisions • Build and maintain automation to reduce manual operational work • Write and maintain runbooks and documentation • Help define scalability, availability, and performance requirements for new systems • Collaborate across engineering teams to adopt reliability and observability best practices • Contribute to CI/CD pipelines and help teams ship changes safely and quickly

🎯 Requisitos

• Strong, hands-on understanding of Kubernetes • Practical experience with SLIs, SLOs, and error budgets • Solid grounding in AWS or Azure • Networking fundamentals, including subnetting and IP addressing • Deep experience with at least one observability stack: OpenTelemetry, Prometheus, Grafana, Datadog, or Elasticsearch • Strong understanding of a CI/CD system; GitHub Actions preferred, with TeamCity, Azure DevOps, or GitLab CI acceptable • Strong programming skills and ability to build web applications • Ideally, working knowledge of .NET and ASP.NET; strong Python with Flask/FastAPI or Java with Spring also accepted • Experience with distributed systems and common failure modes, including retries, timeouts, and cascading failures • Strong production troubleshooting skills • 8–10+ years of relevant hands-on experience • Database experience is nice-to-have and not mandatory

🏖️ Benefícios

• Flexible time off • Learning and development opportunities • Comprehensive onboarding program • Leadership training • Programs and events • Bonusly rewards • Peer-nominated awards • Company-paid medical, dental, and vision insurance • 100% employer-paid options and 90% coverage for dependents • FSA and HSA • 401(k) match • Telehealth options, including One Medical memberships • Parental leave and support • Up to $20k in fertility services • Surrogacy and adoption reimbursement • On-demand maternity support through Maven Maternity • Free breast milk shipping through Maven Milk • Pet insurance • Legal advisory services • Financial planning tools • Annual bonus • Equity • Holistic benefits suite • Flexible/autonomous work support

Candidatar-se

Vagas Similares

🕒 Julho 28

TensorWave

11 - 50

🤖 Inteligência Artificial

🏢 Corporativo

☁️ SaaS

DevOps Software Engineer building and maintaining software layer connecting infrastructure systems for AI workloads. Collaborating with teams to develop automation, orchestration, and communication across platforms.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 28

Direct Care Innovations

51 - 200

🏥 Saúde

☁️ SaaS

🏢 Corporativo

Senior DevOps & Cloud Engineer managing Azure infrastructure for Direct Care Innovations' SaaS platform. Collaborating on scalable, reliable, and secure applications and systems while mentoring engineers.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 28

Repario

51 - 200

💼 Consultoria

📦 Logística

⚖️ Jurídico

DevOps Engineer developing code and scripts for automated deployments in legal eDiscovery sector. Collaborating across teams, ensuring security, and mentoring junior staff.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 28

IGAMINGHUNT

11 - 50

🎯 Recrutamento

🎲 Jogos de Azar

🤝 B2B

DevSecOps Engineer responsible for securing cloud infrastructure and CI/CD pipelines in a fintech and iGaming company. Delivering scalable products with a security-first mindset.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 28

CertifyOS

201 - 500

🏥 Saúde

☁️ SaaS

🤝 B2B

Senior Site Reliability Engineer designing and maintaining reliable infrastructure at CertifyOS for healthcare data systems. Influencing architecture and ensuring uptime for millions of provider records.

🇺🇸 Estados Unidos – Remoto (EUA)

💰 $40.000.000 Series B - Certify em 2025-06

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório