Site Reliability Engineer

Vaga não está no LinkedIn

🕒 Junho 4

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $175.000 - $200.000 / ano

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Tern

Tern

11 - 50 funcionários

💼 Consultoria

📦 Logística

📣 Marketing

Consulting • Logistics • Marketing

A Tern é uma empresa focada em fornecer cartões flexíveis e ferramentas fintech acessíveis, projetadas para aumentar a receita e simplificar processos para negócios e empresas. Eles oferecem uma variedade de soluções bancárias integradas, como cartões pré-pagos virtuais e físicos, transferências bancárias, transações internacionais e suporte de conformidade. A plataforma da Tern está equipada com soluções de baixo código/sem código, APIs e análises de dados inteligentes, tornando-a amigável e eficiente para empresas que desejam lançar rapidamente produtos financeiros. Com o objetivo de democratizar os serviços fintech, a Tern busca tornar essas ferramentas acessíveis a um público amplo, eliminando barreiras e promovendo a inovação no espaço de tecnologia financeira.

Descrição

• Own the migration from Heroku to Google Cloud Platform, architecture, execution, and a cutover that doesn't surprise anyone • Build and maintain the Postgres core, Fivetran pipeline, BigQuery data layer, and Hex reporting infrastructure • Optimize the hot paths that matter most: key backend code paths and our heaviest third-party syncs, so performance holds as volume climbs • Own monitoring, alerting, cost reduction, and proactive scaling: surface problems early, keep spend sane, and stay ahead of growth rather than reacting to it • Lead incident response and write post-mortems that turn an outage into a permanent fix and a smarter team • Set the operational bar across engineering and pull others up to it

🎯 Requisitos

• Production reliability ownership: Track record of personally owning production reliability at meaningful scale. Concrete stories of incidents you led, fixed, and prevented from recurring, not just participated in. This is a primary responsibility, not something you've done on the side. • Infrastructure migrations: Real experience owning a cloud migration end to end, not just contributing to one. Fluent in GCP (or a comparable cloud), infrastructure-as-code, and the failure modes of distributed systems. • Observability and proactive operations: You build monitoring and alerting that surfaces problems before users find them. You know what to instrument, what to alert on, and what's just noise. • High agency: You find the highest leverage reliability problems and go fix them without being assigned to them. You don't wait for an outage to justify the work. • AI in your working habits: Specific examples of how AI has made your debugging, automation, or operational workflows faster or more reliable.

🏖️ Benefícios

• Own the migration from Heroku to Google Cloud Platform • Build and maintain the Postgres core, Fivetran pipeline, BigQuery data layer, and Hex reporting infrastructure • Optimize the hot paths that matter most: key backend code paths and our heaviest third-party syncs, so performance holds as volume climbs • Own monitoring, alerting, cost reduction, and proactive scaling: surface problems early, keep spend sane, and stay ahead of growth rather than reacting to it • Lead incident response and write post-mortems that turn an outage into a permanent fix and a smarter team • Set the operational bar across engineering and pull others up to it

Candidatar-se

Vagas Similares

🕒 Junho 3

AuthZed

11 - 50

🔌 API

🔒 Cibersegurança

☁️ SaaS

Site Reliability Engineer responsible for maintaining systems reliability and performance at AuthZed. Collaborate globally while developing scalable infrastructure solutions for a cutting-edge authorization platform.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 3

Amwell

501 - 1000

💼 Consultoria

🛡️ Seguros

🏥 Saúde

Senior Systems Engineer managing cloud and on-prem infrastructure at Amwell. Enhancing operational efficiency through automation and system management tools.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $129.330 - $140.000 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 3

Real

51 - 200

💼 Consultoria

📦 Logística

🏨 Hospitalidade

Senior DevSecOps Engineer leading DevOps & Security team at Real. Managing AWS, Kubernetes, and infrastructure as code with a focus on security practices.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 3

Endeavor

5001 - 10000

🏨 Hospitalidade

📣 Marketing

Senior DevOps Engineer for WME building Azure infrastructure and improving system reliability in entertainment industry. Leading cloud resources, Kubernetes deployments, and CI/CD pipelines.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 2

Accenture Federal Services

10.000+ funcionários

💼 Consultoria

🎖️ Defesa

📦 Logística

DevOps Engineer responsible for establishing CI/CD pipelines for SAP environments at Accenture Federal Services. Collaborating with teams to enhance operational efficiency and security compliance.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $100.200 - $203.400 / ano

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório