Site Reliability Engineer

🕒 Junho 27

🏄 California – Remoto

infoinfo

💵 $130.000 - $160.000 / ano

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

👻 Score fantasma 14%

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Berkeley Research Group (BRG)

Berkeley Research Group (BRG)

1001 - 5000 funcionários

🏗️ Construção

📦 Logística

🎖️ Defesa

💰 Venture Round em 2020-07

Construction • Logistics • Defense

A BRG combina credenciais acadêmicas de renome mundial com expertise empresarial testada em campo, criada para agilidade e conectividade, o que nos diferencia e coloca nossos clientes à frente. Nossos especialistas de primeira linha incluem líderes experientes da indústria, acadêmicos renomados e cientistas de dados pioneiros. Juntos, eles trazem uma diversidade de experiências comprovadas do mundo real para economia, disputas e investigações; finanças corporativas; e serviços de melhoria de desempenho que abordam os desafios mais complexos para organizações ao redor do mundo. Nossa estrutura única nutre os relacionamentos interdisciplinares que nos dão a vantagem, estabelecendo as bases para insights mais informados e pensamentos mais originais e incisivos de perspectivas diversas que, quando combinados com nosso alcance e recursos globais, nos tornam exclusivamente capazes de enfrentar os desafios de nossos clientes.

Descrição

• Design, implement, and maintain scalable and reliable systems in cloud environments such as Azure Cloud Services. • Provide operational support for full-stack software applications. • Increase system resilience with expert-level coding, bulletproof release, and change management skills. • Develop service-level indicators and objectives to automate release validation. • Improve automation and increase the system’s self-healing capability. • Collect operating system data and report performance metrics to stakeholders. • Ensure security best practices are followed in cloud infrastructure and application deployments. • Manage cloud and database system maintenance, debugging production issues as they arise. • Improve reliability, quality, and time-to-market of our suite of software solutions. • Partner with security and product teams to define and publish policies, processes, and playbooks to facilitate rapid and effective handling of alerts and incidents. • Lead incident management processes; respond to outages and service disruptions promptly.

🎯 Requisitos

• Bachelor’s degree in computer science or similar field. • Five years’ experience as a site reliability engineer or similar role. • Strong programming skills (Golang, Ruby, Python, or similar). • Proven ability to diagnose and monitor performance and reliability issues across the stack. • Expertise in Kubernetes. • Relevant industry certifications, such as through the Site Reliability Engineering (SRE) Foundation. • Proven experience working with cloud-native infrastructure (Azure Cloud Services, AWS, or GCP). • Experience working with observability and incident management tools (Datadog, OpsGenie, PagerDuty). • Experience scripting operating system tasks with Infrastructure as Code. • Impeccable communication skills. • Ability to problem-solve in a fast-paced, high-stakes environment. • Candidate must be able to submit verification of his/her legal right to work in the United States, without company sponsorship.

Candidatar-se

Vagas Similares

🕒 Junho 26

Hearst Health

1001 - 5000

🏥 Saúde

⚕️ Seguro de Saúde

☁️ SaaS

Senior DevOps Engineer helping build and improve healthcare technology platforms at Zynx Health. Operating cloud infrastructure, deployment pipelines, and security practices for clinical decision support solutions.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $112.000 - $135.000 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 26

THEMIS Waste Recovery Technology

11 - 50

🏥 Saúde

📦 Logística

🍽️ Alimentos e Bebidas

DevOps Engineer managing cloud infrastructure and CI/CD pipelines for secure, reliable operations at Themis. Ensuring system health and performance while automating processes and maintaining security controls.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 25

GeekPlus

51 - 200

🔧 Hardware

Controls & Deployment Engineer overseeing the deployment of robotic systems at client sites for Geek+. Responsible for technical integration and project execution in warehouse automation.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 25

Towne Park

10.000+ funcionários

🏨 Hospitalidade

🏥 Saúde

🚗 Transporte

DevSecOps Engineer managing Azure DevOps CI/CD pipelines and security policies. Collaborating across teams to enhance deployment reliability and security for cloud infrastructure.

🇺🇸 Estados Unidos – Remoto (EUA)

💰 $10.000.000 Private Equity Round em 2007-08

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 24

Knock

1 - 10

🔌 API

☁️ SaaS

🏢 Corporativo

DevOps Engineer at Knock, responsible for platform scalability and reliability. Building, scaling, and maintaining core services as a remote-first team.

🗣️🇺🇸🇬🇧 Inglês obrigatório