Principal Site Reliability Engineer

🕒 Junho 30

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $200.000 - $250.000 / ano

⏰ Tempo Integral

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of DraftKings Inc.

DraftKings Inc.

1001 - 5000 funcionários

Fundada em 2012

📣 Marketing

💼 Consultoria

📦 Logística

Marketing • Consulting • Logistics

A DraftKings Inc. é uma empresa de entretenimento esportivo digital que opera uma das principais plataformas online de apostas esportivas, esportes de fantasia diários e cassinos, oferecendo experiências de apostas e jogos por dinheiro real via web e aplicativos móveis. Combina dados esportivos, análises e conteúdo para envolver os fãs, fornece programas de marketing e fidelização/VIP e mantém equipes globais nas áreas de engenharia, produto, conformidade e experiência do cliente, enquanto enfatiza o jogo responsável.

Descrição

• Define and execute the long-term strategy for our Kubernetes platform across Google Kubernetes Engine, Amazon Elastic Kubernetes Service, RKE2, and on-premise environments, ensuring reliability, scalability, and operational consistency. • Drive architectural decisions across critical infrastructure, including cluster lifecycle management, networking, identity and access management, observability, autoscaling, capacity planning, and cost optimization. • Lead large-scale platform initiatives across multiple engineering teams, establishing technical direction, engineering standards, and measurable outcomes that improve platform reliability and developer experience. • Establish and evolve reliability practices by defining service level objectives, service level indicators, and error budget frameworks that align platform performance with business priorities. • Build automation-first infrastructure through Infrastructure as Code, GitOps workflows, self-healing systems, and internal platform tooling that improve engineering velocity and reduce operational overhead. • Champion the responsible adoption of AI-powered engineering capabilities that improve operational efficiency, accelerate incident response, and enhance developer productivity. • Lead critical platform incidents, drive post-incident improvements, and strengthen platform resilience through automation, capacity planning, and operational excellence. • Mentor senior engineers, influence technical strategy across the organization, and elevate engineering excellence through architecture reviews, coaching, and technical leadership.

🎯 Requisitos

• A Bachelor's Degree in Computer Science or a related technical field. • At least 8 years of experience designing, operating, and scaling distributed cloud and on-premise infrastructure, including at least 3 years operating at the Staff, Principal, or equivalent technical leadership level. • Proven experience leading large-scale infrastructure or platform initiatives that require cross-functional alignment and long-term technical ownership. • Deep expertise with Kubernetes, including cluster architecture, networking, storage, security, operators, lifecycle management, and large-scale production operations. • Extensive experience building and operating production infrastructure in AWS and Google Cloud Platform using Infrastructure as Code technologies such as Terraform, Pulumi, or similar tools. • Strong software development experience in Go, Python, or both, with expertise in GitOps, continuous integration and continuous delivery, observability, distributed systems, Linux, and reliability engineering principles. • Experience incorporating AI-powered tools into engineering workflows while applying sound judgment around reliability, security, and operational risk. • Exceptional communication and leadership skills with a proven ability to mentor engineers, influence technical strategy, and drive engineering excellence. • Experience working in regulated industries, hybrid cloud environments, contributing to open-source projects, or holding cloud certifications is preferred.

🏖️ Benefícios

• bonus • equity • benefits as applicable

Candidatar-se

Vagas Similares

🕒 Junho 29

Convoso

201 - 500

💼 Consultoria

📣 Marketing

📦 Logística

Director of DevOps leading a team of engineers at Convoso, an AI-powered contact center platform. Responsible for developing and optimizing the platform and ensuring service reliability.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $220.000 - $260.000 / ano

⏰ Tempo Integral

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 29

FluidStack

11 - 50

🤖 Inteligência Artificial

Principal Operations Engineer overseeing critical operations in data centers for Fluidstack. Leading on-call escalation, root cause analysis, and operational excellence in real-time situations.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $150.000 - $250.000 / ano

⏰ Tempo Integral

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 24

Redox

201 - 500

🏥 Saúde

⚕️ Seguro de Saúde

☁️ SaaS

DevSecOps Engineer ensuring secure software development at Redox, enhancing healthcare data exchange. Collaborating with platform engineers to implement security best practices across the AWS/EKS infrastructure.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 23

Lyric - Clarity in motion.

201 - 500

🏥 Saúde

💼 Consultoria

📦 Logística

Azure DevOps Engineer at Lyric managing Azure infrastructure for healthcare technology solutions. Focus on security, reliability, and operational efficiency in a remote role.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $150.289 - $225.434 / ano

⏰ Tempo Integral

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 23

SAIC

10.000+ funcionários

☁️ SaaS

📣 Marketing

🏢 Corporativo

DevSecOps Engineer providing exceptional DevOps engineering for advancing CI/CD and automating pipelines. Must have deep proficiency in AWS, Azure, and DevSecOps tools.

🇺🇸 Estados Unidos – Remoto (EUA)

🔥 Investimento no último ano

💰 $500.000.000 Post-IPO Debt - SAIC em 2025-09

⏰ Tempo Integral

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório