Principal Site Reliability Engineer

🕒 4 dias atrás

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $200.000 - $250.000 / ano

⏰ Tempo Integral

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of DraftKings Inc.

DraftKings Inc.

1001 - 5000 funcionários

Fundada em 2012

🎮 Jogos

⚽ Esportes

👥 B2C

Gaming • Sports • B2C

A DraftKings Inc. é uma empresa global reconhecida por oferecer produtos e experiências inovadoras, principalmente nos setores de apostas esportivas e esportes fantasy. A companhia tem forte presença em diversos países, com o objetivo de proporcionar momentos excepcionais aos clientes e superar desafios por meio de trabalho em equipe e persistência. Na DraftKings, a inovação em engenharia, analytics e desenvolvimento de produto é essencial, com foco em criar experiências inesquecíveis para os clientes em operações de sportsbook e cassino. A empresa valoriza uma cultura de trabalho dinâmica, inclusão, equidade e colaboração global entre suas equipes diversas.

Descrição

• Define and execute the long-term strategy for the Kubernetes platform across Google Kubernetes Engine, Amazon Elastic Kubernetes Service, RKE2, and on-premise environments • Drive architectural decisions for cluster lifecycle management, networking, identity and access management, observability, autoscaling, capacity planning, and cost optimization • Lead large-scale platform initiatives across multiple engineering teams, establishing technical direction, engineering standards, and measurable outcomes • Establish and evolve reliability practices using service level objectives, service level indicators, and error budget frameworks • Build automation-first infrastructure through Infrastructure as Code, GitOps workflows, self-healing systems, and internal platform tooling • Champion responsible adoption of AI-powered engineering capabilities • Lead critical platform incidents, drive post-incident improvements, and strengthen platform resilience • Mentor senior engineers, influence technical strategy, and elevate engineering excellence through architecture reviews, coaching, and technical leadership

🎯 Requisitos

• Bachelor's Degree in Computer Science or a related technical field • At least 8 years of experience designing, operating, and scaling distributed cloud and on-premise infrastructure • At least 3 years operating at the Staff, Principal, or equivalent technical leadership level • Proven experience leading large-scale infrastructure or platform initiatives requiring cross-functional alignment and long-term technical ownership • Deep expertise with Kubernetes, including cluster architecture, networking, storage, security, operators, lifecycle management, and large-scale production operations • Extensive experience building and operating production infrastructure in AWS and Google Cloud Platform using Infrastructure as Code technologies such as Terraform, Pulumi, or similar tools • Strong software development experience in Go, Python, or both • Expertise in GitOps, continuous integration and continuous delivery, observability, distributed systems, Linux, and reliability engineering principles • Experience incorporating AI-powered tools into engineering workflows • Exceptional communication and leadership skills, with proven ability to mentor engineers, influence technical strategy, and drive engineering excellence • Experience working in regulated industries, hybrid cloud environments, contributing to open-source projects, or holding cloud certifications is preferred • May be required to obtain a gaming license issued by the appropriate state agency as a condition of employment

🏖️ Benefícios

• Bonus • Equity • Benefits as applicable • Guidance through the gaming license process if relevant to the role

Candidatar-se

Vagas Similares

🕒 5 dias atrás

Circle

501 - 1000

💳 Fintech

₿ Cripto

🌐 Web 3

Staff SRE scaling Circle’s blockchain infrastructure, including Kubernetes platforms and full-node networks. Building AI-powered automation for Circle’s regulated digital-dollar and payments ecosystem.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 6 dias atrás

Galaxy

201 - 500

₿ Cripto

💸 Finanças

Galaxy VP leading SRE and infrastructure automation across physical and virtual data-center environments. Driving IaC governance, observability, lifecycle management, and custom tooling for digital assets and AI infrastructure.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 6 dias atrás

Intus Care

11 - 50

💼 Consultoria

📣 Marketing

📦 Logística

Director of SRE leading reliability, QA, observability, and incident management for Intus Care’s cloud-native healthcare EMR platform. Building scalable SRE capabilities and operational standards for systems supporting value-based care.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $175.000 - $200.000 / ano

💰 $13.100.000 Venture Round em 2023-01

⏰ Tempo Integral

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 6 dias atrás

Veeam Software

1001 - 5000

💼 Consultoria

📦 Logística

☁️ SaaS

Site Reliability Engineer building reliability practices for Veeam’s Government and Sovereign Cloud SaaS platform. Designing Azure infrastructure, observability, automation, and incident-response systems in regulated environments.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $138.900 - $231.400 / ano

💰 $500.000.000 Private Equity Round em 2019-01

⏰ Tempo Integral

🟠 Sênior

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 1

Filevine

201 - 500

☁️ SaaS

⚖️ Jurídico

🤖 Inteligência Artificial

Staff Site Reliability Engineer at Filevine shaping engineering culture and driving reliability practices. Leading technical standards and mentorship within a remote engineering team focused on legal AI technology.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $235.000 - $275.000 / ano

💰 $108.000.000 Series D em 2022-04

⏰ Tempo Integral

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório