Search Remote Jobs

Principal Site Reliability Engineer

đŸ”„ 2 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of DraftKings Inc.

DraftKings Inc.

1001 - 5000 employees

Founded 2012

🎼 Gaming

⚜ Sports

đŸ‘„ B2C

Gaming ‱ Sports ‱ B2C

DraftKings Inc. is a global company known for providing innovative products and experiences primarily in the sports betting and fantasy sports sectors. The company boasts a strong presence across multiple countries, aiming to deliver exceptional customer moments and overcoming challenges through teamwork and persistence. At DraftKings, innovation in engineering, analytics, and product development is key, with a focus on creating unforgettable customer experiences in sportsbook and casino operations. The company emphasizes a dynamic work culture, inclusion, equity, and global collaboration within its diverse teams.

📋 Description

‱ Define and execute the long-term strategy for the Kubernetes platform across Google Kubernetes Engine, Amazon Elastic Kubernetes Service, RKE2, and on-premise environments ‱ Drive architectural decisions for cluster lifecycle management, networking, identity and access management, observability, autoscaling, capacity planning, and cost optimization ‱ Lead large-scale platform initiatives across multiple engineering teams, establishing technical direction, engineering standards, and measurable outcomes ‱ Establish and evolve reliability practices using service level objectives, service level indicators, and error budget frameworks ‱ Build automation-first infrastructure through Infrastructure as Code, GitOps workflows, self-healing systems, and internal platform tooling ‱ Champion responsible adoption of AI-powered engineering capabilities ‱ Lead critical platform incidents, drive post-incident improvements, and strengthen platform resilience ‱ Mentor senior engineers, influence technical strategy, and elevate engineering excellence through architecture reviews, coaching, and technical leadership

🎯 Requirements

‱ Bachelor's Degree in Computer Science or a related technical field ‱ At least 8 years of experience designing, operating, and scaling distributed cloud and on-premise infrastructure ‱ At least 3 years operating at the Staff, Principal, or equivalent technical leadership level ‱ Proven experience leading large-scale infrastructure or platform initiatives requiring cross-functional alignment and long-term technical ownership ‱ Deep expertise with Kubernetes, including cluster architecture, networking, storage, security, operators, lifecycle management, and large-scale production operations ‱ Extensive experience building and operating production infrastructure in AWS and Google Cloud Platform using Infrastructure as Code technologies such as Terraform, Pulumi, or similar tools ‱ Strong software development experience in Go, Python, or both ‱ Expertise in GitOps, continuous integration and continuous delivery, observability, distributed systems, Linux, and reliability engineering principles ‱ Experience incorporating AI-powered tools into engineering workflows ‱ Exceptional communication and leadership skills, with proven ability to mentor engineers, influence technical strategy, and drive engineering excellence ‱ Experience working in regulated industries, hybrid cloud environments, contributing to open-source projects, or holding cloud certifications is preferred ‱ May be required to obtain a gaming license issued by the appropriate state agency as a condition of employment

đŸ–ïž Benefits

‱ Bonus ‱ Equity ‱ Benefits as applicable ‱ Guidance through the gaming license process if relevant to the role

Apply Now

Similar Jobs

đŸ”„ 14 hours ago

Circle

501 - 1000

💳 Fintech

₿ Crypto

🌐 Web 3

Staff SRE scaling Circle’s blockchain infrastructure, including Kubernetes platforms and full-node networks. Building AI-powered automation for Circle’s regulated digital-dollar and payments ecosystem.

đŸ”„ 16 hours ago

CACI International Inc

10,000+ employees

đŸ’Œ Consulting

đŸŽ–ïž Defense

Cloud DevOps Engineer securing AWS CI/CD pipelines, containers, and infrastructure for CACI’s national-security customers. Building DevSecOps automation, Kubernetes security, monitoring, and incident-response workflows.

🕒 Yesterday

Arcadia

201 - 500

đŸ„ Healthcare

đŸ’Œ Consulting

Revenue Operations Staff Systems Engineer owning Salesforce, Planhat, RocketLane, and AI automation. Scaling Arcadia’s healthcare go-to-market and customer operations technology ecosystem.

🕒 Yesterday

Galaxy

201 - 500

₿ Crypto

💾 Finance

Galaxy VP leading SRE and infrastructure automation across physical and virtual data-center environments. Driving IaC governance, observability, lifecycle management, and custom tooling for digital assets and AI infrastructure.

🕒 2 days ago

Intus Care

11 - 50

đŸ’Œ Consulting

📣 Marketing

📩 Logistics

Director of SRE leading reliability, QA, observability, and incident management for Intus Care’s cloud-native healthcare EMR platform. Building scalable SRE capabilities and operational standards for systems supporting value-based care.