Search Remote Jobs

Staff DevOps Engineer

Job not on LinkedIn

🔥 6 minutes ago

🏄 California – Remote

infoinfo

💵 $190k - $235k / year

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Cast & Crew

Cast & Crew

501 - 1000 employees

☁️ SaaS

📱 Media

👥 HR Tech

💰 Private equity on 2013-03

SaaS • Media • HR Tech

Cast & Crew is a provider of cloud-based software and services that support the full production lifecycle for the entertainment industry. The company offers production accounting and AP software (PSL+), digital onboarding and timekeeping tools (Start+, Hours+), collaboration and content tools (Studio+), reporting/data integrations, and a full-service payroll and financial services suite including residuals, workers' compensation, tax-incentive guidance, and related production HR resources. Cast & Crew serves film, television, streaming and live entertainment productions, delivering B2B SaaS products plus specialized payroll and workforce support for production employers and crews.

📋 Description

• Serve as a technical anchor for the platform engineering practice • Design and evolve CI/CD pipelines in Azure DevOps, including pipeline-as-code standards, templating strategies, and artifact promotion workflows • Own the health and evolution of AWS EKS clusters, including node lifecycle, autoscaling, networking, RBAC, and cluster upgrades • Design and enforce Infrastructure-as-Code practices and champion GitOps patterns • Drive platform reliability improvements using observability data from New Relic in partnership with SRE • Define and maintain golden-path templates for containerized workloads, including Dockerfile standards and Helm chart libraries • Partner with engineering teams to onboard services and reduce toil through automation • Act as an escalation point for complex infrastructure incidents through PagerDuty • Participate in the on-call rotation and lead post-incident reviews • Identify recurring failure modes and drive systemic fixes to reduce page volume and MTTR • Maintain runbooks and platform documentation in Confluence • Define and socialize DevOps standards across the engineering organization • Conduct architecture reviews and provide technical guidance • Mentor senior and mid-level engineers through pairing, code review, and knowledge sharing • Identify tooling gaps and build business cases for platform investments

🎯 Requirements

• 8+ years of DevOps or platform engineering experience • At least 2 years operating at a Staff or Principal level in an organization of 100+ engineers • Deep hands-on expertise with Kubernetes, including troubleshooting workloads, networking, storage, and cluster operations at scale • Experience with AWS EKS preferred • Strong command of Azure DevOps Pipelines, including YAML pipeline authoring, library management, service connections, and environment promotion gates • Proven experience designing and maintaining CI/CD systems for microservice architectures with multiple independent teams • Experience operating observability platforms such as New Relic or Datadog • Proficiency in Python, Bash, or Go • Proficiency with Infrastructure-as-Code tooling such as Terraform, Pulumi, or CDK • Familiarity with feature flag patterns and progressive delivery; Unleash or equivalent is a plus • Excellent written communication skills • Preferred: experience with data engineering or ML infrastructure workloads on Kubernetes • Preferred: background contributing to or maintaining internal developer portals such as Backstage • Preferred: familiarity with FinOps practices and AWS cost attribution and optimization • Preferred: experience in SRE-adjacent roles and comfort with SLO/SLI definition and error budget policy

🏖️ Benefits

• Comprehensive medical coverage • Dental coverage • Vision coverage • 401(k) match • Generous PTO • Paid parental leave • Health and wellness programs • Tuition reimbursement • Employee discounts

Apply Now

Similar Jobs

🔥 14 minutes ago

Cast & Crew

501 - 1000

💼 Consulting

🏥 Healthcare

📦 Logistics

Staff DevOps Engineer owning AWS EKS, Terraform, and Azure DevOps platforms for Cast & Crew’s entertainment technology business. Improving reliability, developer experience, and infrastructure standards.

🕒 2 days ago

Clinician Nexus

51 - 200

🏥 Healthcare

⚕️ Healthcare Insurance

📚 Education

DevOps Manager leading secure, reliable platform engineering for Clinician Nexus, a healthcare workforce technology company. Balancing team leadership with hands-on AWS, Kubernetes, Terraform, and CI/CD engineering.

🕒 2 days ago

CVS Health

10,000+ employees

🏥 Healthcare

⚕️ Healthcare Insurance

🛒 Retail

Staff DevSecOps Engineer securing CVS Health’s Health 100 application portfolio. Leading CI/CD automation, tool migrations, mobile security, and vulnerability reduction.

🕒 3 days ago

Peraton

10,000+ employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Site Reliability Engineer operating Peraton’s AWS, GovCloud, and ROSA production infrastructure. Improving observability, incident response, resilience, and deployment automation for national security systems.

🕒 3 days ago

Peraton

10,000+ employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Site Reliability Engineer operating Peraton’s AWS and OpenShift production infrastructure. Managing reliability, observability, incident response, releases, and infrastructure automation for national security missions.