
501 - 1000 employees
đź Consulting
đĽ Healthcare
đŚ Logistics
đ° Private equity on 2013-03
Consulting ⢠Healthcare ⢠Logistics
Cast & Crew is a provider of cloud-based software and professional services that support the full lifecycle of film, television, streaming and live-event productions. They offer production accounting, payroll and HR, timekeeping, digital onboarding, content and asset collaboration, reporting/analytics, and industry-specific services such as tax incentives guidance, workers' compensation, residuals, and financing. Cast & Crew serves production companies and entertainment businesses with B2B SaaS solutions and managed services to streamline workflows and ensure compliance across global productions.
đĽ 3 minutes ago
đşđ¸ United States â Remote
đľ $190k - $235k / year
â° Full Time
đ´ Lead
â DevOps & Site Reliability Engineer (SRE)
đŚ H1B Visa Sponsor
đť Ghost score 0%
Improve your chances of getting an interview by checking your resume score before you apply.

501 - 1000 employees
đź Consulting
đĽ Healthcare
đŚ Logistics
đ° Private equity on 2013-03
Consulting ⢠Healthcare ⢠Logistics
Cast & Crew is a provider of cloud-based software and professional services that support the full lifecycle of film, television, streaming and live-event productions. They offer production accounting, payroll and HR, timekeeping, digital onboarding, content and asset collaboration, reporting/analytics, and industry-specific services such as tax incentives guidance, workers' compensation, residuals, and financing. Cast & Crew serves production companies and entertainment businesses with B2B SaaS solutions and managed services to streamline workflows and ensure compliance across global productions.
⢠Architect and continuously improve Azure DevOps CI/CD pipelines, including pipeline-as-code standards, templating strategies, and artifact promotion workflows ⢠Own the health and evolution of AWS EKS clusters, including node lifecycle, autoscaling, networking, RBAC, and upgrades ⢠Design and enforce Infrastructure-as-Code practices and champion GitOps patterns ⢠Drive platform reliability improvements using New Relic observability data in partnership with SRE ⢠Define and maintain golden-path templates for containerized workloads, including Dockerfile standards and Helm chart libraries ⢠Partner with engineering teams to onboard services and reduce toil through automation ⢠Escalate and coordinate complex infrastructure incidents through PagerDuty, participate in on-call rotation, and lead post-incident reviews ⢠Identify recurring failure modes and drive fixes that reduce page volume and MTTR ⢠Maintain runbooks and platform documentation in Confluence ⢠Define and socialize DevOps standards for pipelines, containers, secrets, and deployment safety ⢠Conduct architecture reviews and provide technical guidance ⢠Mentor senior and mid-level engineers through pairing, code review, and knowledge sharing ⢠Identify tooling gaps and build business cases for platform investments
⢠8+ years of DevOps or platform engineering experience ⢠At least 2 years operating at a Staff or Principal level in an organization of 100+ engineers ⢠Deep, hands-on expertise with Kubernetes, specifically AWS EKS preferred, including workloads, networking, storage, and cluster operations at scale ⢠Strong command of Azure DevOps Pipelines, including YAML pipeline authoring, library management, service connections, and environment promotion gates ⢠Proven experience designing and maintaining CI/CD systems for microservice architectures with multiple independent teams ⢠Experience operating observability platforms such as New Relic or Datadog to drive proactive reliability improvements ⢠Proficiency in Python, Bash, or Go ⢠Proficiency with Infrastructure-as-Code tooling such as Terraform, Pulumi, or CDK ⢠Familiarity with feature flag patterns and progressive delivery; Unleash or equivalent is a plus ⢠Excellent written communication skills ⢠Experience with data engineering or ML infrastructure workloads on Kubernetes is preferred ⢠Background contributing to or maintaining internal developer portals is preferred ⢠Familiarity with FinOps practices and tooling for AWS cost attribution and optimization is preferred ⢠Experience in SRE-adjacent roles and comfort with SLO/SLI definition and error budget policy is preferred
⢠Comprehensive medical, dental, and vision coverage ⢠401(k) match ⢠Generous PTO ⢠Paid parental leave ⢠Health and wellness programs ⢠Tuition reimbursement ⢠Employee discounts
Apply NowđĽ 5 hours ago
Principal SRE architecting reliable network infrastructure for Akamaiâs globally distributed cloud and edge platform. Automating operations, defining SLOs, and troubleshooting large-scale network systems.
đşđ¸ United States â Remote
đľ $169.3k - $304.7k / year
đ° Post-IPO Equity on 2001-07
â° Full Time
đ´ Lead
â DevOps & Site Reliability Engineer (SRE)
đŚ H1B Visa Sponsor
đĽ 23 hours ago
Principal DevOps Architect designing secure AWS infrastructure for global multi-tenant SaaS services. Building Terraform, CI/CD, Kubernetes, observability, and HIPAA-compliant AI/ML platforms.
đ Yesterday
Principal DevOps Engineer leading AWS, Kubernetes, and Terraform infrastructure for Campspotâs campground reservation software and camping marketplace. Driving automation, reliability, security, and cloud-cost optimization.
đşđ¸ United States â Remote
đľ $155k - $180k / year
â° Full Time
đ´ Lead
â DevOps & Site Reliability Engineer (SRE)
đ 2 days ago
Staff SRE improving reliability for Fingerprintâs device-intelligence platform. Defining SLOs, strengthening incident operations, and enabling safe AI-assisted system operations.
đşđ¸ United States â Remote
đľ $177k - $240k / year
đ° $32M Series B on 2021-11
â° Full Time
đ´ Lead
â DevOps & Site Reliability Engineer (SRE)
đ 4 days ago
Principal DevOps Architect owning AWS, Terraform, CI/CD, observability, and AI/ML platforms. Ensuring HIPAA and SOC 2 compliance for clinical research SaaS.
đşđ¸ United States â Remote
đľ $155k - $195k / year
đ° Private Equity Round on 2022-01
â° Full Time
đ´ Lead
â DevOps & Site Reliability Engineer (SRE)