Staff DevOps Engineer

Job not on LinkedIn

🕒 5 days ago

🇺🇸 United States – Remote

💵 $200k - $260k / year

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Cadence Solutions

Cadence Solutions

11 - 50 employees

Founded 2013

💼 Consulting

🏢 Enterprise

🤝 B2B

Consulting • Enterprise • B2B

Cadence Solutions is a system integrator and consulting firm focused on information management and digital transformation. They accelerate enterprise adoption of Microsoft 365, SharePoint (including SharePoint Online and Server migrations), Microsoft Purview, the Power Platform, and OpenText. Cadence offers SharePoint implementations, large-scale content migrations (including a migration solution called Peregrine), SharePoint governance and health checks, Copilot readiness and data governance for Copilot, records management, eDiscovery, data loss prevention, and hands-on Microsoft 365 training and workshops. Their services center on migration consulting, content strategy, classification and insights, and helping organizations apply governance and compliance best practices across ECM and M365 ecosystems.

📋 Description

• Own the design and continuous improvement of Cadence's cloud infrastructure, driving reliability, scalability, and secure software delivery across all environments. • Maintain and mature core Kubernetes services, including resilient networking, autoscaling, monitoring, and well-architected patterns for production workloads that support real-time patient monitoring. • Lead Terraform-based infrastructure-as-code practices, including authoring, reviewing, and enforcing standards for AI-generated IaC to ensure correctness and security before deployment. • Define and enforce DevSecOps controls across clusters, including least-privilege IAM, container image scanning, and runtime policy - ensuring infrastructure meets the compliance requirements of a regulated healthcare environment. • Sharpen observability practices using tools like Datadog, improving alerting, incident response, and the feedback loops that keep clinical systems available and performant. • Manage infrastructure spend with discipline, identifying and resolving cost inefficiencies without compromising system resilience. • Mentor engineers across the team on infrastructure best practices, raising the technical bar in pull requests, runbooks, and production operations.

🎯 Requirements

• 8+ years of hands-on DevOps or platform engineering experience, with demonstrated ownership of production cloud infrastructure at scale. • Deep experience with AWS and Kubernetes, including designing, operating, and debugging production clusters under real load. • Proficiency with Terraform, Helm, and CI/CD pipelines using GitHub Actions or comparable tooling. • Strong command of observability tooling, including Datadog or equivalent platforms, with experience building alerting systems and leading incident response. • Experience in healthcare or another highly regulated industry, with working knowledge of relevant compliance and security requirements. • Track record of mentoring engineers and raising infrastructure standards across a team. • Fluency with LLM APIs, prompt engineering, and AI-assisted development tools; demonstrated experience building or evaluating AI-powered systems in production.

🏖️ Benefits

• Competitive pay & equity* • Fully remote • Comprehensive health coverage: Medical, dental & vision • Paid time off • 401k plan + matching • Paid parental leave • Home office stipend

Apply Now

Similar Jobs

🕒 5 days ago

Counterpart Health

51 - 200

🏥 Healthcare

🤖 Artificial Intelligence

☁️ SaaS

Director of Site Reliability Engineering at Counterpart Health managing a team of SREs and enhancing infrastructure reliability. Leading strategic initiatives across multiple regions and supporting healthcare innovation.

🇺🇸 United States – Remote

💵 $187k - $243k / year

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 5 days ago

Finalsite

201 - 500

📚 Education

☁️ SaaS

🤝 B2B

Staff Site Reliability Engineer leading Finalsite's infrastructure evolution and operational excellence practices. Collaborating with engineering leadership to enhance CI/CD and multi-cloud reliability.

🇺🇸 United States – Remote

💵 $180k - $250k / year

💰 Debt financing on 2014-12

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 5 days ago

Whitespace

11 - 50

🎖️ Defense

🏛️ Government

🤖 Artificial Intelligence

Senior DevSecOps Engineer enhancing cybersecurity compliance for federal standards and DoD authorization processes. Leading secure CI/CD implementations and DevSecOps toolchain management for government projects.

🇺🇸 United States – Remote

⏰ Full Time

🟠 Senior

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 5 days ago

Martian Wall

11 - 50

🎯 Recruiter

💼 Consulting

🤝 B2B

DevOps Architect designing and managing multi-stage CI/CD systems for US-based clients. Strong expertise in cloud and DevOps tool chains is essential.

🇺🇸 United States – Remote

⏰ Full Time

🟠 Senior

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 25

Global Enterprise Services, LLC (GES)

11 - 50

💼 Consulting

📦 Logistics

Reliability Engineer responsible for cloud platform performance and incident response, managing compliance. Requires strong technical expertise and 8 years of experience.

🇺🇸 United States – Remote

⏰ Full Time

🟠 Senior

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)