Senior Manager, Site Reliability Engineering – DevOps

🔥 12 hours ago

🇺🇸 United States – Remote

💵 $150k - $170k / year

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

infoinfo

👻 Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Penn Mutual

Penn Mutual

1001 - 5000 employees

Founded 1847

For over 175 years, Penn Mutual has empowered individuals, families and businesses on the journey to achieve their financial goals. Through our partnership with Financial Professionals across the U.S., we help generations grow stronger by instilling the confidence and reliability that comes from a secure financial future. Penn Mutual and its affiliates offer a comprehensive suite of competitive and robust solutions to meet the unique needs of Financial Professionals and their clients, including life insurance, annuities, wealth management and institutional asset management. To learn more, including current financial strength ratings, visit www.pennmutual.com.

📋 Description

• Lead, coach, and develop DevOps engineers through goal setting, performance feedback, career development, workload planning, and delivery oversight • Remain hands-on with critical technical work, provide practical guidance, review designs and automation patterns, and model engineering discipline • Translate DevOps strategy and roadmap priorities into team plans, measurable deliverables, and sustainable operating practices • Manage daily operations for DevOps-supported services, including CI/CD pipelines, deployment automation, infrastructure automation, release support, developer tooling, and production readiness practices • Drive Infrastructure as Code, DevSecOps practices, automated testing, reusable deployment patterns, secure configuration, secret management, monitoring, and documentation • Partner with Software Engineering, SRE, Infrastructure, Architecture, Information Security, QA, Change Management, Risk, and Audit • Support AWS cloud adoption and modernization through standardized automation, secure delivery pipelines, environment management, observability, tagging, cost awareness, and operational support models • Use metrics, incident trends, stakeholder feedback, and platform health indicators to improve performance, reduce manual toil, and strengthen service quality • Participate in incident reviews, problem management, corrective action planning, ITIL change governance, audit discussions, vendor coordination, and operational improvement initiatives

🎯 Requirements

• Bachelor’s degree in Computer Science, Engineering, Information Systems, or related field; or equivalent practical experience • 10+ years of experience in DevOps, SRE, platform engineering, cloud engineering, infrastructure, software engineering, enterprise operations, or a related technology discipline • 2+ years of experience leading engineers, managing technical work, or serving in a formal or informal people leadership role • Practical knowledge of CI/CD, Infrastructure as Code, cloud services, release engineering, deployment automation, observability, DevSecOps practices, and operational support • Experience applying standards, controls, documentation, metrics, production readiness practices, and continuous improvement methods in an enterprise technology environment • Strong communication and collaboration skills, with the ability to manage priorities and explain technical issues in terms of business impact, risk, and operational outcomes • Working knowledge of IT service management, incident management, change management, compliance, audit readiness, business continuity, and technology risk practices • Strong problem-solving skills with the ability to analyze complex technical problems and propose/implement effective solutions • Preferred: experience supporting DevOps, platform engineering, cloud enablement, application modernization, or developer productivity initiatives in a regulated enterprise environment • Preferred: hands-on experience with AWS, CI/CD platforms, Infrastructure as Code, container platforms, observability tools, security scanning, secrets management, and automated deployment practices • Preferred: experience embedding security, compliance, logging, monitoring, disaster recovery, tagging, and cost management expectations into cloud and DevOps delivery practices • Preferred: experience with automation workflows, agents, and other toil-reduction tools • Preferred: experience improving release quality, change success rates, deployment frequency, incident reduction, automation maturity, and developer self-service capabilities • Preferred: relevant certifications or demonstrated expertise in cloud, DevOps, platform engineering, containerization, security, IT service management, or reliability engineering • Preferred: knowledge of Agile software development methodologies such as Agile or Kanban

Apply Now

Similar Jobs

🔥 20 hours ago

Koniag Government Services

1001 - 5000

🏛️ Government

🎖️ Defense

💼 Consulting

Senior Microsoft SRE engineering Azure, AKS, and automation platforms for Koniag’s federal government customers. Improving reliability, observability, security, and production operations.

🔥 20 hours ago

Koniag Government Services

1001 - 5000

🏛️ Government

🎖️ Defense

💼 Consulting

Senior DevSecOps Engineer securing and automating Microsoft Azure platforms. Supporting Koniag’s federal government customers with cloud security, CI/CD, AKS, and reliable digital services.

🕒 Yesterday

Gormat

11 - 50

🔒 Cybersecurity

🏛️ Government

🎖️ Defense

Cloud DevOps Engineer developing and integrating cloud-based solutions. Improving system performance, configuration, reliability, and release processes with up to 25% travel.

🕒 Yesterday

Zigabyte

51 - 200

💼 Consulting

🏥 Healthcare

📦 Logistics

DevSecOps Engineer building secure CI/CD pipelines and cloud-native infrastructure. Integrating cybersecurity, automation, and compliance controls for consulting solutions.

🕒 Yesterday

Identiq

51 - 200

💳 Fintech

🛍️ eCommerce

🔒 Cybersecurity

Founding Site Reliability Engineer building observability, incident management, and SRE practices for Incident IQ’s K-12 district workflow platform. Defining SLIs, SLOs, and reliability automation.