Manager – Site Reliability Engineer

🔥 1 hour ago

🇺🇸 United States – Remote

💵 $123.5k - $163.9k / year

⏰ Full Time

🟠 Senior

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

infoinfo

👻 Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Frontier Airlines

Frontier Airlines

5001 - 10000 employees

Founded 1994

✈️ Travel

📦 Logistics

🚗 Transport

Travel • Logistics • Transport

Frontier Airlines is a U. S. -based low-cost commercial airline that operates scheduled passenger flights across domestic and select international routes. It provides online booking and trip management, a frequent-flyer program (FRONTIER Miles), ancillary services (bag fees, seat selection, bundles like Discount Den and GoWild Pass), co-branded credit card offers, and customer support and travel tools. The airline emphasizes cost-conscious travel, promotions and sustainability initiatives (branding as a green airline).

📋 Description

• Lead the design and implementation of highly available, resilient, and scalable cloud platforms using AWS and cloud-native technologies • Define and execute reliability strategies, SLOs, SLIs, and error budgets across critical business services • Architect and operate Kubernetes-based platforms for containerized and microservices applications • Establish enterprise standards for reliability, availability, disaster recovery, resiliency testing, and operational excellence • Partner with software engineering, platform engineering, security, and operations teams to improve reliability and reduce operational risk • Drive DevSecOps, CI/CD automation, Infrastructure as Code, and self-service platform capabilities • Define observability standards covering monitoring, logging, distributed tracing, synthetic monitoring, and operational intelligence • Lead incident response, post-incident reviews, root cause analyses, and continuous improvement • Eliminate single points of failure through proactive reliability engineering and resilient architecture • Implement automated remediation, self-healing capabilities, and proactive alerting • Lead capacity planning, performance engineering, availability management, and scalability assessments • Optimize cloud resource utilization with FinOps and engineering teams • Evaluate AIOps, Generative AI, and intelligent automation for platform reliability • Mentor SREs, platform engineers, and software engineers • Develop executive-level reliability roadmaps, operational strategies, and platform investment recommendations • Advise executive leadership on reliability strategy, operational risk, and service resiliency • Provide technical leadership during major incidents, outages, and high-severity escalations • Support PCI-DSS, SOC 2, security governance, and operational resilience compliance initiatives • Contribute to cloud transformation, platform engineering, and application modernization programs • Provide technical leadership and operational oversight for outsourced/offshore incident management resources • Manage vendor performance, incident response execution, operational effectiveness, and adherence to service-level objectives

🎯 Requirements

• Bachelor’s degree in computer science, Engineering, Information Technology, or a related discipline • 10+ years of experience in software engineering, cloud architecture, infrastructure engineering, or enterprise architecture • 5+ years of hands-on AWS architecture and cloud transformation experience • 5+ years of leadership or management experience • Proven success leading large-scale cloud migrations and modernization initiatives • Experience designing and supporting highly available, mission-critical, customer-facing platforms • Deep understanding of Kubernetes, container orchestration, microservices, and distributed systems • Extensive experience with DevSecOps, CI/CD pipelines, Infrastructure as Code, and automation frameworks • Strong knowledge of Site Reliability Engineering (SRE), operational excellence, and platform reliability practices • Experience implementing cloud governance, FinOps, and cost optimization programs • Preferred: experience supporting large-scale enterprise, eCommerce, aviation, travel, SaaS, or high-volume digital platforms • Preferred: experience building internal developer platforms and platform engineering capabilities • AWS Professional and/or Kubernetes certifications preferred • Experience with AIOps, intelligent automation, and reliability analytics preferred • Knowledge of cloud networking and security, high availability and disaster recovery, performance engineering, capacity planning, observability, incident response, root cause analysis, CI/CD automation, configuration management, and reliability automation

🏖️ Benefits

• Medical, dental and vision coverage • 401(k) retirement savings options • Paid holidays, vacation time and sick time • Travel privileges on Frontier Airlines and participating partner airlines • Buddy passes • Travel-related discounts and employee discounts on select products, services and vendors • A hybrid schedule for eligible headquarters roles based in Denver, Colorado • Business casual dress options for eligible corporate and support roles • Employee support programs and resources, including the HOPE League

Apply Now

Similar Jobs

🔥 3 hours ago

DaVita Kidney Care

10,000+ employees

🏥 Healthcare

⚕️ Healthcare Insurance

Senior DevOps Engineer building FinOps automation, governance, and cloud cost optimization for DaVita’s healthcare operations. Designing Kubernetes, billing-data, and infrastructure-as-code solutions.

🔥 5 hours ago

Inizio Evoke

1001 - 5000

🏥 Healthcare

💼 Consulting

📣 Marketing

Director leading AWS infrastructure, DevOps, security, and SaaS operations for Inizio Evoke, a healthcare-focused creative agency. Driving reliable, scalable, and cost-efficient technology delivery.

🔥 6 hours ago

Salesforce

10,000+ employees

💼 Consulting

📣 Marketing

☁️ SaaS

DevOps Engineer/Architect designing Salesforce deployment and CI/CD solutions for enterprise Salesforce implementations. Advising clients, governing releases, and enabling AI-first DevOps workflows.

🔥 6 hours ago

Branch

201 - 500

💳 Fintech

👥 HR Tech

🤝 B2B

Staff Cloud Operations Engineer managing secure, scalable GCP infrastructure for Branch’s financial-services platform. Automating operations, monitoring systems, and leading incident response.

🔥 7 hours ago

RELX

10,000+ employees

💼 Consulting

🏥 Healthcare

🛡️ Insurance

Senior Site Reliability Engineer improving AWS platform reliability, automation, and observability for Elsevier’s scientific and medical information services. Deploying secure AI-powered capabilities and supporting engineering teams.