Site Reliability Engineering Manager

🕒 vor 1 Monat

🐊 Florida – Remote

infoinfo

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🦅 H1B-Visum-Sponsor

infoinfo

👻 Geisterscore 12%

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of NationsBenefits

NationsBenefits

1001 - 5000 Mitarbeiter

Gegründet 2013

🏥 Gesundheitswesen

💼 Beratung

📦 Logistik

💰 Private Equity Round im 2022-04

Healthcare • Consulting • Logistics

NationsBenefits ist ein führendes Unternehmen im Bereich Gesundheitstechnologie, das sich auf die Bereitstellung von Finanztechnologielösungen und das Management von Zusatzleistungen spezialisiert hat. Das Unternehmen bietet eine Vielzahl von Dienstleistungen an, darunter Hörvorsorge über NationsHearing, Gesundheits- und Wellnessprodukte über NationsOTC und Essenslieferungen durch NationsMarket. NationsBenefits bietet auch spezialisierte Dienstleistungen wie Notfallhilfe, Gesundheitsversorgungstransporte und personalisierte Gesundheitsberatung unter Einsatz von künstlicher Intelligenz an. Ihre eigenen Plattformen, einschließlich Benefits Pro™, unterstützen die Mitgliederbetreuung, die Konfiguration von Leistungen und E-Commerce-Transaktionen. NationsBenefits konzentriert sich darauf, die Ergebnisse für Mitglieder zu verbessern, Versorgungslücken zu schließen und die Zufriedenheit durch fortschrittliche Analysen und maßgeschneiderte Programme zu steigern.

Beschreibung

• Lead, mentor, and develop a US-based team of Site Reliability Engineers • Conduct regular 1:1s, performance reviews, and career development discussions • Own hiring, onboarding, and retention efforts as the team scales • Foster a culture of ownership, blameless postmortems, and continuous improvement • Lead day-to-day production operations and ensure timely incident triage, resolution, and escalation • Serve as an escalation point and incident commander for major production incidents • Drive problem management and root cause analysis processes • Carry PagerDuty on-call escalation responsibilities for critical issues • Track and report operational KPIs, SLAs, and SLOs, including availability, MTTR, and incident trends • Improve system reliability, observability, and resilience using Datadog and related tooling • Drive automation, self-healing capabilities, and runbook maturity • Partner with Development, DevOps, DevSecOps, and Engineering teams to embed reliability into the SDLC • Contribute hands-on to tooling, automation, and technical reviews as needed • Coordinate closely with SRE leadership in India to ensure seamless follow-the-sun coverage • Represent the US SRE organization in cross-functional planning and operational reviews • Communicate effectively with both technical and non-technical stakeholders • Maintain high-quality documentation for incidents, postmortems, runbooks, and operational procedures • Ensure adherence to healthcare and fintech compliance standards, including HIPAA, PCI DSS, SOC 2, ISO 27001, and HITRUST

🎯 Anforderungen

• 5–8 years of experience in Site Reliability Engineering, DevOps, Production Support, or Platform Engineering • 1–2+ years of experience leading, mentoring, or managing engineers • Demonstrated success operating in a player-coach leadership model • Strong hands-on experience with production incident management and escalation processes • Proficiency with Datadog or similar observability platforms • Hands-on experience with Kubernetes and Docker in production environments • Strong scripting or programming skills in PowerShell, Bash, Python, Java, or C# • Experience with Helm, CI/CD pipelines, and deployment automation • Working knowledge of ITIL processes and Agile methodologies • Experience working with SQL, MySQL, or NoSQL databases • Excellent communication and stakeholder management skills • Willingness to participate in PagerDuty on-call escalation and work within a global follow-the-sun operating model

🏖️ Vorteile

• Competitive compensation and comprehensive benefits • Unlimited PTO • Fully remote work environment (US-based) • Opportunity to lead and grow a high-impact SRE organization • Exposure to modern cloud-native technologies and large-scale reliability challenges • Collaborative culture focused on innovation, learning, and continuous improvement • Meaningful work that directly impacts healthcare technology and millions of members

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 1 Monat

C5MI

201 - 500

💼 Beratung

🏢 Unternehmen

📦 Logistik

Azure DevOps Administrator managing ADO environment and coaching delivery teams to align with SDLC standards. Ensuring configuration and governance for C5MI's Azure DevOps and SDLC compliance.

🇺🇸 Vereinigte Staaten – Remote

💵 $100.000 - $120.000 / Jahr

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

Smithfield Foods

10.000+ Mitarbeiter

🏭 Fertigung

🌾 Landwirtschaft

🍽️ Lebensmittel & Getränke

Sr. Utilities Engineer optimizing and managing utility systems at Smithfield Foods. Focusing on industrial refrigeration, boiler, and compressed air systems for manufacturing processes.

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

Koniag Government Services

1001 - 5000

🏛️ Regierung

🎖️ Verteidigung

💼 Beratung

DevOps Engineer supporting cloud modernization and cybersecurity efforts for Koniag Data Solutions. Collaborating with cross-functional teams to enhance cloud delivery and continuous authorization processes.

🇺🇸 Vereinigte Staaten – Remote

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

Dropzone AI

51 - 200

🤖 Künstliche Intelligenz

Senior DevOps Engineer at Dropzone AI enhancing infrastructure for our AI cybersecurity platform. Collaborating with engineering teams on scalable, resilient, and secure systems.

🇺🇸 Vereinigte Staaten – Remote

💵 $170.000 - $185.000 / Jahr

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

Golden 1 Credit Union

1001 - 5000

🏦 Bankwesen

💸 Finanzen

DevOps Engineer responsible for leading automation processes and managing infrastructure in Azure and On-Premises environments. Collaborating closely with development teams for reliable deployments and CI/CD pipelines.

🇺🇸 Vereinigte Staaten – Remote

💵 $123.600 - $135.000 / Jahr

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich