Principal Site Reliability Engineer, SRE

Stelle nicht auf LinkedIn

🕒 vor 2 Monaten

🇺🇸 Vereinigte Staaten – Remote

⏰ Vollzeit

🔴 Experte

⛑ DevOps- und Site Reliability Engineer (SRE)

👻 Geisterscore 37%

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of SoluStaff

SoluStaff

51 - 200 Mitarbeiter

🏥 Gesundheitswesen

🎖️ Verteidigung

🏭 Fertigung

Healthcare • Defense • Manufacturing

Symmetrio ist ein Full-Service-Unternehmen für Recruiting, Personaldienstleistungen und Beratung mit jahrzehntelanger Erfahrung in verschiedenen wachstumsstarken Branchen. Sie spezialisieren sich auf die Vermittlung von Festangestellten, Übernahme von Zeitarbeitskräften und Personalverstärkung, wobei der Fokus auf der Abstimmung von Talenten mit den Zielen und der Kultur der Kunden liegt. Symmetrio bietet maßgeschneiderte Rekrutierungslösungen und Beratungsdienste in Bereichen wie Life Sciences, Informationstechnologie, Ingenieurwesen, Medizintechnik, Logistiklösungen, Fertigung und Gebäudeautomation. Ihr Team von Experten im Bereich Talentakquise engagiert sich dafür, die organisatorischen Bedürfnisse der Kunden zu verstehen und spezialisierte Fachkräfte bereitzustellen, um diesen Herausforderungen zu begegnen. Mit einem starken Fokus auf Vertrauen, Verständnis und Zusammenarbeit zielt Symmetrio darauf ab, die Abläufe zu optimieren und Innovation und Exzellenz für ihre Kunden voranzutreiben.

Beschreibung

• Serve as the primary technical owner for production reliability across U.S. customer environments. • Investigate and resolve complex issues spanning web applications, APIs, backend services, data pipelines, cloud infrastructure, and customer integrations. • Lead production incident response efforts, coordinating cross-functional teams to restore service and minimize customer impact. • Perform root cause analysis and drive corrective actions that improve long-term system stability and resilience. • Partner with software engineering and platform teams to identify recurring reliability risks and implement sustainable solutions. • Design, configure, and validate secure customer connectivity solutions including Site-to-Site VPNs, Transit Gateway integrations, routing configurations, and secure network paths. • Support customer onboarding initiatives by troubleshooting connectivity challenges and ensuring consistent implementation processes. • Enhance platform observability through improvements in monitoring, logging, alerting, tracing, and operational dashboards. • Contribute to CI/CD, infrastructure automation, and deployment processes that improve release safety and operational consistency. • Develop operational tooling that supports incident response, troubleshooting, onboarding, and system monitoring activities. • Collaborate with engineering leadership to improve cloud architecture, scalability, security, and operational readiness. • Partner with customer-facing teams to communicate technical issues, remediation plans, and reliability improvements in a clear and effective manner. • Support compliance, security, and risk management initiatives within highly regulated healthcare environments.

🎯 Anforderungen

• 6+ years of hands-on experience supporting and managing AWS-based production environments. • 4+ years of experience supporting web applications and backend services (Python/Django experience strongly preferred). • Experience with AWS networking technologies including VPCs, Site-to-Site VPNs, Transit Gateways, routing, NAT gateways, and security groups. • Strong experience with Terraform and infrastructure-as-code deployment practices. • Experience with containerized environments including ECS, Fargate, Kubernetes, or similar technologies. • Experience building and supporting CI/CD pipelines and release automation processes. • Familiarity with monitoring and observability platforms such as Datadog, CloudWatch, Sentry, Grafana, or similar tools. • Experience leading production incidents, outage management, and root cause analysis initiatives. • Exposure to Windows Server environments, Active Directory, Kerberos, and enterprise infrastructure concepts is preferred. • Healthcare technology, healthcare SaaS, clinical software, or other regulated industry experience is highly preferred. • Bachelor’s degree in Computer Science, Engineering, Information Technology, or a related technical field preferred.

🏖️ Vorteile

• Health Care Plan (Medical, Dental & Vision) • Retirement Plan (401k, IRA) • Paid Time Off (Vacation, Sick & Public Holidays)

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 2 Monaten

Coinbase

1001 - 5000

💼 Beratung

₿ Crypto

💸 Finanzen

Staff Site Reliability Engineer driving AI transformation by ensuring reliability and automation at Coinbase. Collaborating with infrastructure teams and leading critical incident responses to maintain service excellence.

🇺🇸 Vereinigte Staaten – Remote

💵 $218.025 - $256.500 / Jahr

💰 €21.400.000 Post-IPO Equity im 2022-11

⏰ Vollzeit

🔴 Experte

⛑ DevOps- und Site Reliability Engineer (SRE)

🦅 H1B-Visum-Sponsor

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 3 Monaten

Agilent Technologies

10.000+ Mitarbeiter

🍽️ Lebensmittel & Getränke

🏥 Gesundheitswesen

💼 Beratung

DevOps Software Engineer designing and maintaining CI/CD pipelines and cloud infrastructure for Agilent’s CrossLab Connect team. Supporting application development and optimizing deployment processes.

🇺🇸 Vereinigte Staaten – Remote

💵 $143.760 - $224.625 / Jahr

💰 €500.000.000 Post-IPO Debt im 2019-09

⏰ Vollzeit

🟠 Senior

🔴 Experte

⛑ DevOps- und Site Reliability Engineer (SRE)

🦅 H1B-Visum-Sponsor

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 3 Monaten

Ad Hoc LLC

501 - 1000

💼 Beratung

🏥 Gesundheitswesen

📦 Logistik

Staff DevOps Engineer at Ad Hoc, leading technical solutions and improving software engineering processes. Expert in CI/CD and infrastructure, with a focus on federal service delivery.

🇺🇸 Vereinigte Staaten – Remote

💵 $130.000 - $150.000 / Jahr

⏰ Vollzeit

🔴 Experte

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 3 Monaten

Capgemini

10.000+ Mitarbeiter

💼 Beratung

🏥 Gesundheitswesen

📦 Logistik

Mainframe DevOps Migration Consultant at Capgemini Engineering supporting application migration projects utilizing client’s DBB/Git/IDD Solutions.

🗣️🇺🇸🇬🇧 Englisch erforderlich

Groovy

🕒 vor 3 Monaten

Capgemini

10.000+ Mitarbeiter

💼 Beratung

🏥 Gesundheitswesen

📦 Logistik

Software Change Management Consultant supporting application migration projects utilizing IBM DBB/Git/IDD solutions. Leading technical training and troubleshooting in a remote capacity across North America.

🗣️🇺🇸🇬🇧 Englisch erforderlich

Groovy