Lead Site Reliability Engineer

🕒 vor 2 Monaten

🇺🇸 Vereinigte Staaten – Remote

💵 $114.000 - $165.300 / Jahr

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🦅 H1B-Visum-Sponsor

infoinfo

👻 Geisterscore 21%

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of Empower

Empower

10.000+ Mitarbeiter

💸 Finanzen

💳 Fintech

👥 B2C

Finance • Fintech • B2C

Empower ist ein führender Anbieter von Finanzdienstleistungen, der sich darauf konzentriert, Einzelpersonen und Organisationen dabei zu helfen, finanzielle Freiheit durch Altersvorsorge und Investmentmanagement zu erreichen. Mit über 19 Millionen Kunden in den USA bietet Empower eine umfassende Palette an finanzbezogenen Dienstleistungen an, darunter intelligente Planung und Anlageberatung sowie Tools wie das Empower Personal Dashboard™ für einen vollständigen Finanzüberblick. Das Unternehmen ist bekannt als führender Anbieter von Altersvorsorgeplänen und arbeitet eng mit privaten Anlegern, Arbeitsplatzsparer, Plan-Sponsoren und Finanzfachleuten zusammen. Empower wird auch für Initiativen in den Bereichen Diversity, Equity, Inclusion anerkannt und hat ein soziales Engagement, das die Wirkung in der Gemeinschaft stärkt.

Beschreibung

• Lead cross-functional reliability initiatives across multiple value streams and coordinate execution across teams. • Define and evolve SRE best practices, tools, and methodologies across the organization. • Architect enterprise-scale, multi-region AWS infrastructure that balances reliability, cost, performance, and security. • Establish and operate SLOs, SLIs, and error budgets for critical services, using them to drive prioritization decisions. • Serve as incident commander for major incidents and drive postmortems that produce completed action items and organizational learning. • Lead disaster recovery planning for critical financial services infrastructure. • Build shared Infrastructure as Code foundations in Terraform (reusable modules, standards, and patterns adopted across teams). • Design and implement production-scale Kubernetes patterns, including multi-tenancy, security policies, and advanced scheduling. • Establish observability standards and strategies using Datadog and Splunk (metrics, logging, tracing, dashboards, and alerting). • Set CI/CD standards and patterns, including pipeline-as-code and progressive delivery at scale. • Lead chaos engineering, game days, and systematic reliability testing initiatives. • Drive FinOps initiatives to optimize cloud spend while maintaining reliability targets. • Lead a functional team of SREs (without direct reports) on projects and operational initiatives. • Mentor SREs at multiple levels through coaching, design reviews, code reviews, and training sessions. • Partner with Engineering, Product, and Security leadership to align reliability work with business priorities, zero-trust architecture, and compliance controls.

🎯 Anforderungen

• Bachelor’s degree in Computer Science, Information Technology, or related field (or equivalent practical experience) • 7 to 10 years of Site Reliability Engineering experience (or equivalent), with demonstrated technical leadership • Proven ability to lead technical teams and drive complex projects to completion • Expert AWS knowledge, including designing large-scale, multi-region architectures • Deep Kubernetes expertise, including advanced features, security, and production-scale operations • Mastery of Infrastructure as Code using Terraform, including building shared platforms and frameworks • Strong software engineering background with production experience in Python and/or Go • Extensive experience with observability platforms (Datadog, Splunk) and implementing monitoring at scale • Deep understanding of CI/CD principles and experience implementing enterprise-grade pipelines • Proven track record leading major incidents and conducting effective postmortems • Strong understanding of security, networking, and infrastructure design patterns • Strong communication skills with ability to explain complex technical concepts to diverse audiences • Experience mentoring engineers and building technical capabilities in teams.

🏖️ Vorteile

• Medical, dental, vision and life insurance • Retirement savings – 401(k) plan with generous company matching contributions (up to 6%), financial advisory services, potential company discretionary contribution, and a broad investment lineup • Tuition reimbursement up to $5,250/year • Business-casual environment that includes the option to wear jeans • Generous paid time off upon hire – including a paid time off program plus ten paid company holidays and three floating holidays each calendar year • Paid volunteer time — 16 hours per calendar year • Leave of absence programs – including paid parental leave, paid short- and long-term disability, and Family and Medical Leave (FMLA) • Business Resource Groups (BRGs) – BRGs facilitate inclusion and collaboration across our business internally and throughout the communities where we live, work and play. BRGs are open to all.

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 2 Monaten

Infarsight

51 - 200

🤖 Künstliche Intelligenz

✈️ Reisen

📦 Logistik

Senior DevOps & Cloud Infrastructure Engineer optimizing AWS environments for automation and product innovation. Leading deployment strategies and resource management across AWS, Vercel, and RackSpace.

🇺🇸 Vereinigte Staaten – Remote

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 2 Monaten

Vytalize Health

201 - 500

🏥 Gesundheitswesen

☁️ SaaS

⚕️ Krankenversicherung

DevSecOps Engineer responsible for implementing secure development and cloud security practices. Leading vulnerability management and integrating security within engineering and IT workflows.

🇺🇸 Vereinigte Staaten – Remote

💰 €100.000.000 Series C - Vytalize Health im 2023-02

⏰ Vollzeit

🟠 Senior

🔴 Experte

⛑ DevOps- und Site Reliability Engineer (SRE)

🦅 H1B-Visum-Sponsor

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 2 Monaten

Prominent Edge

11 - 50

💼 Beratung

🎖️ Verteidigung

📦 Logistik

Lead DevOps engineer working at Prominent Edge on scalable solutions using AWS technologies. Engage in diverse projects, automate deployments, and enjoy a supportive work environment.

🇺🇸 Vereinigte Staaten – Remote

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 2 Monaten

Planned Systems International

1001 - 5000

💼 Beratung

🎖️ Verteidigung

📦 Logistik

Site Reliability Engineer for PSI ensuring reliability, security, and performance in Federal Government IT solutions. Collaborating with teams to enhance DevOps practices and system health.

🇺🇸 Vereinigte Staaten – Remote

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 2 Monaten

EITACIES Inc.

51 - 200

💼 Beratung

🏥 Gesundheitswesen

🏭 Fertigung

Linux Site Reliability Engineer supporting large-scale hybrid infrastructure environments across cloud and private data centers. Requires deep Linux administration and automation skills.

🇺🇸 Vereinigte Staaten – Remote

💵 $58 - $64 / Stunde

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich