Senior Site Reliability Engineer – SRE

Stelle nicht auf LinkedIn

🕒 vor 5 Monaten

🌽 Illinois – Remote

infoinfo

💵 $165.000 - $225.000 / Jahr

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

👻 Geisterscore 54%

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of Moonlite

Moonlite

1 - 10 Mitarbeiter

📚 Bildung

🏪 Marktplatz

👥 B2C

Education • Marketplace • B2C

Moonlite ist eine gemeinschaftsorientierte Webplattform, die Menschen dabei hilft, bewährte Wege zur Geldverdienung zu entdecken, zu vergleichen und zu vertrauen. Sie kuratiert tausende von Geschäftsideen, Ressourcen, Erstellern und Kursen, die alle von echten Nutzern validiert und bewertet wurden, sodass Suchende den Hype vermeiden und sich auf das konzentrieren können, was funktioniert. Moonlite bietet Community-Diskussionen, Bewertungen, Vergleiche nebeneinander und eine schnelle Umfrage, um Nutzer mit passenden Einkommenswegen abzustimmen, mit dem Ziel, Einzelpersonen zu helfen, mit Zuversicht finanzielle Freiheit zu erlangen.

Beschreibung

• Design, build, and operate production Kubernetes clusters on bare-metal infrastructure. • Implement and operate custom Kubernetes networking solutions. • Develop and maintain custom Kubernetes operators and controllers. • Deploy and optimize NVIDIA GPU operators and custom scheduling logic for GPU workloads. • Build deep integrations between Kubernetes and underlying infrastructure. • Design and implement automation using Terraform, Ansible, Helm, and custom operators. • Manage production bare-metal infrastructure across multiple regions ensuring high availability, fault tolerance, and graceful degradation. • Build comprehensive monitoring, logging, and alerting using Prometheus, Grafana, and ELK stack. • Identify and resolve performance bottlenecks across infrastructure domains.

🎯 Anforderungen

• 5+ years in SRE, DevOps, or infrastructure engineering roles with proven experience operating production infrastructure at scale. • Deep hands-on experience building and operating production Kubernetes clusters on bare-metal infrastructure. • Strong understanding of Kubernetes internals including custom resource definitions (CRDs), operators, controllers, admission webhooks, and scheduling. • Strong fundamentals in Linux systems administration, performance tuning, troubleshooting, and automation in production environments. • Proficiency with infrastructure-as-code tools (Terraform, Ansible, Helm) and building automation to reduce operational overhead. • Solid understanding of networking concepts including IPAM, DNS, DHCP, VLAN/VXLAN, routing, load balancing, and experience troubleshooting network issues in production. • Experience building and maintaining comprehensive monitoring solutions using tools like Prometheus, Grafana, and centralized logging systems. • Understanding of SRE principles including SLIs/SLOs/SLAs, error budgets, incident management, and blameless postmortems. • Strong scripting skills in Go, Python, or Bash for automation, tooling development, and operational efficiency. • Demonstrated ability to troubleshoot complex issues under pressure, manage incidents effectively, and communicate clearly during outages. • Excellent communication skills and ability to work across teams including systems engineers, network engineers, and software developers.

🏖️ Vorteile

• 6% 401(k) match • Fully covered health insurance premiums • Other comprehensive offerings to support your well-being and success as we grow together.

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 5 Monaten

Vytwo Technologies Inc

201 - 500

💼 Beratung

📦 Logistik

🎯 Rekrutierung

Meanstack Architect with DevOps expertise for TCoE, designing scalable applications and leading technical teams in a fully remote environment.

🇺🇸 Vereinigte Staaten – Remote

💵 $45 - $50 / Stunde

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 5 Monaten

Panopto

51 - 200

☁️ SaaS

📚 Bildung

🏢 Unternehmen

Mid-Level DevOps Engineer at Panopto transforming outdated build processes into automated pipelines. Elevate the engineering experience by enhancing delivery lifecycle and collaboration.

🇺🇸 Vereinigte Staaten – Remote

💵 $155.000 - $175.000 / Jahr

💰 Private Equity Round im 2021-04

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🦅 H1B-Visum-Sponsor

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 5 Monaten

Sword Health

201 - 500

🏥 Gesundheitswesen

💼 Beratung

📦 Logistik

DevOps Engineer at Sword Health designing scalable infrastructure and automating processes to enhance AI healthcare solutions. Collaborating with cross-functional teams in a remote work environment.

🇺🇸 Vereinigte Staaten – Remote

💵 $140.000 - $220.000 / Jahr

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 5 Monaten

Resolve Tech Solutions

501 - 1000

💼 Beratung

🏥 Gesundheitswesen

📦 Logistik

DevOps Lead Engineer at RTS responsible for scalable cloud infrastructure design and CI/CD pipeline optimization. Collaborating across teams to drive automation, governance, and cost optimization.

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 5 Monaten

Keeper Security, Inc.

501 - 1000

🔒 Cybersecurity

☁️ SaaS

🏢 Unternehmen

Senior DevOps Engineer managing IL5-compliant infrastructure for Keeper Security, working in high-security environments and collaborating with various engineering teams.

🇺🇸 Vereinigte Staaten – Remote

💰 Private Equity Round - Keeper Security im 2023-05

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich