Site Reliability Engineer

🕒 vor 1 Monat

🏄 California – Remote

infoinfo

💵 $138.900 - $231.400 / Jahr

⏰ Vollzeit

🟠 Senior

🔴 Experte

⛑ DevOps- und Site Reliability Engineer (SRE)

🦅 H1B-Visum-Sponsor

infoinfo

👻 Geisterscore 1%

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of Veeam Software

Veeam Software

1001 - 5000 Mitarbeiter

Gegründet 2006

💼 Beratung

📦 Logistik

☁️ SaaS

💰 €500.000.000 Private Equity Round im 2019-01

Consulting • Logistics • SaaS

Veeam Software ist ein globaler Marktführer für Datenresilienz und -schutz und bietet selbstverwaltete Datensicherungssoftware für Hybrid- und Multi-Cloud-Umgebungen. Die Veeam Data Platform stellt umfassende Lösungen für Datensicherung, Wiederherstellung und Sicherheit bereit und setzt auf Zero-Trust-Prinzipien sowie KI-gestützte Tools für Data Intelligence. Das Angebot von Veeam umfasst sichere Backup- und Storage-Services für Plattformen wie Microsoft 365, AWS und Google Cloud und unterstützt unterschiedlichste Workloads in virtuellen, physischen und SaaS-Umgebungen. Mit einem starken Ruf für Innovation und hohem Kundenvertrauen bedient Veeam ein breites Spektrum an Branchen und sorgt für Datenresilienz gegenüber Störungen wie Ransomware-Angriffen. Die Lösungen ermöglichen Unternehmen Datenfreiheit, sichere Speicherung und effizientes Management und untermauern die Position als ein führender Anbieter von Enterprise-Backup- und -Wiederherstellungssoftware weltweit.

Beschreibung

• Support the Veeam Data Cloud SaaS platform’s Government and Sovereign Cloud environment • Map platform systems, workloads, dependencies, and risk areas • Work with subject-matter experts to fill knowledge gaps and build onboarding material • Write and maintain runbooks, architecture documentation, and operational guides • Design highly available and fault-tolerant infrastructure on Azure, including Azure Government • Define SLIs, SLOs, and error budgets • Lead incident response and blameless postmortems, converting incidents into improvements • Identify reliability risks and develop remediation plans within compliance constraints • Define observability instrumentation requirements and drive implementation • Establish alerting, telemetry, and monitoring standards • Build automation to reduce toil and support fleet management • Participate in on-call rotations • Work with infrastructure as code, CI/CD, deployment automation, and configuration management in air-gapped or compliance-restricted environments • Build and maintain testing, canary deployment, and release validation pipelines • Integrate chaos engineering and monitoring tools • Collaborate across product, platform, security, legal, compliance, and operations teams • Own reliability problems end-to-end and drive solutions • Mentor engineers and spread SRE practices across the organization

🎯 Anforderungen

• 7+ years in Software Engineering, including 3+ years in SRE, Platform Engineering, or similar roles • Experience across multi-service platforms • Experience with Government or Sovereign Cloud, such as Azure Government or AWS GovCloud • Experience in regulated compliance environments, including FedRAMP, CMMC, IL2/IL4/IL5, PCI-DSS, SOX, HIPAA, or HITRUST • Strong experience building and running production services on cloud infrastructure; Azure preferred, including Azure Government • Ability to learn large, complex platforms quickly with limited guidance and restricted environment access • Ability to independently investigate systems and produce clear documentation, risk assessments, and improvement plans • Experience with programming in TypeScript/JavaScript, Go, Java, C#, or similar • Experience with monitoring and observability tools such as Prometheus, Grafana, OpenTelemetry, or ELK Stack • Experience with infrastructure as code, including Terraform, Terragrunt, or Pulumi • Experience with container orchestration, especially Kubernetes • Experience with CI/CD and GitOps tooling, including GitHub Actions, Azure DevOps, GitLab CI, ArgoCD, FluxCD, or Dagger • Strong understanding of distributed systems, networking, and cloud-native architecture • Clear written and verbal communication skills

🏖️ Vorteile

• Unlimited paid time off • 12 paid holidays, including 4 global VeeaMe Days for self-care • 24 paid volunteer hours annually through Veeam Cares • Paid parental leave: 8 weeks for all parents, 16 weeks for birthing parents • Medical, dental, and vision coverage starting on the first day • Mental health support, therapy sessions, and digital wellness tools via the Employee Assistance Program • 401(k) retirement plan with company matching contributions • Fertility, adoption, and surrogacy support through Maven • AirVet: 24/7 virtual veterinary care at no cost • Legal services, identity protection, and supplemental health insurance options • Tax-advantaged spending accounts for healthcare, dependent care, and commuting • On-demand learning libraries, mentoring, workshops, and learning events including the annual Global Day of Learning • Competitive compensation and benefits

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 1 Monat

Syniti

1001 - 5000

🤝 B2B

🏢 Unternehmen

Senior SRE automating Azure and AWS infrastructure for Syniti’s enterprise data platform. Supporting Kubernetes, CI/CD, observability, security, and compliance across global SaaS workloads.

🇺🇸 Vereinigte Staaten – Remote

💵 $134.941 - $171.411 / Jahr

💰 Private Equity Round im 2017-08

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🦅 H1B-Visum-Sponsor

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

Karat

201 - 500

👥 HR Tech

🏢 Unternehmen

☁️ SaaS

Senior Deployment Engineer helping Karat, a technical interviewing company, implement and optimize enterprise interview frameworks. Advising clients, analyzing hiring performance, and delivering executive training.

🇺🇸 Vereinigte Staaten – Remote

💵 $116.875 - $148.187 / Jahr

💰 Funding Round im 2022-04

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🦅 H1B-Visum-Sponsor

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

Net Health

501 - 1000

🏥 Gesundheitswesen

☁️ SaaS

🤖 Künstliche Intelligenz

DevOps Engineer designing secure AWS platforms and CI/CD automation for Net Health’s healthcare SaaS products. Owning cloud architecture, database performance, observability, security, and cost optimization.

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

CLARA Analytics

51 - 200

💼 Beratung

🏥 Gesundheitswesen

⚖️ Rechtswesen

DevOps Engineer at CLARA Analytics improving infrastructure-as-code practices in a fully remote environment. Collaborating with cross-functional teams and automating workflows for an AI-powered analytics platform.

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

Filevine

201 - 500

☁️ SaaS

⚖️ Rechtswesen

🤖 Künstliche Intelligenz

Staff Site Reliability Engineer at Filevine shaping engineering culture and driving reliability practices. Leading technical standards and mentorship within a remote engineering team focused on legal AI technology.

🇺🇸 Vereinigte Staaten – Remote

💵 $235.000 - $275.000 / Jahr

💰 €108.000.000 Series D im 2022-04

⏰ Vollzeit

🔴 Experte

⛑ DevOps- und Site Reliability Engineer (SRE)

🦅 H1B-Visum-Sponsor

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich