Site Reliability Engineer

🕒 Junho 10

🌪️ Kansas – Remoto

info

💵 $109.800 - $183.000 / ano

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Veeam Software

Veeam Software

1001 - 5000 funcionários

Fundada em 2006

💼 Consultoria

📦 Logística

☁️ SaaS

💰 $500.000.000 Private Equity Round em 2019-01

Consulting • Logistics • SaaS

A Veeam Software é líder global em resiliência e proteção de dados, oferecendo software de proteção de dados autogerenciável para ambientes híbridos e multi-cloud. Sua Veeam Data Platform oferece soluções abrangentes para backup, recuperação e segurança de dados, com princípios de confiança zero e ferramentas impulsionadas por inteligência artificial para inteligência de dados. As ofertas da Veeam incluem serviços de backup e armazenamento seguros para plataformas como Microsoft 365, AWS e Google Cloud, suportando cargas de trabalho diversas, incluindo ambientes virtuais, físicos e SaaS. Com uma reputação de inovação e confiança do cliente, a Veeam atende a uma ampla gama de indústrias, garantindo resiliência de dados contra interrupções, como ataques de ransomware. Suas soluções permitem que as empresas alcancem liberdade de dados, armazenamento seguro e gestão eficiente, reforçando sua posição como um dos principais fornecedores de software de backup e recuperação empresarial em todo o mundo.

Descrição

• Get up to speed on VDC workloads, dependencies, and operational workflows by reading code, docs, and working with SMEs. • Write and maintain runbooks, incident guides, and operational documentation. • Support knowledge transfer and contribute to onboarding materials for the team. • Participate in incident response including triage, investigation, mitigation, and postmortems. • Help implement and maintain SLIs, SLOs, and error budgets defined by the team. • Identify reliability issues during incidents or reviews and propose concrete improvements. • Support high availability and fault tolerance work on Azure, including Azure Government. • Close monitoring gaps by implementing instrumentation, alerting, and dashboards based on team standards. • Contribute to toil reduction through automation and tooling improvements. • Participate in on-call rotations. • Work with IaC, CI/CD pipelines, and deployment tooling in compliance-restricted environments. • Support testing, canary deployments, and release validation workflows. • Implement changes to infrastructure and configuration following established patterns and review processes. • Work with engineering, security, compliance, and operations teams to execute on reliability improvements. • Communicate clearly about system behavior, risk, and status — in writing and in meetings. • Raise blockers and gaps proactively; don't wait for problems to escalate.

🎯 Requisitos

• 3+ years in Software Engineering, with at least 1 year in SRE, Platform Engineering, or DevOps working on cloud-hosted services. • Experience with cloud infrastructure on Azure or a comparable cloud provider. • Familiarity with regulated or compliance-oriented environments such as government (FedRAMP, CMMC), financial (PCI-DSS), or healthcare (HIPAA). You understand that compliance shapes what you can and can't do operationally. • Able to read and understand code well enough to investigate system behavior without always having someone walk you through it. • Experience with monitoring and observability tools (e.g., Prometheus, Grafana, OpenTelemetry, ELK stack). • Experience with IaC tools (Terraform, Terragrunt, or Pulumi) and container orchestration (Kubernetes). • Experience with CI/CD tooling such as GitHub Actions, Azure DevOps, GitLab CI, or ArgoCD. • Strong programming skills in one or more of: TypeScript/JS, Go, Java, C#, or similar. • Solid understanding of distributed systems fundamentals and networking basics. • Clear written and verbal communication skills.

🏖️ Benefícios

• Unlimited paid time off, 12 paid holidays including 4 global VeeaMe Days for self-care and 24 paid volunteer hours annually through Veeam Cares • Paid parental leave: 8 weeks for all parents, 16 weeks for birthing parents • Medical, dental, and vision coverage starting on your first day • Mental health support, therapy sessions, and digital wellness tools via our Employee Assistance Program • 401(k) retirement plan with company matching contributions • Fertility, adoption, and surrogacy support through Maven, plus paid volunteer time • AirVet: 24/7 virtual veterinary care at no cost • Legal services, identity protection, and supplemental health insurance options • Tax-advantaged spending accounts for healthcare, dependent care, and commuting • Opportunities to learn and grow through on-demand libraries (LinkedIn Learning, O’Reilly), mentoring, workshops, and learning events like our annual Global Day of Learning

Candidatar-se

Vagas Similares

🕒 Junho 10

Senior Site Reliability Engineer ensuring reliability and operational excellence at Priority Technology Holdings. Collaborating with product and infrastructure teams to enhance service resilience and observability.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $129.000 - $161.000 / ano

🔥 Investimento no último ano

💰 $50.000.000 Post-IPO Debt - Priority Commerce em 2025-08

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 10

Granicus

501 - 1000

🏛️ Governo

☁️ SaaS

📋 Conformidade

DevOps Engineer II automating CI/CD and improving platform stability at Granicus. Leveraging cloud and security best practices for effective technology solutions in the Govtech industry.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 9

ClickUp

1001 - 5000

☁️ SaaS

⚡ Produtividade

🏢 Corporativo

GTM DevOps Engineer at ClickUp responsible for reliability and automation of Go-To-Market technology stack. Collaborating with developers to build CI/CD pipelines and manage cloud infrastructure.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $160.000 - $210.000 / ano

💰 $400.000.000 Series C - ClickUp em 2021-10

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 9

DMI (Digital Management, LLC)

1001 - 5000

💼 Consultoria

🏥 Saúde

📦 Logística

Mid-level DevSecOps Engineer supporting hybrid cloud infrastructure for federal agency client. Focus on automation, security, and CI/CD practices.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 9

Coinbase

1001 - 5000

💼 Consultoria

₿ Cripto

💸 Finanças

Senior Site Reliability Engineer managing AI infrastructure at Coinbase. Driving automation, reliability, and observability in critical AI operations.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $186.065 - $218.900 / ano

💰 $21.400.000 Post-IPO Equity em 2022-11

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório