Senior Site Reliability Engineer

🕒 Agosto 4

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $165.000 - $185.000 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of The Access Group

The Access Group

5001 - 10000 funcionários

💼 Consultoria

🏥 Saúde

🏨 Hospitalidade

Consulting • Healthcare • Hospitality

O Access Group é um provedor de software e serviços de gestão empresarial em nuvem, focado em setores industriais. A empresa oferece soluções SaaS modulares — incluindo finanças e contabilidade, RH e folha de pagamento, aprendizagem e conformidade, CRM, ERP, pagamentos e TI gerenciada — adaptadas a setores como instituições de caridade, educação, construção, saúde, hospitalidade, recrutamento, armazenagem e atacado. A companhia atende outras organizações com ferramentas integradas de nível empresarial, suportadas por serviços profissionais, sucesso do cliente e operações globais para ajudar seus clientes a otimizar operações e atender às necessidades regulatórias e específicas de cada setor.

Descrição

• Serve as the senior escalation point for complex P1/P2 production incidents, owning cross-system triage and permanent architectural remediation • Lead platform-level architecture reviews for reliability, scalability, security, and operational standards • Identify systemic failure patterns and translate them into architectural changes, design standards, and platform improvements • Own availability, reliability, performance, and scalability of production systems • Define, track, and improve SLOs, SLIs, and operational KPIs • Develop and maintain Terraform infrastructure-as-code solutions, including modules, state management, and governance • Eliminate operational toil through automation, self-service capabilities, and platform tooling • Build and maintain automation frameworks using Bash, PowerShell, and related scripting technologies • Administer and architect Microsoft Azure solutions, with AWS as a secondary platform • Operate Kubernetes in production, including cluster management and platform maintenance • Manage hybrid-cloud environments, virtual machines, networking, and distributed infrastructure • Maintain Datadog observability and PagerDuty alerting configurations • Design infrastructure controls for PCI-DSS, SOC 1/2, and ISO 27001 compliance and support audits • Partner with Engineering, Product, Security, and Operations on CI/CD, releases, and DevOps maturity • Mentor engineers and influence organizational infrastructure design standards

🎯 Requisitos

• 8+ years in Site Reliability Engineering, Platform Engineering, or Infrastructure Engineering • Direct ownership of complex production platforms at scale • Senior technical escalation experience for cross-team incidents and architectural remediation • Expert Microsoft Azure infrastructure experience and working knowledge of AWS • Deep Kubernetes production expertise • Advanced Terraform and Infrastructure-as-Code skills, including module design, state management, and governance • Strong Bash scripting and operational automation development • Networking fundamentals including firewalls, DNS, routing, VPN, troubleshooting, and Cloudflare edge services • Active Directory administration and hybrid identity experience • CI/CD pipeline design and deployment workflow improvement experience • Datadog, PagerDuty, or equivalent observability and alerting experience • Experience in compliance-regulated environments, including PCI-DSS, SOC 1/2, and ISO 27001 • Ability to influence across organizational boundaries without direct authority • Applicants must reside within the Eastern or Central time zones • Authorization to work in the U.S. without employer sponsorship is required • Preferred: Puppet or equivalent configuration management administration • Preferred: Microsoft SQL Server administration • Preferred: Meraki firewall policy management • Preferred: AI-driven operational workflows and Model Context Protocol (MCP) development • Preferred: Internal developer platform or platform engineering initiative leadership • Preferred: Large-scale SaaS or high-availability platform support • Preferred: Scala and/or Java infrastructure-level knowledge • Preferred: Azure Solutions Architect Expert, Azure Administrator Associate, AWS Solutions Architect, or CKA certification

🏖️ Benefícios

• 22 days paid time off • 11 company paid holidays • Medical insurance • Dental insurance • Vision insurance • 5% 401(k) company match • Range of other selectable benefits • Competitive salary • Development and career progression opportunities • Home-based/remote work arrangement

Candidatar-se

Vagas Similares

🕒 Agosto 4

AAA Life Insurance Company

501 - 1000

⚕️ Seguro de Saúde

💸 Finanças

🛡️ Seguros

Senior DevOps Engineer modernizing AAA Life’s cloud-native infrastructure and middleware. Designing CI/CD, automation, observability, security, and disaster recovery for life insurance systems.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 4

Net Health

501 - 1000

🏥 Saúde

☁️ SaaS

🤖 Inteligência Artificial

DevOps Engineer designing secure AWS platforms and CI/CD automation for Net Health’s healthcare SaaS products. Owning cloud architecture, database performance, observability, security, and cost optimization.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 4

Worth AI

11 - 50

💼 Consultoria

🛡️ Seguros

🤖 Inteligência Artificial

Senior DevOps Engineer improving Worth AI’s cloud infrastructure, Kubernetes platform, and deployment reliability. Automating infrastructure, strengthening observability, optimizing costs, and enabling engineering teams.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 3

CLARA Analytics

51 - 200

💼 Consultoria

🏥 Saúde

⚖️ Jurídico

DevOps Engineer at CLARA Analytics improving infrastructure-as-code practices in a fully remote environment. Collaborating with cross-functional teams and automating workflows for an AI-powered analytics platform.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 2

Defcon AI

11 - 50

🤖 Inteligência Artificial

🚗 Transporte

📦 Logística

DevSecOps Lead building and operating AI program delivery platforms in government cloud environments. Leading teams to ensure secure, efficient CI/CD pipelines for government deployment.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $175.000 - $215.000 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório