Senior Site Reliability Engineer, Infrastructure

🕒 Maio 29

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $125.000 - $135.000 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Vultr

Vultr

201 - 500 funcionários

Fundada em 2014

🤖 Inteligência Artificial

🤝 B2B

🔧 Hardware

💰 $329.000.000 Debt Financing - Vultr em 2025-06

Artificial Intelligence • B2B • Hardware

A Vultr é um provedor global de infraestrutura em nuvem que oferece máquinas virtuais sob demanda, servidores bare-metal, instâncias aceleradas por GPU, bancos de dados gerenciados, armazenamento de objetos e em blocos, Kubernetes e serviços de rede. A plataforma enfatiza cargas de trabalho de IA e HPC com uma ampla seleção de GPUs AMD e NVIDIA, rede rápida e mais de 32 regiões de data centers, além de um marketplace de aplicativos implantáveis e APIs amigáveis para desenvolvedores. A Vultr tem como público-alvo desenvolvedores e empresas que buscam alternativas de computação e armazenamento em nuvem acessíveis, escaláveis e em conformidade com os regulamentos em relação aos hyperscalers.

Descrição

• Design and build the observability pipeline for datacenter infrastructure including CDUs, PDUs, bare metal servers, and provisioning workflows, collecting telemetry via Redfish, IPMI, SNMP, and OpenTelemetry. • Own the full stack from data collection through to visualization and alerting in Grafana, Loki, and Mimir. • Build dashboards and alerting that are actionable and meaningful for stakeholder teams including Datacenter Ops, SysAdmin, Network, and Provisioning. • Establish standards and patterns for how datacenter infrastructure telemetry is collected, stored, and visualized across Vultr's global footprint. • Partner closely with stakeholder teams to understand their operational needs and translate them into observable, measurable signals. • Drive infrastructure-as-code practices across the observability pipeline to ensure consistency, repeatability, and maintainability.

🎯 Requisitos

• 5+ years of experience in site reliability, platform, or infrastructure engineering in a production environment. • Hands-on experience building and operating observability pipelines including metrics, logs, and alerting using Grafana, Loki, Mimir, or equivalent tooling. • Working knowledge of datacenter hardware telemetry protocols including Redfish, IPMI, and/or SNMP. • Strong Linux fundamentals and operational experience in production infrastructure environments. • Demonstrated experience with infrastructure-as-code and configuration management tooling (Terraform, Ansible, Chef or similar). • Strong cross-functional communication skills and experience delivering tooling for operational stakeholder teams.

🏖️ Benefícios

• 100% company-paid insurance premiums for employee medical, dental and vision plans. • 401(k) plan that matches 100% up to 4%, with immediate vesting • Professional Development Reimbursement of $2,500 each year • 11 Holidays + Paid Time Off Accrual + Rollover Plan • Commitment matters to Vultr! Increased PTO at 3 year and 10 year anniversary + 1 month paid sabbatical every 5 years + Anniversary Bonus each year • $500 stipend for remote office setup in first year + $400 each following year • Internet reimbursement up to $75 per month • Gym membership reimbursement up to $50 per month • Company paid Wellable subscription

Candidatar-se

Vagas Similares

🕒 Maio 29

PVcase

201 - 500

⚡ Energia

☁️ SaaS

🏢 Corporativo

DevOps/Platform Engineer managing AWS infrastructure and ensuring application performance. Collaborating with global teams to enhance operational workflows for the PVcase Prospect SaaS application.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $126.600 - $180.000 / ano

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Maio 29

Innovative Solutions

51 - 200

💼 Consultoria

📣 Marketing

📦 Logística

DevOps Engineer designing scalable, secure AWS infrastructure while collaborating with multiple clients. Responsibilities include implementing CI/CD pipelines and ensuring system reliability.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Maio 29

Primordial Labs

11 - 50

🎖️ Defesa

📦 Logística

💼 Consultoria

Field Deployment Engineer at Primordial Labs ensuring Anura performs in military operations. Bridging engineering and field operations while providing technical support for autonomous systems deployment.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $120.000 - $200.000 / ano

💰 $4.000.000 Seed Round - Primordial Labs em 2022-10

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Maio 29

ITW

10.000+ funcionários

🚘 Automotivo

🏗️ Construção

🍽️ Alimentos e Bebidas

DevOps Release Engineer coordinating and automating the software release lifecycle at ITW. Collaborating across engineering, sales, and product management to ensure quality and timeliness of releases.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Maio 29

Capgemini

10.000+ funcionários

💼 Consultoria

🏥 Saúde

📦 Logística

Software Change Management Consultant supporting application migration projects utilizing IBM DBB/Git/IDD solutions. Leading technical training and troubleshooting in a remote capacity across North America.

🗣️🇺🇸🇬🇧 Inglês obrigatório

Groovy