Senior Data Reliability Engineer, AWS

🕒 Agosto 1

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $105.700 - $149.275 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

infoinfo

👻 Score fantasma 39%

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Empower

Empower

10.000+ funcionários

💸 Finanças

💳 Fintech

👥 B2C

Finance • Fintech • B2C

A Empower é uma fornecedora líder de serviços financeiros focada em ajudar indivíduos e organizações a alcançar a liberdade financeira através do planejamento de aposentadoria e da gestão de investimentos. Atendendo a mais de 19 milhões de americanos, a Empower oferece um conjunto abrangente de serviços relacionados a finanças, incluindo planejamento inteligente e aconselhamento de investimento, além de ferramentas como o Empower Personal Dashboard™ para uma visão financeira completa. A empresa é renomada como uma das principais provedoras de planos de aposentadoria e trabalha de perto com investidores pessoais, poupadores de planos no local de trabalho, patrocinadores de planos e profissionais financeiros. A Empower também é reconhecida por suas iniciativas em Diversidade, Equidade, Inclusão e tem um compromisso social que fortalece o impacto na comunidade.

Descrição

• Own the reliability and stability of production data pipelines and data platform services • Diagnose and resolve data pipeline failures, delays, and data quality issues in production • Investigate distributed data systems, including Spark/EMR workloads, ingestion pipelines, and warehouse performance • Lead or support incident response through triage, mitigation, and long-term resolution • Perform root cause analysis and implement durable fixes • Define and improve data SLAs for freshness, latency, and completeness • Design and enhance monitoring, alerting, and observability for data systems • Develop automation and tooling to reduce operational toil and improve resilience • Contribute to disaster recovery and resiliency planning, including backup validation and recovery workflows • Partner with engineering teams to improve pipeline design, reliability, and operational readiness • Create and maintain runbooks, SOPs, and operational documentation • Participate in occasional off-hours support and on-call rotation for production data systems

🎯 Requisitos

• Minimum 5 years of experience working with production data platforms in AWS environments • Prior experience building data pipelines through production, including exposure to real-world failures and operational challenges • Strong experience with Python and SQL in real data systems • Hands-on experience troubleshooting distributed data processing systems, such as Spark/EMR, Redshift, and streaming systems • Proven ability to debug and resolve production issues in data pipelines and data platforms • Experience with AWS data services such as EMR, Redshift, DynamoDB, S3, or similar • Experience handling production incidents and performing root cause analysis • Strong problem-solving mindset and ability to work through ambiguous production issues • Must be authorized to work for any employer in the U.S.; employment visa sponsorship is unavailable, including CPT/OPT • Reliable high-speed internet with a wired connection and suitable home workspace for remote work • Fiber, cable, or DSL internet connection required for remote work • Ability to participate in an on-call rotation and occasional after-hours change windows

🏖️ Benefícios

• Flexible work environment • Medical, dental, vision and life insurance • Retirement savings – 401(k) plan with generous company matching contributions (up to 6%), financial advisory services, potential company discretionary contribution, and a broad investment lineup • Tuition reimbursement up to $5,250/year • Business-casual environment with the option to wear jeans • Generous paid time off upon hire, including a paid time off program, ten paid company holidays and three floating holidays each calendar year • Paid volunteer time — 16 hours per calendar year • Leave of absence programs, including paid parental leave, paid short- and long-term disability, and Family and Medical Leave (FMLA) • Business Resource Groups (BRGs) open to all • Bonus program opportunity for non-sales positions • Necessary computer equipment provided • Remote-work high-speed internet and home-workspace requirements

Candidatar-se

Vagas Similares

🕒 Agosto 1

Scientific Games

10.000+ funcionários

🎮 Jogos

🤝 B2B

Senior DevOps Engineer enhancing infrastructure and deployment reliability for Scientific Games. Collaborating across teams to scale secure, high-performing platforms and improve delivery efficiency.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 31

Andromeda

11 - 50

🏥 Saúde

💼 Consultoria

🏨 Hospitalidade

Engineer embedded with teams running large-scale training and inference on GPU clusters in production. Responsible for onboarding, debugging, and improving performance while ensuring reliability.

🇺🇸 Estados Unidos – Remoto (EUA)

💰 $15.142.238 Series A - Andromeda Robotics em 2025-09

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 31

Smithfield Foods

10.000+ funcionários

🏭 Manufatura

🌾 Agricultura

🍽️ Alimentos e Bebidas

Senior Utilities Engineer at Smithfield Foods optimizing utility systems for industrial refrigeration and ensuring compliance. Collaborating with facilities teams for operational efficiency and system improvements across various locations.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 31

Rocket.net

11 - 50

☁️ SaaS

🛍️ Comércio Eletrônico

🏢 Corporativo

Site Reliability Engineer ensuring high standards for servers, services, and customer environments at Rocket.net. Providing advanced technical support and ensuring platform reliability.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 31

Hearst Health

1001 - 5000

🏥 Saúde

⚕️ Seguro de Saúde

☁️ SaaS

Senior DevOps Engineer II at Bring a Trailer modernizing the infrastructure and security of a trusted automotive marketplace. Developing applications to enable engineering teams to work efficiently.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $133.000 - $150.000 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório