Cloud Site Reliability Engineer

🕒 Setembro 21

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $120.000 - $130.000 / ano

⏰ Tempo Integral

🟠 Sênior

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

👻 Score fantasma 0%

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Cadwell

Cadwell

51 - 200 funcionários

Fundada em 1979

🏥 Saúde

🏭 Manufatura

🔧 Hardware

Healthcare • Manufacturing • Hardware

A <Cadwell> é uma empresa de tecnologia médica que projeta, fabrica e oferece suporte a hardwares e softwares de neurodiagnóstico, monitoramento neurocirúrgico intraoperatório e diagnóstico do sono. Ela fornece sistemas de EEG, EMG, condução nervosa, potenciais evocados, ultrassom neuromuscular, IONM, polissonografia e testes de apneia do sono domiciliar, além de softwares clínicos (Sierra, Arc), gerenciamento de dados na nuvem (CadLink), eletrodos de consumo e acessórios, além de treinamento, educação e suporte técnico para prestadores de serviços de saúde.

Descrição

• Build, maintain, and improve AWS cloud infrastructure for hosted customer environments • Automate build, test, and deployment pipelines • Build and maintain log ingestion, monitoring, alerting, and observability systems • Lead incident response and conduct post-mortems with corrective actions • Design and test backup and disaster recovery strategies • Optimize cloud compute, storage, lifecycle policies, performance, retention, and cost • Implement cybersecurity practices including IAM, network segmentation, encryption, patching, and vulnerability remediation • Partner with software engineering on application reliability, scalability, performance, and architecture decisions • Support hosted environment onboarding, migrations, and upgrades with enterprise support and project delivery teams • Ensure compliance with HIPAA, GDPR, and IEC 62304-related quality processes • Document infrastructure architecture, runbooks, and escalation procedures • Perform other assigned activities

🎯 Requisitos

• Expert-level knowledge of core AWS services and architectural patterns, including compute, object storage, networking, managed databases, and ECS • Command of infrastructure as code and deployment automation using Terraform, YAML, JSON/Jinja, CI/CD pipelines, and Bash, Python, or JavaScript/TypeScript • Practical knowledge of cloud security, IAM, least-privilege design, secret management, encryption, network segmentation, and vulnerability remediation • Experience building monitoring, logging, and observability systems • Knowledge of backup, disaster recovery, business continuity, replication, restore validation, and recovery objectives • Managed database administration and cloud cost management experience • Familiarity with HIPAA, GDPR, and IEC 62304 • Bachelor's degree in Computer Science, Information Technology, or related field, or equivalent work experience • 8+ years of experience in site reliability engineering, cloud infrastructure, or DevOps, including recent hands-on production AWS experience • Experience operating production infrastructure under formal on-call and incident management practices • Reliable high-speed internet • Required participation in on-call rotation, including occasional after-hours and weekend response • Flexibility to work evenings or weekends for planned changes • AWS certification, regulated-environment experience, healthcare support experience, virtualization, and HL7 tooling are preferred

🏖️ Benefícios

• Reliable high-speed internet required • Participation in an on-call rotation, including occasional after-hours and weekend response • Flexibility to work evenings or weekends for planned infrastructure changes • Travel up to 10% for company meetings

Candidatar-se

Vagas Similares

🕒 Setembro 21

Raya

51 - 200

🌍 Impacto Social

👥 B2C

📱 Mídia

Senior infrastructure engineer building Raya’s scalable AWS and Kubernetes platform. Improving reliability, performance, security, automation, and multi-region infrastructure.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Setembro 21

Arize AI

51 - 200

🤖 Inteligência Artificial

☁️ SaaS

🏢 Corporativo

DevOps Engineer supporting Arize AI’s observability platform across SaaS and on-prem environments. Managing Kubernetes, cloud infrastructure, monitoring, and release automation for customer deployments.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Setembro 21

AuthZed

11 - 50

🔌 API

🔒 Cibersegurança

☁️ SaaS

Senior Site Reliability Engineer securing AuthZed’s cloud infrastructure and authorization platform, including SpiceDB. Building Kubernetes guardrails, supply-chain security, vulnerability management, and incident response.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Setembro 21

KASHIO

51 - 200

💳 Fintech

☁️ SaaS

🤝 B2B

DevOps Engineer gestionando infraestructura AWS, Kubernetes y automatización para Kashio. Optimizando CI/CD, observabilidad, seguridad y confiabilidad de entornos cloud.

🇺🇸 Estados Unidos – Remoto (EUA)

💰 $500.000 Seed Round - KashIO em 2021-04

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🗣️🇪🇸 Espanhol obrigatório

🕒 Setembro 19

Clinician Nexus

51 - 200

🏥 Saúde

⚕️ Seguro de Saúde

📚 Educação

DevOps Manager leading secure, reliable platform engineering for Clinician Nexus, a healthcare workforce technology company. Balancing team leadership with hands-on AWS, Kubernetes, Terraform, and CI/CD engineering.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $133.100 - $221.900 / ano

💰 Seed Round em 2019-12

⏰ Tempo Integral

🟠 Sênior

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório