Senior Site Reliability Engineer – Fedramp

🕒 Setembro 17

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $85.000 - $141.000 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

infoinfo

👻 Score fantasma 0%

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Coalfire

Coalfire

1001 - 5000 funcionários

Fundada em 2001

💼 Consultoria

🏥 Saúde

📦 Logística

Consulting • Healthcare • Logistics

A Coalfire é uma fornecedora de serviços de cibersegurança que ajuda empresas a melhorarem sua resiliência em segurança e a simplificarem a conformidade regulatória. A empresa oferece serviços especializados, incluindo programas de cibersegurança focados em ameaças, automação de conformidade, gestão de riscos e serviços de consultoria em segurança em diversos setores, como serviços financeiros, saúde, varejo e tecnologia. A Coalfire é conhecida por sua expertise em hackers e defensores, e suas plataformas são projetadas para fortalecer a resiliência cibernética dos clientes, reduzir superfícies de ataque e acelerar o alcance de objetivos de conformidade como FedRAMP e HITRUST.

Descrição

• Own an operational capability for the managed estate, including automation, runbooks, and service standards • Design observability for regulated cloud environments, including telemetry and log pipelines, service-level objectives, alert quality, and escalation paths • Build and maintain continuous-monitoring evidence pipelines • Own backup and recovery engineering, including tested recovery procedures, measurable recovery objectives, and outage automation • Serve as the senior escalation point in client environments and resolve complex operational incidents • Lead incident and problem management, including major-event incident command, blameless post-incident reviews, and corrective actions • Automate operational toil using infrastructure-as-code, pipelines, and scripting • Partner with Engagement Architects and Build teams on transition into managed operations • Hold on-call responsibility and improve rotation coverage, alert actionability, and team load • Represent operational posture to clients and support renewals and expansions • Mentor and lead Site Reliability Engineers and junior staff • Author and peer review code, runbooks, operational design documentation, and compliance artifacts

🎯 Requisitos

• BS or above in a related Information Technology field or equivalent combination of education and experience • Bachelor’s degree or equivalent combination of education and work experience • Professional- or specialty-level certification in AWS, Azure, or GCP; associate-level certification considered with equivalent demonstrated depth • 5+ years in site reliability engineering, cloud operations, platform engineering, or managed services • 5+ years operating production cloud environments in AWS, Azure, or GCP, including monitoring, incident response, and automation • Automation-first mindset with deep Infrastructure-as-Code, CI/CD, scripting, and policy-as-code • Deep operational command of at least one major cloud platform and working knowledge of a second • Observability engineering experience with metrics, logging, log pipelines, distributed tracing, SLI/SLO definition, and alert design • Demonstrated incident response and incident command capability • Backup, recovery, and resilience engineering experience • Working command of NIST 800-53, FedRAMP, or comparable security control frameworks • Ability to lead technical client conversations about operational posture, risk, and trade-offs • Demonstrated ability to mentor engineers and improve team output • Excellent communication, organizational, and problem-solving skills • Effective documentation skills, including technical diagrams, runbooks, and written descriptions • Ability to work independently and as part of a team • Critical thinking and ability to balance security and availability requirements against mission needs • Demonstrated experience owning an operational capability, monitoring platform, or reusable automation used by multiple teams or clients • Experience as the senior operational escalation point on client-facing managed services, including incident command on major events • Advanced experience with Infrastructure-as-Code and orchestration/automation tools such as Terraform and Ansible • Experience transitioning environments from build into steady-state operations

🏖️ Benefícios

• Flexible work model allowing employees to choose when and where they work • Paid parental leave • Flexible time off • Certification and training reimbursement • Digital mental health and wellbeing support membership • Comprehensive insurance options • Employee resource groups • In-person and virtual events • Annual incentive, commission, and/or recognition programs may be available

Candidatar-se

Vagas Similares

🕒 Setembro 17

OnePay

501 - 1000

💳 Fintech

🏦 Bancário

₿ Cripto

SRE Lead building reliable infrastructure for OnePay’s consumer fintech platform. Leading senior engineers while coding, automating operations, and improving incident response.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $250.000 - $280.000 / ano

💰 $300.000.000 Series unknown em 2025-01

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Setembro 17

Ad Hoc LLC

501 - 1000

💼 Consultoria

🏥 Saúde

📦 Logística

Senior DevOps Engineer building AWS infrastructure and CI/CD pipelines for Ad Hoc’s Veterans Affairs digital services. Improving security, reliability, developer experience, and software delivery speed.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $130.000 - $140.000 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Setembro 17

Octus

501 - 1000

💼 Consultoria

⚖️ Jurídico

📚 Educação

Lead DevOps Engineer leading cloud infrastructure, CI/CD, and security for Octus, a global credit intelligence and analytics provider. Mentoring DevOps engineers and ensuring reliable, scalable systems.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $175.000 - $225.000 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Setembro 17

PhoenixTeam

51 - 200

💳 Fintech

🏠 Imobiliário

🤖 Inteligência Artificial

DevOps Manager modernizing Jenkins-based CI/CD and Fortify quality controls for PhoenixTeam's federal FHA mortgage program. Coordinating releases, documentation, and delivery across development teams.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟠 Sênior

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Setembro 17

Sprezzatura

51 - 200

🏛️ Governo

💼 Consultoria

🏥 Saúde

DevSecOps Engineer building AWS, Kubernetes, and Terraform infrastructure for VA.gov. Automating secure deployments and platform services supporting millions of Veterans.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $110.000 - $135.000 / ano

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório