DevOps Team Lead

Vaga não está no LinkedIn

🕒 Agosto 25

🧀 Wisconsin – Remoto

infoinfo

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

infoinfo

👻 Score fantasma 10%

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of URUS Group

URUS Group

1001 - 5000 funcionários

Fundada em 2018

🌾 Agricultura

🤝 B2B

🧬 Biotecnologia

Agriculture • B2B • Biotechnology

O URUS Group é uma holding formada em 2018, com foco na indústria agrícola global, atendendo especificamente produtores de leite e de carne bovina. Engloba várias empresas, incluindo AgSource, Alta Genetics e GENEX, oferecendo soluções de genética e de gestão de fazendas voltadas a melhorar a qualidade e a produtividade dos rebanhos. A URUS é comprometida com a sustentabilidade e com o desenvolvimento de tecnologias inovadoras que ampliam a rentabilidade das operações pecuárias comerciais em todo o mundo.

Descrição

• Own delivery, reliability, and day-to-day operation of the AWS cloud platform • Design, build, review, and operate Infrastructure-as-Code across a multi-account AWS environment • Own and improve infrastructure CI/CD pipelines with automated validation, security, compliance controls, and keyless cloud authentication • Build and maintain AWS networking, containers, databases, storage, messaging, systems management, backup, monitoring, and cost controls • Establish engineering standards for IaC, IAM, tagging, state management, versioning, security, documentation, and operational readiness • Translate product, engineering, security, and data needs into requirements, milestones, estimates, dependencies, and acceptance criteria • Own and refine the DevOps backlog and facilitate Scrum ceremonies • Strengthen reliability through observability, SLOs, alerting, runbooks, disaster recovery, backup and restore, and reduction of operational toil • Participate in a rotating night and weekend on-call schedule and provide technical leadership during incidents • Partner with security on identity, least privilege, network segmentation, encryption, secrets management, audit logging, and compliance • Lead cloud cost visibility and optimization, including tagging, right-sizing, commitment planning, anomaly response, and multi-region cost analysis • Evolve the platform toward globally distributed workloads, including regional rollout, data residency, replication, latency-aware routing, and cross-region failover • Use AI tools in engineering workflows and help operate infrastructure supporting emerging AI capabilities • Set technical direction with architecture and coach engineers through reviews, pairing, design discussions, documentation, and shared ownership

🎯 Requisitos

• Extensive experience in software engineering, infrastructure, or DevOps, with demonstrated experience operating at a senior or technical lead level in complex cloud environments • Expert-level, current hands-on experience with Terraform or equivalent Infrastructure as Code • Experience with IaC at scale, including reusable module design and versioning, state architecture, safe refactoring, upgrades, drift detection and remediation, and bringing legacy infrastructure under IaC management • Deep hands-on AWS experience across compute, storage, networking, IAM, and security, ideally within a multi-account AWS Organization • Proven ownership of infrastructure CI/CD pipelines, including automated policy, validation, and security controls • Experience designing and operating containerized or distributed production workloads with appropriate networking, security, and observability • Scripting proficiency and a track record of automating operational work • Demonstrated ownership of production operations, including on-call, incident response, runbook development, and independently troubleshooting and resolving platform issues • Experience leading Agile/Scrum practices, including backlog ownership, refinement, estimation, planning, and stakeholder negotiation • Ability to turn ambiguous needs into actionable technical plans accounting for requirements, dependencies, risks, milestones, and estimates • Strong understanding of cloud security, compliance, and risk management, including least privilege, zero trust, encryption, and exception management • Excellent written communication, documentation, and stakeholder engagement skills • Self-directed, collaborative approach focused on improving systems • Willingness and ability to participate in a rotating night and weekend on-call schedule and independently manage production incidents • Preferred: Microsoft Azure experience • Preferred: globally distributed or multi-region applications experience • Preferred: monorepo IaC, policy-as-code, IaC security tooling, or automated code-quality platforms • Preferred: FinOps experience • Preferred: SRE practices • Preferred: analytics, data engineering, database, or integration workload experience • Preferred: AWS certification • Preferred: hybrid on-premises connectivity or software development lifecycle improvement experience • Preferred: AI/ML infrastructure experience

Candidatar-se

Vagas Similares

🕒 Agosto 25

System Automation Corporation

51 - 200

💼 Consultoria

🏥 Saúde

⚖️ Jurídico

Site Reliability Engineer operating Azure infrastructure for System Automation’s regulatory-agency SaaS platform. Automating reliability, observability, CI/CD, security, and incident response.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $120.000 - $140.000 / ano

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 25

Blue River Technology

201 - 500

🌾 Agricultura

🤖 Inteligência Artificial

🔧 Hardware

Senior Site Reliability Engineer scaling Kubernetes platforms and cloud infrastructure. Supporting Blue River Technology’s autonomous robotics products through reliability, security, and observability.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 24

The Home Depot

10.000+ funcionários

🏗️ Construção

📦 Logística

🛒 Varejo

Senior Principal Reliability Engineer designing resilient infrastructure for Home Depot store systems, payments, and COM platforms. Guiding multiple engineering teams on reliability, cloud costs, and technology strategy.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $170.000 - $280.000 / ano

💰 Debt Financing em 2007-07

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 24

Symbotic

501 - 1000

🔧 Hardware

📦 Logística

🤖 Inteligência Artificial

Senior reliability manager scaling maintenance and asset performance across Exol’s automated warehouses. Driving uptime, safety, launches, vendor governance, and enterprise reliability standards.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 24

PingWind Inc. (SDVOSB)

51 - 200

💼 Consultoria

📦 Logística

🏥 Saúde

DevSecOps Engineer building and deploying secure cloud-based IAM systems for federal government clients. Maintaining highly available architectures, automated delivery, compliance, and infrastructure upgrades.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $93.000 - $128.000 / ano

⏰ Tempo Integral

🟠 Sênior

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório