Senior Site Reliability Engineer

🕒 Junho 18

🥔 Idaho, Texas – Remoto

infoinfo

💵 $117.000 - $209.330 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

👻 Score fantasma 17%

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Autodesk

Autodesk

10.000+ funcionários

Fundada em 1982

🏗️ Construção

🏭 Manufatura

💼 Consultoria

Construction • Manufacturing • Consulting

A Autodesk é líder global em software para designers, engenheiros, construtores e criadores. A empresa oferece um portfólio abrangente de aplicativos de design e engenharia, incluindo produtos populares como AutoCAD, Revit e 3ds Max. Por meio de sua plataforma Design and Make, a Autodesk capacita profissionais de diversos setores a projetar, visualizar e gerenciar projetos com eficiência, fomentando a inovação e a sustentabilidade em arquitetura, engenharia, construção e manufatura.

Descrição

• Serve as a primary owner for the reliability, availability, performance, operability, and capacity of one or more production services • Deploy, operate, maintain, and continuously improve production services running in Autodesk GovCloud environments • Partner with engineering teams to ensure services are designed with reliability, scalability, security, and operability in mind • Define and operate reliability practices such as SLOs/SLIs, error budgets, production readiness reviews, service reviews, and operational health reviews • Build automation to improve deployment safety, operational efficiency, incident response, and service recovery • Design, develop, and maintain software, automation, and tooling that improve the reliability, scalability, and efficiency of production systems • Implement and improve monitoring, alerting, logging, tracing, and observability capabilities across supported services • Lead and participate in incident response, troubleshooting, and post-incident reviews focused on learning and continuous improvement • Develop and maintain operational documentation, runbooks, and recovery procedures • Scale and enhance resilience testing and Gameday practices to validate system behavior, recovery capabilities, and operational readiness • Continuously identify and eliminate operational toil through software engineering, automation, and process improvement • Ensure supported services remain compliant with Autodesk security, privacy, and regulatory requirements, including FedRAMP and related controls where applicable • Participate in a 24x7 on-call rotation for production services

🎯 Requisitos

• B.S. or higher in Computer Science, Engineering, or a related technical discipline, or equivalent practical experience • 7+ years of experience in Site Reliability Engineering, Software Engineering, Platform Engineering, Cloud Infrastructure, or Production Operations • Experience operating and supporting customer-facing production services in large-scale cloud environments • Strong understanding of reliability engineering principles, including SLOs/SLIs, observability, incident management, capacity planning, production readiness, and automation • Experience with AWS, Azure, or other public cloud platforms • Experience developing automation using languages such as Python, Go, Java, PowerShell, Bash, or similar • Experience with Infrastructure as Code, CI/CD pipelines, deployment automation, and modern cloud operations practices • Understanding of security, compliance, and operational risk management in production environments • Strong written and verbal communication skills.

🏖️ Benefícios

• Health and financial benefits • Time away and everyday wellness

Candidatar-se

Vagas Similares

🕒 Junho 18

Coupa Software

1001 - 5000

💼 Consultoria

📦 Logística

🏥 Saúde

Senior Database Reliability Engineer overseeing Cloud based SQL Server infrastructures at Coupa. Leading database architecture and ensuring reliable, high-performance data solutions.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 18

Pinterest

1001 - 5000

📱 Mídia

👥 B2C

Site Reliability Engineer enhancing AWS-based platform reliability at Pinterest and scaling Kubernetes workloads. Operating and improving cloud-native infrastructure with a focus on automation and resilience.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $114.297 - $235.319 / ano

💰 Post IPO equity em 2022-08

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 17

Intermedia Cloud Communications

1001 - 5000

💼 Consultoria

🏥 Saúde

⚖️ Jurídico

DevOps Engineer managing GCP infrastructure for cloud communications. Collaborating with development teams to maintain application deployment and infrastructure.

🇺🇸 Estados Unidos – Remoto (EUA)

💰 Venture Round em 2017-02

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 17

ClassWallet

11 - 50

💳 Fintech

📚 Educação

🏛️ Governo

DevOps Engineer optimizing AWS infrastructure, GitHub Actions, and observability for ClassWallet's public-funds digital wallet platform. Ensuring scalable, compliant, and reliable systems for government agencies.

🇺🇸 Estados Unidos – Remoto (EUA)

💰 $500.000 Debt Financing em 2020-05

⏰ Tempo Integral

🟠 Sênior

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 16

DexCare

51 - 200

🏥 Saúde

💼 Consultoria

📦 Logística

Senior Site Reliability Engineer at DexCare managing cloud-native infrastructure in AWS. Designing secure systems and ensuring observability while collaborating in an Agile environment.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $125.000 - $165.000 / ano

💰 $50.000.000 Series B em 2022-01

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório