Senior Site Reliability Engineer

🕒 Junho 18

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Autodesk

Autodesk

10.000+ funcionários

Fundada em 1982

🏗️ Construção

🏭 Manufatura

💼 Consultoria

Construction • Manufacturing • Consulting

A Autodesk é líder global em software para designers, engenheiros, construtores e criadores. A empresa oferece um portfólio abrangente de aplicativos de design e engenharia, incluindo produtos populares como AutoCAD, Revit e 3ds Max. Por meio de sua plataforma Design and Make, a Autodesk capacita profissionais de diversos setores a projetar, visualizar e gerenciar projetos com eficiência, fomentando a inovação e a sustentabilidade em arquitetura, engenharia, construção e manufatura.

Descrição

• Serve as a primary owner for the reliability, availability, performance, operability, and capacity of one or more production services • Deploy, operate, maintain, and continuously improve production services running in Autodesk GovCloud environments • Partner with engineering teams to ensure services are designed with reliability, scalability, security, and operability in mind • Define and operate reliability practices such as SLOs/SLIs, error budgets, production readiness reviews, service reviews, and operational health reviews • Build automation to improve deployment safety, operational efficiency, incident response, and service recovery • Design, develop, and maintain software, automation, and tooling that improve the reliability, scalability, and efficiency of production systems • Implement and improve monitoring, alerting, logging, tracing, and observability capabilities across supported services • Lead and participate in incident response, troubleshooting, and post-incident reviews focused on learning and continuous improvement • Develop and maintain operational documentation, runbooks, and recovery procedures • Scale and enhance resilience testing and Gameday practices to validate system behavior, recovery capabilities, and operational readiness • Continuously identify and eliminate operational toil through software engineering, automation, and process improvement • Ensure supported services remain compliant with Autodesk security, privacy, and regulatory requirements, including FedRAMP and related controls where applicable • Participate in a 24x7 on-call rotation for production services • Function effectively in a fast-paced environment while helping establish and mature operational excellence practices for Autodesk GovCloud

🎯 Requisitos

• B.S. or higher in Computer Science, Engineering, or a related technical discipline, or equivalent practical experience • 7+ years of experience in Site Reliability Engineering, Software Engineering, Platform Engineering, Cloud Infrastructure, or Production Operations • Experience operating and supporting customer-facing production services in large-scale cloud environments • Strong understanding of reliability engineering principles, including SLOs/SLIs, observability, incident management, capacity planning, production readiness, and automation • Experience with AWS, Azure, or other public cloud platforms • Experience developing automation using languages such as Python, Go, Java, PowerShell, Bash, or similar • Experience with Infrastructure as Code, CI/CD pipelines, deployment automation, and modern cloud operations practices • Understanding of security, compliance, and operational risk management in production environments • Strong written and verbal communication skills.

🏖️ Benefícios

• Health and financial benefits • Time away • Everyday wellness

Candidatar-se

Vagas Similares

🕒 Junho 18

Coupa Software

1001 - 5000

💼 Consultoria

📦 Logística

🏥 Saúde

Senior Database Reliability Engineer overseeing Cloud based SQL Server infrastructures at Coupa. Leading database architecture and ensuring reliable, high-performance data solutions.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 18

Pinterest

1001 - 5000

📱 Mídia

👥 B2C

Site Reliability Engineer enhancing AWS-based platform reliability at Pinterest and scaling Kubernetes workloads. Operating and improving cloud-native infrastructure with a focus on automation and resilience.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $114.297 - $235.319 / ano

💰 Post IPO equity em 2022-08

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 18

Guidehouse

10.000+ funcionários

🏥 Saúde

🎖️ Defesa

📦 Logística

Site Reliability Engineer collaborating with teams to establish SRE practices and participate in system design reviews at Guidehouse. Focused on AWS cloud infrastructure and promoting automation.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $80.000 - $133.000 / ano

💰 Grant em 2023-02

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 17

Intermedia Cloud Communications

1001 - 5000

💼 Consultoria

🏥 Saúde

⚖️ Jurídico

DevOps Engineer managing GCP infrastructure for cloud communications. Collaborating with development teams to maintain application deployment and infrastructure.

🇺🇸 Estados Unidos – Remoto (EUA)

💰 Venture Round em 2017-02

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 17

Accela

201 - 500

💼 Consultoria

📦 Logística

🏗️ Construção

Lead Site Reliability Engineer ensuring scalability, performance, and operational excellence of Accela's Civic Platform through technical leadership and cloud modernization. Collaborating with cross-functional teams to enhance SaaS offerings for government software solutions.

🗣️🇺🇸🇬🇧 Inglês obrigatório