Site Reliability Engineer

🕒 Junho 22

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Stack AV

Stack AV

51 - 200 funcionários

📦 Logística

🏭 Manufatura

💼 Consultoria

Logistics • Manufacturing • Consulting

A Stack AV é uma empresa que está revolucionando a indústria de transporte por meio de suas soluções de caminhões autônomos, impulsionadas por inteligência artificial avançada. A empresa se concentra no desenvolvimento de sistemas autônomos movidos a IA para aumentar a segurança, a confiabilidade e a eficiência nas operações de transporte. Stack AV está comprometida em enfrentar os desafios da indústria de caminhões projetando soluções inteligentes para melhorar a inteligência da cadeia de suprimentos, os resultados empresariais e a velocidade de entrega. A segurança é um princípio fundamental, e a empresa utiliza tecnologias de ponta em IA, aprendizado de máquina e nuvem para inovar no setor.

Descrição

• Instrument systems scheduling and executing large-scale batch workloads across Kubernetes clusters. • Diagnose and triage job failures for customers. • Collaborate with teams across the company to understand workload requirements and improve platform capabilities. • Scale the reliability and velocity of our systems and processes through increased automation. • Document actions to build a comprehensive library of runbooks, which will act as a knowledge base and foundation for automation. • Participate in an on-call rotation to uphold the SLOs and SLAs of production services. • Contribute to platform tooling, automation, and CI/CD workflows.

🎯 Requisitos

• Fundamental understanding of Linux operating system internals, TCP/IP networking, and storage subsystems. • Strong experience with Kubernetes and container orchestration in production grade environments. • Understanding of engineering design limitations and ability to provide guidance to teams to scale their services to achieve desired performance within budget. • Strong experience implementing and debugging cloud native and open source tools such as Kubernetes, etcd, Prometheus, OpenTelemetry. • Strong communication skills and the ability to work effectively in a diverse and distributed team.

🏖️ Benefícios

• We are proud to be an equal opportunity workplace. • We believe that diverse teams produce the best ideas and outcomes. • We are committed to building a culture of inclusion, entrepreneurship, and innovation across gender, race, age, sexual orientation, religion, disability, and identity.

Candidatar-se

Vagas Similares

🕒 Junho 22

nDeavour Consulting

1 - 10

💼 Consultoria

📦 Logística

📣 Marketing

Site Reliability Engineer ensuring health, performance, and delivery of infrastructure systems at Mobile Wave Solutions. Working collaboratively with engineers to automate processes and improve operational reliability.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 21

Tkxel

501 - 1000

💼 Consultoria

📣 Marketing

🏥 Saúde

Senior Azure DevOps Engineer at Tkxel designing, implementing, and maintaining Azure DevOps infrastructure while mentoring junior team members in a dynamic environment.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 20

Gorilla Logic

501 - 1000

💼 Consultoria

📣 Marketing

📦 Logística

Technical Engineering Manager leading high-performing cloud and DevOps teams. Guiding architecture and delivery of scalable, reliable, and secure cloud solutions for clients.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟠 Sênior

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 19

Planned Systems International

1001 - 5000

💼 Consultoria

🎖️ Defesa

📦 Logística

DevOps Software Engineer responsible for cloud automation and software development. Supporting government research and development activities in the Health and Defense sectors.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 19

Skydio

501 - 1000

🎖️ Defesa

🏭 Manufatura

📦 Logística

Deployment Engineer at Skydio leading technical implementations and maintaining customer success with cloud connected products. Focused on WiFi, networking, and client communication within the Northeast region.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $115.000 - $135.000 / ano

💰 $170.000.000 Series E - Skydio em 2024-11

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório