Senior Site Reliability Engineer

🕒 Março 30

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Akamai Technologies

Akamai Technologies

5001 - 10000 funcionários

🔒 Cibersegurança

🏢 Corporativo

📱 Mídia

Cybersecurity • Enterprise • Media

A Akamai Technologies é uma empresa global de plataforma de borda e serviços em nuvem que oferece soluções de entrega de conteúdo, computação na borda e segurança. A empresa opera uma das maiores redes distribuídas do mundo para acelerar e proteger o tráfego da web, mídia e aplicações, oferecendo produtos para entrega de conteúdo, proteção contra DDoS, segurança de API e aplicativos, gerenciamento de bots, computação na borda (funções sem servidor/funções de borda) e inferência de IA na borda. A Akamai também fornece serviços de segurança focados em empresas (confiança zero, gerenciamento de identidade e acesso, acesso seguro à internet) e ferramentas de infraestrutura em nuvem/IA, e recentemente expandiu suas capacidades por meio de aquisições (por exemplo, LayerX) para adicionar controle de uso de IA baseado em navegador.

Descrição

• Own reliability workstreams for Akamai's serverless inference platform • Build automation and tooling • Contribute to architecture and operational decisions • Take ownership of critical reliability problems end-to-end • Partner with product engineering teams • Develop expertise in GPU infrastructure, Kubernetes at scale, and AI inference workloads • Build and maintain observability for AI workloads, including telemetry, dashboards, alerts, SLO/SLI tracking • Write automation and tooling to reduce operational toil, improve deployment safety, and accelerate incident response • Integrate AI workloads into Akamai's incident management processes • Build and maintain CI/CD integrations, deployment safety checks, and rollback automation • Collaborate with product engineering teams to improve reliability and ensure operational readiness for product releases • Contribute to capacity planning, autoscaling configuration, and workload scheduling for AI compute infrastructure

🎯 Requisitos

• 5+ years of experience in SRE, infrastructure engineering, or platform engineering, working with large-scale distributed systems • Extensive experience with Kubernetes and containerization at scale • Experience defining SLOs and working with observability tools such as Prometheus, Grafana, and distributed tracing • Coding ability in Python or Go for automation and tooling, with experience in CI/CD pipelines, deployment safety, and infrastructure-as-code • Interest in or experience with AI/ML infrastructure, model serving, or GPU workloads • Ability to take ownership of problems and drive them to resolution independently

🏖️ Benefícios

• Healthcare • 401K savings plan • Company holidays • Vacation (in the form of PTO) • Sick time • Family friendly benefits including parental leave • Employee assistance program focusing on mental and financial wellness • Flexible working arrangements

Candidatar-se

Vagas Similares

🕒 Março 30

Expert Executive Recruiters (EER Global)

51 - 200

💼 Consultoria

📦 Logística

🏥 Saúde

Senior DevOps/Infra Engineer needed to design, automate, and secure high-load infrastructure. Role involves kernel tuning, VPNs, monitoring, and CI/CD in a collaborative remote team.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Março 28

Close

51 - 200

💼 Consultoria

📣 Marketing

☁️ SaaS

Join Infrastructure Team at Close as a Site Reliability Engineer. Work on robust systems supporting a modern communication-focused CRM service for small scaling businesses.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Março 27

Red River

501 - 1000

💼 Consultoria

📦 Logística

Senior Wireless Deployment Engineer managing deployment, optimization, and lifecycle of enterprise wireless networks. Providing technical leadership and support for Aruba and Juniper Mist solutions.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Março 27

SGNL

11 - 50

🔒 Cibersegurança

🔐 Segurança

☁️ SaaS

Senior DevOps Engineer at SGNL solving authorization challenges for major companies. Collaborating and leading teams in a dynamic, scale-oriented environment.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Março 27

Cority

201 - 500

🏥 Saúde

📦 Logística

💼 Consultoria

Sr. DevOps Engineer working to deploy and operate systems at Cority, the global EHS software provider. Collaborating with engineering for continuous delivery and monitoring towards operational excellence.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório