Senior Site Reliability Engineer

🕒 Agosto 24

🇮🇳 Índia – Remoto

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

👻 Score fantasma 10%

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Akamai Technologies

Akamai Technologies

5001 - 10000 funcionários

🔒 Cibersegurança

💰 Post-IPO Equity em 2001-07

Cloud Computing • Cybersecurity • Content Delivery

A Akamai Technologies é um provedor líder de serviços em nuvem especializado em oferecer soluções de segurança, computação em nuvem e entrega de conteúdo. A empresa oferece uma gama de serviços, como segurança de API, proteção contra DDoS e otimização de desempenho para aplicativos web, garantindo experiências de usuário seguras e confiáveis. Com uma infraestrutura global robusta, a Akamai capacita empresas a otimizar sua presença digital, protegendo contra diversas ameaças cibernéticas e melhorando o desempenho de aplicativos.

Descrição

• Oversee, scale, and optimize next-generation dedicated AI hardware infrastructure • Ensure uptime and reliability of AI hardware infrastructure offerings • Enhance reliability, scalability, and performance across high-density hardware and software infrastructure in regional data centers • Define KPIs, implement proactive monitoring, automate operations, and resolve urgent issues • Develop and scale Python tooling and infrastructure-as-code utilities to eliminate operational toil and automate fleet-wide provisioning • Integrate automated workflows across corporate ticketing systems for hardware and network break-fix incidents • Use AI tools and LLM-based development approaches for technical execution, script creation, and system evaluation • Improve availability, latency, and systemic health of private cloud and compute environments • Design telemetry pipelines, Prometheus/Grafana dashboards, and AI-based anomaly detection for bare-metal and virtualized environments • Participate in 24x7x365 on-call rotations and lead real-time incident management through PagerDuty and Slack workflows • Partner with infrastructure vendors and coordinate on-site field technicians

🎯 Requisitos

• 5+ years of relevant experience • Bachelor's degree in Computer Science or related field • Exceptional proficiency in Python and tooling/coding for scalable operational tools, API integrations, and automation frameworks • Hands-on experience with Prometheus, Grafana, OpenTelemetry, and Loki • Working understanding of advanced networking topologies, high-bandwidth routing/switching infrastructure, BGP, and dual-stack IPv4/IPv6 networks • Expertise designing new service rollouts, operational readiness criteria, telemetry baselines, and alerting thresholds • Extensive experience building technical runbooks, leading complex incident response bridges, and conducting blameless post-mortems • Ability to own ambiguous technical challenges, coordinate cross-functional teams, and drive production-grade solutions

🏖️ Benefícios

• Health and well-being benefits • Financial benefits • FlexBase workplace flexibility: work at home, in an office, or a combination of both

Candidatar-se

Vagas Similares

🕒 Agosto 21

Simbian

11 - 50

🤖 Inteligência Artificial

🔒 Cibersegurança

Forward Deployment Engineer implementing Simbian’s AI SOC platform for enterprise and MSSP customers in India. Integrating SIEM, EDR, XDR, IAM, cloud tools, APIs, and webhooks.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 13

Lingaro

1001 - 5000

💼 Consultoria

📣 Marketing

📦 Logística

DevOps Engineer maintaining observability platforms and infrastructure for Lingaros Group in India. Building Grafana and Prometheus monitoring, CI pipelines, automated diagnostics, and failover systems.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 12

Provenir

201 - 500

☁️ SaaS

💳 Fintech

🤖 Inteligência Artificial

Senior DevOps Engineer architecting Provenir’s AWS and Kubernetes platform for AI-powered decision intelligence. Driving GitOps, AIOps, CI/CD modernization, data infrastructure, observability, and cloud cost optimization.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 12

BETSOL

501 - 1000

💼 Consultoria

🏥 Saúde

📦 Logística

Senior Cloud Engineer securing Azure and GCP platforms for BETSOL, an AI-powered enterprise cloud transformation company. Building infrastructure, CI/CD automation, monitoring, and portal applications.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 5

Signalmash

51 - 200

💼 Consultoria

📦 Logística

🏥 Saúde

DevOps Engineer owning Kubernetes, CI/CD, PostgreSQL, observability, and security for Signalmash’s cloud communications platform. Improving reliability, deployment speed, recovery, and infrastructure costs from India.

🗣️🇺🇸🇬🇧 Inglês obrigatório