DevOps Engineer

🕒 Agosto 5

🇮🇳 Índia – Remoto

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

👻 Score fantasma 39%

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Signalmash

Signalmash

51 - 200 funcionários

Fundada em 2020

💼 Consultoria

📦 Logística

🏥 Saúde

Consulting • Logistics • Healthcare

A Signalmash é uma plataforma de comunicações boutique (CPaaS) que oferece a empresas serviços de mensagens e voz com qualidade de operadora, APIs amigáveis para desenvolvedores e suporte personalizado. Suas ofertas incluem APIs de Mensagens (10DLC, toll-free, short code), RCS e MMS/SMS, SIP/Voice trunking, CCaaS/PBX hospedado, provisionamento de números de telefone e identificador de chamadas personalizado, além de conformidade, KYC, faturamento e operações de comunicações como serviços gerenciados. A Signalmash tem como público-alvo desenvolvedores, ISVs, revendedores e empresas que buscam uma infraestrutura de comunicações confiável e em conformidade, com onboarding e suporte personalizados.

Descrição

• Own CI/CD pipelines end to end using GitHub Actions, including self-hosted runner infrastructure, build speed optimization, caching, elimination of flaky jobs, and safe automated deployments • Operate and improve Kubernetes environments, including self-managed K3s on bare metal and GCP cloud deployments • Manage PostgreSQL operations, including schema migrations, backup strategy, point-in-time recovery, and performance tuning • Build and maintain observability through metrics, logs, dashboards, and alerting • Own backup and disaster recovery processes, restore testing, and operational runbooks • Harden platform security through secrets management, TLS, least-privilege access, overlay networking, and dependency management • Coordinate production incidents and document root causes and corrective actions • Reduce infrastructure costs while maintaining reliability • Introduce responsible AI-assisted engineering workflows using Claude Code, Codex, Cursor, Copilot, or similar tools • Work alongside internal engineers and external development partners • Connect with the US management team on priorities, risks, and incidents • Own production issues from alert through documented root cause

🎯 Requisitos

• 4+ years of DevOps, SRE, or Platform Engineering experience • Strong Kubernetes experience, including deployment, networking, storage, upgrades, and ideally self-managed or bare-metal clusters • Deep CI/CD experience using GitHub Actions or equivalent • PostgreSQL administration, including migrations, backup, restore, and performance • Experience with Docker, Linux administration, shell scripting, and Python or Node.js • Cloud experience with GCP or equivalent AWS/Azure • Experience with Prometheus, Grafana, or equivalent monitoring platforms • Proven experience operating production customer-facing platforms • Practical use of AI coding tools in production repositories • Strong written and spoken English • Based in India • Available to work until approximately 11:30 PM IST to overlap with the US management team • Preferred: experience with communications or CPaaS platforms • Preferred: hybrid bare-metal and cloud infrastructure experience • Preferred: Cloudflare, GitOps (ArgoCD/Flux), Infrastructure as Code • Preferred: ORM migration workflows such as Prisma • Preferred: cost optimization achievements • Preferred: SOC 2 or similar security environments

🏖️ Benefícios

• Competitive salary • Flexible remote-first work environment • Direct collaboration with US leadership and engineering partners • Opportunity to shape infrastructure strategy from the ground up • Exposure to AI-assisted software engineering workflows • Career growth as the engineering organization expands • Potential opportunity to relocate to Kochi if a local engineering office is established

Candidatar-se

Vagas Similares

🕒 Agosto 4

Neo4j

501 - 1000

☁️ SaaS

🤖 Inteligência Artificial

🏢 Corporativo

Cloud Operations Engineer managing and troubleshooting customer Neo4j database infrastructure. Supporting deployments, monitoring, upgrades, and incidents across AWS, Azure, Google Cloud, virtual, and bare-metal environments.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 29

Outmarket AI

11 - 50

🤖 Inteligência Artificial

🛡️ Seguros

☁️ SaaS

DevOps Engineer managing infrastructure and delivery platform for AI products. Focus on security, reliability, and observability while working in an AI-first environment.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 29

Granicus

501 - 1000

🏛️ Governo

☁️ SaaS

📋 Conformidade

Site Reliability Engineer 3 modernizing reliability engineering for Granicus with a focus on AIOps and automation. Improve service reliability and build scalable, resilient platforms for various workloads.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 28

Sezzle

201 - 500

💳 Fintech

👥 B2C

🛍️ Comércio Eletrônico

Senior Site Reliability Engineer at Sezzle resolving infrastructure challenges and enhancing reliability through scalable solutions. Seeking innovative and experienced candidates to drive technical excellence.

🇮🇳 Índia – Remoto

💵 $5.000 - $9.500 / mês

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 28

fal

51 - 200

🤖 Inteligência Artificial

🔌 API

☁️ SaaS

Machine Learning Engineer focusing on the reliability and security of generative media model APIs at fal. Working with cutting-edge models and infrastructure in a remote setting.

🗣️🇺🇸🇬🇧 Inglês obrigatório