DevOps Engineer

🕒 August 5

🇮🇳 India – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 40%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Signalmash

Signalmash

51 - 200 employees

Founded 2020

💼 Consulting

📦 Logistics

🏥 Healthcare

Consulting • Logistics • Healthcare

Signalmash is a boutique communications platform (CPaaS) that provides businesses with carrier-grade messaging and voice services, developer-friendly APIs, and white‑glove support. Their offerings include Messaging APIs (10DLC, toll-free, short code), RCS and MMS/SMS, SIP/Voice trunking, CCaaS/hosted PBX, phone number provisioning and branded caller ID, plus compliance, KYC, billing and communications operations as managed services. Signalmash targets developers, ISVs, resellers and enterprises seeking reliable, compliant communications infrastructure and personalized onboarding and support.

📋 Description

• Own CI/CD pipelines end to end using GitHub Actions, including self-hosted runner infrastructure, build speed optimization, caching, elimination of flaky jobs, and safe automated deployments • Operate and improve Kubernetes environments, including self-managed K3s on bare metal and GCP cloud deployments • Manage PostgreSQL operations, including schema migrations, backup strategy, point-in-time recovery, and performance tuning • Build and maintain observability through metrics, logs, dashboards, and alerting • Own backup and disaster recovery processes, restore testing, and operational runbooks • Harden platform security through secrets management, TLS, least-privilege access, overlay networking, and dependency management • Coordinate production incidents and document root causes and corrective actions • Reduce infrastructure costs while maintaining reliability • Introduce responsible AI-assisted engineering workflows using Claude Code, Codex, Cursor, Copilot, or similar tools • Work alongside internal engineers and external development partners • Connect with the US management team on priorities, risks, and incidents • Own production issues from alert through documented root cause

🎯 Requirements

• 4+ years of DevOps, SRE, or Platform Engineering experience • Strong Kubernetes experience, including deployment, networking, storage, upgrades, and ideally self-managed or bare-metal clusters • Deep CI/CD experience using GitHub Actions or equivalent • PostgreSQL administration, including migrations, backup, restore, and performance • Experience with Docker, Linux administration, shell scripting, and Python or Node.js • Cloud experience with GCP or equivalent AWS/Azure • Experience with Prometheus, Grafana, or equivalent monitoring platforms • Proven experience operating production customer-facing platforms • Practical use of AI coding tools in production repositories • Strong written and spoken English • Based in India • Available to work until approximately 11:30 PM IST to overlap with the US management team • Preferred: experience with communications or CPaaS platforms • Preferred: hybrid bare-metal and cloud infrastructure experience • Preferred: Cloudflare, GitOps (ArgoCD/Flux), Infrastructure as Code • Preferred: ORM migration workflows such as Prisma • Preferred: cost optimization achievements • Preferred: SOC 2 or similar security environments

🏖️ Benefits

• Competitive salary • Flexible remote-first work environment • Direct collaboration with US leadership and engineering partners • Opportunity to shape infrastructure strategy from the ground up • Exposure to AI-assisted software engineering workflows • Career growth as the engineering organization expands • Potential opportunity to relocate to Kochi if a local engineering office is established

Apply Now

Similar Jobs

🕒 August 4

Neo4j

501 - 1000

☁️ SaaS

🤖 Artificial Intelligence

🏢 Enterprise

Cloud Operations Engineer managing and troubleshooting customer Neo4j database infrastructure. Supporting deployments, monitoring, upgrades, and incidents across AWS, Azure, Google Cloud, virtual, and bare-metal environments.

Ansible

AWS

Azure

Cloud

Google Cloud Platform

Kubernetes

Linux

Neo4j

Prometheus

Terraform

🕒 July 29

Outmarket AI

11 - 50

🤖 Artificial Intelligence

🛡️ Insurance

☁️ SaaS

DevOps Engineer managing infrastructure and delivery platform for AI products. Focus on security, reliability, and observability while working in an AI-first environment.

Airflow

AWS

Azure

Cloud

Google Cloud Platform

Kubernetes

Python

Terraform

🕒 July 29

Granicus

501 - 1000

🏛️ Government

☁️ SaaS

📋 Compliance

Site Reliability Engineer 3 modernizing reliability engineering for Granicus with a focus on AIOps and automation. Improve service reliability and build scalable, resilient platforms for various workloads.

Ansible

AWS

Azure

Cloud

Distributed Systems

ElasticSearch

Google Cloud Platform

ITSM

Kubernetes

Linux

Logstash

Terraform

Unix

🕒 July 28

Sezzle

201 - 500

💳 Fintech

👥 B2C

🛍️ eCommerce

Senior Site Reliability Engineer at Sezzle resolving infrastructure challenges and enhancing reliability through scalable solutions. Seeking innovative and experienced candidates to drive technical excellence.

AWS

Distributed Systems

Grafana

Kubernetes

Microservices

MySQL

Postgres

Prometheus

RDBMS

SQL

Go

🕒 July 28

fal

51 - 200

🤖 Artificial Intelligence

🔌 API

☁️ SaaS

Machine Learning Engineer focusing on the reliability and security of generative media model APIs at fal. Working with cutting-edge models and infrastructure in a remote setting.

Distributed Systems