Senior DevOps Engineer

🕒 July 24

🇮🇳 India – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 13%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of HighLevel

HighLevel

201 - 500 employees

Founded 2018

💼 Consulting

📦 Logistics

☁️ SaaS

💰 Series A on 2021-11

Consulting • Logistics • SaaS

HighLevel is an all-in-one marketing and sales platform designed to help businesses grow and succeed. The platform consolidates various marketing tools into a single solution, providing features such as lead capture through landing pages, surveys, forms, and calendars, as well as tools for nurturing leads via automated messaging across multiple channels including phone, SMS, email, and social media. HighLevel offers customizable solutions like online appointment scheduling, multi-channel follow-up campaigns, and pipeline management. Additionally, businesses can build websites, funnels, and landing pages using the intuitive page builder. HighLevel supports integrating with existing systems via API, and offers a membership platform for community building and course management. The platform is targeted towards marketers and offers white-labeling options for businesses to brand the software as their own. With a community-driven development approach and award-winning support, HighLevel is focused on empowering businesses to streamline their operations and enhance their marketing efficiencies.

📋 Description

• Design and implement strategies to reduce operational costs across GCP, AWS, Firebase, and managed services like MongoDB Atlas and Elastic.co • Build and maintain cost dashboards (using DoiT, GCP BigQuery, AWS Cost Explorer, etc.) to track spend across cloud providers and services at the product, team, and project level • Define and maintain Cost per Operation (CPO) metrics per product/service; collaborate with product owners to align infra cost with revenue generation • Set up policies, guardrails, budgets, and alerts to prevent cost overruns and enforce efficient resource usage across Kubernetes clusters and cloud services • Develop automation scripts (using Python, Bash, or Terraform) to detect idle resources, right-size workloads, and enforce tagging strategies • Break down cloud bills by team, sub_team, project, and service using label-based and usage-based filtering; provide actionable insights and recommendations • Partner with finance, DevOps, platform, and product teams to tie infrastructure cost back to product growth and engineering impact • Provide guidance on designing cost-efficient cloud-native systems and help teams adopt best practices for storage, compute, and networking usage

🎯 Requirements

• 4+ years in Cloud Engineering roles with a focus on cost optimization • Deep hands-on experience with GCP (BigQuery, GKE, Firebase, Pub/Sub), AWS (EC2, S3, Lambda, EKS), and managed services like MongoDB Atlas, Elastic.co, ClickHouse • Strong working knowledge of DoiT Cloud Analytics, GCP Billing Export, AWS CUR (Cost & Usage Reports), and BigQuery • Proficient in Python, Bash, and automation frameworks for cost cleanup and reporting • Comfortable querying and visualizing cloud billing data to derive unit economics (e.g., cost per user, per API call, per deployment) • Familiar with cost management in Kubernetes (e.g., node cost allocation, workload optimization, spot/preemptible usage) • Ability to translate complex cost insights into actionable plans for engineering and business stakeholders

🏖️ Benefits

• Health insurance • 401(k) matching • Flexible work hours • Paid time off • Remote work options

Apply Now

Similar Jobs

🕒 July 21

Granicus

501 - 1000

🏛️ Government

☁️ SaaS

📋 Compliance

Senior DevOps Engineer focused on cloud automation and operational reliability for Govtech solutions. Leading technical projects and mentoring engineers in complex application environments.

Cloud

Linux

🕒 July 20

Five9

1001 - 5000

☁️ SaaS

🤖 Artificial Intelligence

📡 Telecommunications

Network Engineer maintaining scalable CI/CD pipelines for a cloud contact center software. Driving automation and network design in 24/7 SaaS environments.

Ansible

Cloud

Firewalls

Jenkins

Kubernetes

Linux

Python

Terraform

VoIP

🕒 July 18

BETSOL

501 - 1000

💼 Consulting

🏥 Healthcare

📦 Logistics

Senior Cloud Engineer developing and operating cloud portal workloads across Azure and GCP using AI-first practices. Collaborating on security and automation in a global enterprise environment.

Ansible

AWS

Azure

Cloud

Google Cloud Platform

JavaScript

Jenkins

Kubernetes

Python

Terraform

TypeScript

🕒 July 16

Akamai Technologies

5001 - 10000

🔒 Cybersecurity

Site Reliability Engineer ensuring performance and reliability of Akamai's media and web delivery platform. Leading investigations into complex reliability and performance across global systems.

Distributed Systems

DNS

Linux

Python

SQL

TCP/IP

Unix

Go

🕒 July 13

MRSOOL | مرسول

201 - 500

🍽️ Food & Beverage

✈️ Travel

💼 Consulting

Site Reliability Engineer II for Mrsool, enhancing infrastructure and supporting development teams in a dynamic environment. Ensuring reliability for a leading delivery platform in the MENA region.

Ansible

AWS

Azure

Chef

Cloud

Distributed Systems

Docker

Google Cloud Platform

Grafana

Java

Kubernetes

Prometheus

Puppet

Python

Ruby

Terraform

Go