Senior Site Reliability Engineer

🔥 1 minute ago

🇮🇳 India – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Level AI

Level AI

51 - 200 employees

Founded 2018

🤖 Artificial Intelligence

☁️ SaaS

🏢 Enterprise

💰 $39.4M Series C - Level AI on 2024-07

Artificial Intelligence • SaaS • Enterprise

Level AI is a customer experience (CX) platform that uses proprietary AI models to analyze and automate contact center interactions. It scores 100% of customer interactions, deploys AI voice agents, provides real-time agent assist and coaching, automates quality assurance (Auto-QA), and delivers voice-of-the-customer analytics. Level AI Latitude comprises seven task-specific models designed for enterprise CX, claiming improved cost-efficiency and speed versus general-purpose LLMs. The company targets high-volume, regulated contact centers in industries such as financial services, healthcare, retail, and insurance, and emphasizes security and compliance (ISO 27001, SOC2, HIPAA, PCI, GDPR).

📋 Description

• Own continued reduction of Kubernetes overprovisioning • Drive infrastructure right-sizing programs • Maintain cost telemetry for backend team decision-making • Run structured experimentation on on-premise GPU clusters in partnership with AI service owners • Build tooling, dashboards, and processes enabling backend teams to own their cost and reliability budgets • Ensure infrastructure surface area is properly instrumented for cost-at-scale and reliability • Take on defined security workstreams involving platform-security changes • Provide hands-on support across backend engineering, infrastructure operations, and FinOps

🎯 Requirements

• 4-5 years of hands-on systems experience • Production experience in Python, Go/Rust • Comfortable owning services end to end and reasoning about backend code across teams • Kubernetes at scale: scheduler behaviour, resource requests/limits, HPA/VPA, node pool design, cost-aware autoscaling (Cast AI, Karpenter, or equivalent) • GCP fluency • Infrastructure as code with Terraform • CI/CD experience • Comfort operating in hybrid setups, including on-prem GPU clusters • Familiarity with throughput profiling, batching, KV-cache behavior, inference server tuning, and GPU utilisation metrics • Experience with metrics, traces, logs, SLOs, and proper systems instrumentation • Demonstrated history of converting infrastructure choices into measurable cost outcomes • Ability to take on platform-security workstreams without constant handoff to the DevOps team • Must be able to work in the EST time zone

Apply Now

Similar Jobs

🔥 12 hours ago

Miratech

501 - 1000

🤝 B2B

💼 Consulting

☁️ SaaS

DevOps Engineer designing highly available Kubernetes infrastructure for Miratech’s global IT services and consulting clients. Managing event-driven systems, GitOps, cloud platforms, databases, and observability.

AWS

Cloud

Consul

Distributed Systems

Flux

Grafana

Kafka

Kubernetes

Linux

MongoDB

MySQL

NGINX

Postgres

Prometheus

RabbitMQ

Redis

Terraform

Go

🔥 13 hours ago

Miratech

501 - 1000

🤝 B2B

💼 Consulting

☁️ SaaS

DevOps Engineer building highly available Kubernetes infrastructure for Miratech’s global IT services and consulting clients. Managing cloud-native, event-driven systems across distributed environments.

AWS

Cloud

Consul

Distributed Systems

Flux

Grafana

Kafka

Kubernetes

Linux

MongoDB

MySQL

NGINX

Postgres

Prometheus

RabbitMQ

Redis

Terraform

Go

🔥 13 hours ago

Miratech

501 - 1000

🤝 B2B

💼 Consulting

☁️ SaaS

DevOps Engineer architecting Kubernetes and cloud-native infrastructure for Miratech’s global IT services clients. Managing event-driven systems, GitOps deployments, databases, and observability.

AWS

Cloud

Consul

Distributed Systems

Flux

Grafana

Kafka

Kubernetes

Linux

Microservices

MongoDB

MySQL

Postgres

Prometheus

RabbitMQ

Redis

Terraform

Go

🔥 20 hours ago

Miratech

501 - 1000

🤝 B2B

💼 Consulting

☁️ SaaS

Senior DevOps Engineer leading enterprise GitHub repository and CI/CD migrations for Miratech, a global IT services and consulting company. Automating platforms across GitHub, Terraform, AWS/EKS, Kubernetes, and Jenkins.

AWS

Cloud

Jenkins

Kubernetes

Python

Terraform

🔥 20 hours ago

Miratech

501 - 1000

🤝 B2B

💼 Consulting

☁️ SaaS

Senior Observability/DevOps Engineer migrating enterprise Datadog monitoring platforms for global IT services company. Automating observability across AWS and Kubernetes using Terraform and Datadog APIs.

AWS

Kubernetes

ServiceNow

Terraform