Senior Site Reliability Engineer

Job not on LinkedIn

🕒 June 10

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of The Leaflet

The Leaflet

11 - 50 employees

💼 Consulting

⚖️ Legal

🔌 API

Consulting • Legal • API

The Leaflet is an open-source JavaScript library for building mobile-friendly interactive maps. It is lightweight (around 42 KB), designed for simplicity, performance and usability, and provides core mapping features such as tile layers, markers, vector layers, popups, and interaction handlers. Leaflet is highly extensible via a large plugin ecosystem, well-documented, and maintained by a broad community of contributors and organizations.

📋 Description

• Ensure the availability, reliability, and performance of high-traffic Java-based applications in a distributed environment • Troubleshoot and resolve complex issues across production and non-production environments • Participate in pre- and post-deployment performance testing and monitoring to continuously improve application performance • Design, build, and operate agentic AI workflows that automate operational tasks such as alert triage and root cause analysis

🎯 Requirements

• Degree in Computer Science or related field, or equivalent professional experience • 5+ years in SRE, DevOps, or similar infrastructure roles with experience managing large-scale, high-availability production systems • 3+ years hands-on experience managing production Kubernetes clusters, including deep understanding of architecture, networking, storage, and security • Advanced expertise with the Grafana observability stack: dashboards, alerting, visualization, and Grafana Alloy for telemetry collection • Strong scripting abilities in Python, Bash, or Go, with experience building CI/CD pipelines and deployment automation • 1+ years of practical experience building or operating AI/LLM-powered tools, agents, or workflows

🏖️ Benefits

• Fully remote position • Opportunity to work with cutting-edge AI tools • Collaborative team environment

Apply Now

Similar Jobs

🕒 June 10

Recruiting.com

11 - 50

🎯 Recruiter

☁️ SaaS

🤝 B2B

Site Reliability Engineer focusing on maintaining high service levels and monitoring production environments at Cencora. Collaborating with Development and DevOps to enhance global product platform reliability.

Azure

Kubernetes

MySQL

Python

Terraform

Go

🕒 June 9

MARGO

201 - 500

💼 Consulting

🛡️ Insurance

📦 Logistics

Network Reliability Engineer for building AI infrastructure with monitoring and production incident remediation. Collaborating with teams on high-impact production issues in a remote work setup.

Ansible

DNS

Grafana

Linux

MariaDB

Prometheus

Python

SaltStack

TCP/IP

Go

🕒 June 8

Netguru

501 - 1000

💼 Consulting

🏥 Healthcare

📣 Marketing

Regular DevOps Engineer working remotely on projects for various industries. Collaborating with experienced developers at Netguru to modernize digital commerce solutions.

Grafana

Kafka

Kubernetes

Postgres

🕒 June 8

Netguru

501 - 1000

💼 Consulting

🏥 Healthcare

📣 Marketing

Senior DevOps Engineer at Netguru managing diverse projects remotely. Collaborating as part of an experienced team with flexibility over hours and tasks.

Grafana

Kafka

Kubernetes

Postgres

🕒 June 4

GoReel

51 - 200

💼 Consulting

📣 Marketing

🎮 Gaming

SRE Lead responsible for designing, implementing, and maintaining cloud infrastructure in the iGaming industry. Collaborating with development teams to ensure system reliability and streamline deployment processes.

AWS

Cloud

Docker

EC2

ElasticSearch

Grafana

Jenkins

Kubernetes

Prometheus

Python