Senior Site Reliability Engineer, SRE

🔥 0 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Mirantis

Mirantis

501 - 1000 employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Consulting • Healthcare • Logistics

Mirantis is a company that specializes in container management and cloud infrastructure solutions. It offers a range of products, including Mirantis Kubernetes Engine (MKE), Mirantis OpenStack for Kubernetes (MOSK), and Mirantis Container Cloud (MCC), which provide enterprise-level Kubernetes and container management platforms. Mirantis also develops tools for secure software supply chains, such as the Mirantis Container Runtime (MCR) and Mirantis Secure Registry (MSR). As an advocate for open source technologies, Mirantis supports various projects and provides resources like Lens Desktop, a popular Kubernetes IDE, and technical support for enterprises adopting cloud-native technologies. Their solutions cater to sectors such as public services, financial services, and broader SaaS and technology services industries.

📋 Description

• Work with geographically distributed international teams on technical challenges and process improvements • Develop, implement, maintain, and troubleshoot cloud and AI infrastructure solutions based on open source software • Deploy AI infrastructure built on NVIDIA-certified hardware according to engineering architecture and implementation designs • Collaborate with stakeholders to gather and refine technical requirements • Optimize system performance, reliability, and scalability • Troubleshoot, debug, and resolve complex technical issues • Participate in code reviews to maintain high quality standards • Stay current with cloud operations and development trends and best practices • Design and implement AI-driven automation across the DevOps lifecycle, including code development and maintenance • Facilitate knowledge transfer to customers during delivery phases • Mentor team members and Mirantis customers • Work with stakeholders to define technical strategies and ensure seamless integration of cloud and software services

🎯 Requirements

• 5+ years of professional experience in DevOps, focused on cloud and infrastructure technologies • Experience with Kubernetes and/or OpenStack • Experience with high-performance data center processing, networking, and storage • Exposure to Golang and working knowledge of Python and JavaScript • Strong knowledge of distributed systems, microservices architecture, and CI/CD pipelines • Problem-solving and debugging skills across networking, storage, Linux, and Kubernetes • Knowledge of performance optimization and security • Ability to lead technical tasks and collaborate with diverse teams • Ability to make independent judgment calls when working directly with customers • Excellent written and spoken English • Excellent customer-facing communication skills • Commitment to innovation, continuous learning, and high-quality results • Ability to travel up to 25%, including internationally • Bachelor's degree in Computer Science or related field, or equivalent experience • At least 5 years of DevOps or Software Development experience or similar experience • Nice to have: network and/or storage architecture experience • Nice to have: high-performance computing or GPU infrastructure experience, including GPU scheduling, MIG/vGPU, RDMA/RoCE or InfiniBand, NVLink, DCGM health-checking, GPU driver/firmware lifecycle, or NVIDIA AI Enterprise • Nice to have: open source community presence, upstream contributions, or conference presentations • Nice to have: experience with Rancher, OpenShift, or VMware

🏖️ Benefits

• Professional development and training • Attend conferences and working groups • Company outings, happy hours, hackathons, and tech talks • Competitive compensation package with a strong benefits plan • Opportunity to work with passionate, talented colleagues • Open-source innovation environment • High-energy environment valuing openness, collaboration, risk-taking, and continuous growth

Apply Now

Similar Jobs

🕒 4 days ago

Point Wild (Formerly Pango Group)

51 - 200

🔒 Cybersecurity

☁️ SaaS

🏢 Enterprise

Lead DevOps Engineer architecting and securing Point Wild’s GCP infrastructure for cybersecurity solutions. Automating Terraform, GitOps, CI/CD, networking, observability, and reliable production operations.

BigQuery

Cloud

ElasticSearch

Google Cloud Platform

Jenkins

Postgres

SQL

Terraform

🕒 August 4

DraftKings Inc.

1001 - 5000

🎮 Gaming

⚽ Sports

👥 B2C

Senior SRE scaling DraftKings’ cloud and on-premise infrastructure for digital sports entertainment and gaming. Building automation, observability, deployment platforms, and reliability standards across distributed systems.

Ansible

Chef

Cloud

Distributed Systems

DNS

Kubernetes

Python

Ruby

Terraform

Go

🕒 July 28

Point Wild (Formerly Pango Group)

51 - 200

🔒 Cybersecurity

☁️ SaaS

🏢 Enterprise

Lead DevOps Engineer managing cloud infrastructure and delivery strategies at Point Wild. Working with GCP to ensure reliability, scalability, and security of production environments.

Cloud

Google Cloud Platform

Kubernetes

Terraform

🕒 June 17

CluneTech

1001 - 5000

💼 Consulting

📣 Marketing

📦 Logistics

Senior AWS DevOps Engineer managing AWS infrastructure and databases for financial technology applications. Implementing CI/CD pipelines and contributing to application modernization efforts.

AWS

Cloud

EC2

Linux

Postgres

SQL

Terraform

🕒 June 15

Perkbox

201 - 500

💼 Consulting

🏥 Healthcare

📣 Marketing

Senior DevSecOps Engineer improving security practices across the Perkbox platform. Collaborating with teams to ensure secure delivery and managing cloud architecture.

🇧🇬 Bulgaria – Remote

💵 €4.3k / month

💰 Venture Round on 2019-04

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

AWS

Azure

Cloud

Google Cloud Platform

Grafana

Jenkins

Kubernetes

Prometheus

Python

SDLC

Splunk

Terraform