Senior Site Reliability Engineer – IT Developer Enablement Applications

🕒 August 4

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Red Hat

Red Hat

10,000+ employees

Founded 1993

🏢 Enterprise

💰 Corporate Round on 1999-03

Enterprise • Cloud

Red Hat is a leading provider of enterprise open source software solutions, helping companies worldwide to build and deploy applications across hybrid cloud infrastructures. With a strong focus on developing secure, stable, and innovative technologies, Red Hat offers a broad portfolio including products like Red Hat Enterprise Linux, Red Hat OpenShift, and Red Hat Ansible Automation Platform. These products support IT services on any infrastructure efficiently. Trusted by more than 90% of the U. S. Fortune 500, Red Hat empowers organizations to modernize their IT environments, leveraging open source communities to drive technological advancement.

📋 Description

• Execute the full software development life cycle within an agile environment • Design, implement, and maintain scalable, reliable, and highly available infrastructure systems for business applications and services • Automate repetitive tasks and processes using Python, Bash, Terraform, Ansible, and similar tools • Manage OpenShift cloud infrastructure and optimize resources for reliability, performance, and security • Monitor system performance, troubleshoot issues, and ensure uptime through alerting, metrics, and incident response • Collaborate with developers to improve application resilience, deploy code efficiently, and integrate CI/CD pipelines • Conduct root cause analysis for incidents, document findings, and implement preventive measures • Improve system architecture, deployment processes, and tooling to enhance operational efficiency and scalability • Apply security best practices across infrastructure, including access controls, network configurations, and compliance requirements • Communicate with stakeholders and project team members to provide visibility into development efforts • Respond to customer tickets and inquiries

🎯 Requirements

• 5+ years of experience with Linux Systems Engineering, including RHEL and CentOS • 5+ years of experience delivering enterprise software applications • 2+ years of experience with Git; GitLab experience is a plus • Strong knowledge of Linux and networking internals • Strong programming proficiency in Python and experience with the full SDLC • Experience with containers and orchestration tools such as Docker, Podman, Kubernetes, and OpenShift • Good knowledge of CI/CD pipelines and tooling such as GitLab CI and GitHub Actions • Experience with configuration management tools such as Ansible • Familiarity with monitoring and logging suites such as Sumo Logic, Prometheus, or ELK • Familiarity with SRE fundamentals and principles • Familiarity with bare metal environments and APIs such as RedFish • Data-driven and observability-first mindset involving logs, metrics, and traces

🏖️ Benefits

• Flexible work environments, depending on role requirements, including fully remote work • Reasonable accommodations for job applicants with disabilities

Apply Now

Similar Jobs

🕒 August 4

Arista Networks

1001 - 5000

🏢 Enterprise

📡 Telecommunications

Senior Site Reliability Engineer building and operating scalable, secure production infrastructure for Arista Networks’ cloud networking platforms. Automating operations, improving observability, and ensuring reliable deployments from Ireland.

AWS

Azure

Cloud

Distributed Systems

Docker

Google Cloud Platform

Grafana

Kubernetes

Linux

Postgres

Prometheus

Python

Shell Scripting

Spinnaker

Terraform

Unix

Go

🕒 August 3

Arista Networks

1001 - 5000

🏢 Enterprise

📡 Telecommunications

Site Reliability Engineer at Arista Networks focusing on building and operating critical production systems for scalability and reliability. Engaging in a collaborative remote role from Ireland.

AWS

Azure

Cloud

Distributed Systems

Docker

Google Cloud Platform

Grafana

Kubernetes

Linux

Postgres

Prometheus

Python

Shell Scripting

Spinnaker

Terraform

Unix

Go

🕒 July 28

Astreya

1001 - 5000

💼 Consulting

📦 Logistics

📣 Marketing

IT Infrastructure Support Engineer optimizing critical physical security systems for a global IT provider. Engineering automation tools for a reliable and scalable infrastructure environment.

Ansible

Chef

Cloud

Grafana

IoT

Kubernetes

Linux

Prometheus

Puppet

Python

Terraform

Go

🕒 July 27

Sardine

51 - 200

🔒 Cybersecurity

📋 Compliance

💳 Fintech

DevOps Engineer at Sardine improving infrastructure and tooling for a remote-first financial crime platform. Collaborating to ensure reliable, scalable, and cost-efficient systems.

AWS

Cloud

Distributed Systems

Google Cloud Platform

Kubernetes

Prometheus

Python

Terraform

Go

🕒 July 27

Holafly

501 - 1000

✈️ Travel

📡 Telecommunications

👥 B2C

DevSecOps Engineer at Holafly, designing secure GCP foundations and automating with Terraform. Safeguarding connectivity for millions of travelers through infrastructure and security enhancements.

Ansible

Cloud

Docker

Google Cloud Platform

Kubernetes

Python

Terraform