Senior Platform Operations Engineer / Site Reliability Engineer

Job not on LinkedIn

🔥 0 minutes ago

🇩🇪 Germany – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 12%

infoinfo

🗣️🇩🇪 German Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Netlution GmbH

Netlution GmbH

201 - 500 employees

Founded 2001

💼 Consulting

🏢 Enterprise

🤖 Artificial Intelligence

Consulting • Enterprise • Artificial Intelligence

Netlution GmbH is a German IT consulting and services firm specializing in enterprise service management and IT infrastructure. They deliver tailored IT and application services—especially around ServiceNow—covering Service Operations, AI-enhanced ServiceNow integrations, DevOps, cloud and application engineering, 3rd-level support and transition/EOL services for mid-sized and large companies. Netlution positions itself as a trusted advisor offering on-site and remote expert teams focused on automation, high-quality operations and business-aligned IT solutions.

📋 Description

• Operate, maintain, and continuously develop monitoring and observability platforms • Administer and optimize Prometheus, Grafana, and OpenSearch / ELK • Monitor production environments and continuously improve monitoring and alerting concepts • Handle incidents and support the recovery of critical services • Conduct root cause analyses and permanently eliminate the causes of incidents • Support major incidents and coordinate technical remediation measures • Operate and optimize containerized platforms based on Kubernetes • Support CI/CD processes using Jenkins and ArgoCD • Create, maintain, and continuously improve runbooks, operational processes, and technical documentation • Automate recurring operational tasks • Participate in on-call rotations and shift schedules within a 24/7 operations organization

🎯 Requirements

• Several years of experience in Platform Operations, Site Reliability Engineering, Systems Engineering, or IT Operations • Willingness to undergo, or possession of, a German SÜ2 security clearance • Very good knowledge of Linux-based environments • Understanding of Kubernetes and containerized platforms • Hands-on experience with Prometheus • Hands-on experience with Grafana • Hands-on experience with the ELK Stack or OpenSearch • Experience with Elasticsearch or OpenSearch • Experience in monitoring, alerting, and observability environments • Good knowledge of networking fundamentals and communication protocols • Experience working with REST APIs • Proficiency with Git • Analytical approach to troubleshooting and incident resolution • Good written and spoken English • Willingness to participate in on-call rotations and shift work • Nice to have: Experience with ArgoCD • Nice to have: Knowledge of Jenkins • Nice to have: Experience with Helm • Nice to have: Bash scripting skills • Nice to have: Python experience for operational automation and operational excellence • Nice to have: Experience in Site Reliability Engineering (SRE) • Nice to have: Knowledge of modern cloud or platform architectures

🏖️ Benefits

• Permanent employment contract • Attractive compensation • Professional development and training opportunities • Hybrid and fully remote work options • PC equipment provided

Apply Now

Similar Jobs

🔥 1 minute ago

coeo Group

501 - 1000

💳 Fintech

🤝 B2B

🤖 Artificial Intelligence

Azure Cloud Operations Engineer für coeo, datengetriebenen Forderungsmanagement-Spezialisten. Betrieb, Automatisierung und Weiterentwicklung von Azure- und Windows-Server-Infrastrukturen.

🗣️🇩🇪 German Required

Azure

Cloud

Python

🕒 Yesterday

adorsys

51 - 200

💳 Fintech

🔌 API

🔒 Cybersecurity

DevOps Engineer bei adorsys, einem Nürnberger Technologieunternehmen für Software-, Cloud- und Architekturlösungen. Automatisierung von Infrastruktur, CI/CD und Kubernetes-Deployments für Kundenumgebungen.

🗣️🇩🇪 German Required

Ansible

AWS

Azure

Cloud

Docker

Flux

Kubernetes

OpenShift

Python

Terraform

🕒 2 days ago

Sigma Software Group

1001 - 5000

💼 Consulting

🏥 Healthcare

🚘 Automotive

Forward AI Deployment Engineer redesigning client operations with production AI at Sigma Software. Building agentic systems, RAG, and data pipelines that deliver lasting workflow improvements.

Cloud

🕒 3 days ago

Interlead GmbH

51 - 200

🛡️ Insurance

📦 Logistics

🏥 Healthcare

DevSecOps Engineer securing servers, cloud infrastructure, and CI/CD pipelines for Interlead’s qualified-lead platform. Automating operations and embedding security across the technology stack.

🗣️🇩🇪 German Required

Ansible

AWS

Azure

Cloud

Docker

Firewalls

Google Cloud Platform

Grafana

Kubernetes

Linux

Prometheus

Terraform

🕒 5 days ago

IT42morrow IFT GmbH

1 - 10

💼 Consulting

📦 Logistics

📣 Marketing

DevOps Engineer managing cloud environments, CI/CD pipelines, automation, and security. Supporting diverse customer projects for German IT consulting company IT42morrow.

🗣️🇩🇪 German Required

Ansible

AWS

Azure

Chef

Cloud

Docker

Google Cloud Platform

Grafana

Java

Jenkins

Kubernetes

OpenShift

Prometheus

Puppet

Python

SaltStack

Go