Senior Site Reliability Engineer – Kubernetes

Job not on LinkedIn

🕒 August 17

🇵🇱 Poland – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 24%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Software Mind

Software Mind

1001 - 5000 employees

Founded 1999

🤖 Artificial Intelligence

☁️ SaaS

📡 Telecommunications

💰 Private Equity Round on 2020-12

Artificial Intelligence • SaaS • Telecommunications

Software Mind is a technology company that specializes in software development and digital transformation services. With a focus on AI and cloud solutions, the company offers a wide range of services including custom software development, mobile app development, and cloud consulting. Software Mind serves various industries such as financial services, telecom, biotech, and media, providing tailored solutions to accelerate digital transformations and business growth globally.

📋 Description

• Support the deployment, operation, and reliability of production services running on Kubernetes • Monitor service health and investigate production incidents across distributed applications • Participate in on-call support, incident response, root cause analysis, postmortems, and reliability improvements • Troubleshoot application runtime, networking, and service-to-service issues with engineering teams • Support CI/CD, GitOps-based deployments, observability, and production monitoring • Work within a client-directed backlog and established priorities • Own production reliability for the AI Experience Framework stack end to end, including Kubernetes, observability, and troubleshooting Node.js and JVM systems

🎯 Requirements

• 5+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, Production Engineering, or a closely related role • Strong recent hands-on experience supporting Kubernetes-based production services • 3+ years of hands-on production Kubernetes experience strongly preferred • Kubernetes production operations, including deployment, scaling, rollout/rollback, resource tuning, and service-to-service troubleshooting • Strong production incident response experience, including on-call, runbooks, postmortems, and paging hygiene • Splunk experience for log aggregation, search, and production troubleshooting • Prometheus and Grafana experience building alert rules and dashboards • CI/CD and infrastructure-as-code for containerized deployments, including Helm and GitOps tools such as ArgoCD or Flux • Strong Linux and networking fundamentals, including DNS, load balancing, TCP/HTTP, HTTP/2, and Kubernetes networking • Production troubleshooting across Node.js and JVM/Java services, with strong depth in at least one runtime environment • Service-to-service authentication experience, including mTLS, certificate rotation, certificate format conversion, and JWT-based authentication • Very good spoken and written English • Web Components/Lit experience, server-side rendering or isomorphic runtime experience, canary rollout/multi-version production operations, distributed tracing, KEDA or event-driven autoscaling, and enterprise platform integration experience are additional skills

🏖️ Benefits

• Flexible employment and remote work • International projects with leading global clients • International business trips • Non-corporate atmosphere • Language classes • Internal & external training • Private healthcare and insurance • Multisport card • Well-being initiatives

Apply Now

Similar Jobs

🕒 August 13

Railsware

201 - 500

💼 Consulting

📣 Marketing

💳 Fintech

Senior DevOps Engineer operating Mailtrap's multi-region AWS email platform. Leading partial migration to rented bare metal while maintaining reliability, security, and cost efficiency.

Ansible

AWS

Cloud

DNS

Docker

Firewalls

Google Cloud Platform

Grafana

HAProxy

Kafka

Kubernetes

Linux

NGINX

Packer

Postgres

Prometheus

Python

Redis

Ruby

SMTP

Terraform

Go

🕒 August 12

Sigma Software Group

1001 - 5000

💼 Consulting

🏥 Healthcare

🚘 Automotive

AI Deployment Engineer building production AI systems for Sigma Software’s intelligent technology solutions. Designing LLM workflows, RAG pipelines, agentic systems, and custom data pipelines for customers.

🕒 August 12

Akamai Technologies

5001 - 10000

🔒 Cybersecurity

Senior SRE maintaining Akamai's distributed Compute cloud infrastructure. Improving observability, automation, performance, and uptime across cloud interfaces and APIs.

Ansible

Cloud

Docker

Grafana

HAProxy

Jenkins

Linux

NGINX

Prometheus

Python

Redis

SaltStack

Terraform

Go

🕒 August 11

intive

1001 - 5000

💼 Consulting

🏥 Healthcare

📣 Marketing

DevOps Engineer operating AWS, Kubernetes, and Terraform infrastructure for intive, an AI-native software engineering company. Supporting reliable digital product delivery, observability, automation, and cloud deployments.

Ansible

AWS

Cloud

Docker

Grafana

Kubernetes

Linux

Postgres

Prometheus

Python

Shell Scripting

Terraform

TypeScript

🕒 August 11

FYUL

1001 - 5000

🛍️ eCommerce

🏭 Manufacturing

🤝 B2B

Senior SRE operating AWS, Kubernetes, and observability platforms for FYUL’s global on-demand commerce. Driving automation, reliability, security, and cost optimization across engineering teams.

AWS

Cloud

Google Cloud Platform

Grafana

Jenkins

Kafka

Kubernetes

Linux

MongoDB

MySQL

Postgres

Prometheus

Python

Terraform