Site Reliability Engineer – SRE

🔥 0 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Software Mind

Software Mind

1001 - 5000 employees

Founded 1999

🤖 Artificial Intelligence

☁️ SaaS

📡 Telecommunications

💰 Private Equity Round on 2020-12

Artificial Intelligence • SaaS • Telecommunications

Software Mind is a technology company that specializes in software development and digital transformation services. With a focus on AI and cloud solutions, the company offers a wide range of services including custom software development, mobile app development, and cloud consulting. Software Mind serves various industries such as financial services, telecom, biotech, and media, providing tailored solutions to accelerate digital transformations and business growth globally.

📋 Description

• Support the deployment, operation, and ongoing maintenance of the Karuna service running on Kubernetes • Monitor production environments to ensure high availability, reliability, and performance • Investigate, troubleshoot, and resolve production incidents, performing root cause analysis where appropriate • Analyze application logs and debug production issues using Splunk • Perform first-level troubleshooting of UI-related issues involving Web Components, collaborating with front-end engineers when deeper investigation is required • Support deployment activities and contribute to maintaining and improving CI/CD pipelines • Identify opportunities for automation and operational improvements • Work closely with engineering teams and technical stakeholders in an international environment to improve operational processes, enhance service resilience, and optimize observability • Take ownership of operational tasks and proactively drive issues to resolution while working independently with minimal supervision

🎯 Requirements

• Commercial experience as a Site Reliability Engineer, DevOps Engineer, Platform Engineer, or in a similar role • Hands-on experience supporting production services running on Kubernetes • Experience monitoring distributed applications and responding to production incidents • Practical knowledge of log analysis and troubleshooting using Splunk or similar monitoring tools • Understanding of cloud-native applications and modern operational practices • Familiarity with CI/CD pipelines and deployment processes • Basic understanding of Web Components and the ability to perform first-level UI troubleshooting • Strong analytical and problem-solving skills • Ability to work independently, prioritize tasks, and make sound technical decisions in ambiguous situations • Excellent communication skills and confidence collaborating with distributed engineering teams • Very good spoken and written English • Experience with one or more major cloud platforms • Experience supporting enterprise SaaS or security-focused products

🏖️ Benefits

• Flexible employment and remote work • International projects with leading global clients • International business trips • Non-corporate atmosphere • Language classes • Internal & external training • Private healthcare and insurance • Multisport card • Well-being initiatives

Apply Now

Similar Jobs

🔥 11 hours ago

Crystal Intelligence

51 - 200

💼 Consulting

⚖️ Legal

📦 Logistics

Middle DevOps Engineer managing Kubernetes, cloud, and hybrid infrastructure for Crystal Intelligence’s blockchain analytics platform. Supporting monitoring, deployment, security, networking, and production reliability across distributed teams.

Ansible

AWS

Cloud

Firewalls

Google Cloud Platform

Kubernetes

Linux

NoSQL

Python

SaltStack

SQL

Terraform

Unix

🔥 13 hours ago

Sigma Software Group

1001 - 5000

💼 Consulting

🏥 Healthcare

🚘 Automotive

AI Deployment Engineer building production AI systems for Sigma Software’s intelligent technology solutions. Designing LLM workflows, RAG pipelines, agentic systems, and custom data pipelines for customers.

🔥 17 hours ago

Akamai Technologies

5001 - 10000

🔒 Cybersecurity

Senior SRE maintaining Akamai's distributed Compute cloud infrastructure. Improving observability, automation, performance, and uptime across cloud interfaces and APIs.

Ansible

Cloud

Docker

Grafana

HAProxy

Jenkins

Linux

NGINX

Prometheus

Python

Redis

SaltStack

Terraform

Go

🕒 Yesterday

intive

1001 - 5000

💼 Consulting

🏥 Healthcare

📣 Marketing

DevOps Engineer operating AWS, Kubernetes, and Terraform infrastructure for intive, an AI-native software engineering company. Supporting reliable digital product delivery, observability, automation, and cloud deployments.

Ansible

AWS

Cloud

Docker

Grafana

Kubernetes

Linux

Postgres

Prometheus

Python

Shell Scripting

Terraform

TypeScript

🕒 Yesterday

Creatio

501 - 1000

💼 Consulting

📣 Marketing

📦 Logistics

DevOps Engineer automating CI/CD pipelines and internal tools for Creatio, an AI CRM and workflow platform. Building monitoring, alerting, and fault-tolerance systems for engineering and security teams in Poland.

Ansible

AWS

Azure

Docker

Grafana

Groovy

Jenkins

Kubernetes

Python

SQL

Subversion

Terraform

Unix