
51 - 200 employees
Founded 2022
☁️ SaaS
🤖 Artificial Intelligence
🏢 Enterprise
SaaS • Artificial Intelligence • Enterprise
Kestra is an open-source orchestration platform for data, AI, and infrastructure workflows. It is event-driven, language-agnostic, and built for enterprise scale, offering declarative YAML workflows, an API-first design, CI/CD-native features, and 1800+ plugins to integrate cloud, data, infra and SaaS tools. Kestra ships as self-hosted (Docker/Kubernetes) and a managed cloud edition, and provides enterprise capabilities (SSO, RBAC, audit logs, multi-tenancy, isolated workers) alongside AI-native features like RAG pipelines, agents, and a Copilot to operationalize and govern AI workflows.
🕒 July 27
🇳🇱 Netherlands – Remote
⏰ Full Time
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
👻 Ghost score 26%
Improve your chances of getting an interview by checking your resume score before you apply.

51 - 200 employees
Founded 2022
☁️ SaaS
🤖 Artificial Intelligence
🏢 Enterprise
SaaS • Artificial Intelligence • Enterprise
Kestra is an open-source orchestration platform for data, AI, and infrastructure workflows. It is event-driven, language-agnostic, and built for enterprise scale, offering declarative YAML workflows, an API-first design, CI/CD-native features, and 1800+ plugins to integrate cloud, data, infra and SaaS tools. Kestra ships as self-hosted (Docker/Kubernetes) and a managed cloud edition, and provides enterprise capabilities (SSO, RBAC, audit logs, multi-tenancy, isolated workers) alongside AI-native features like RAG pipelines, agents, and a Copilot to operationalize and govern AI workflows.
• Architect, build, and scale the infrastructure for Kestra’s SaaS platform. • Automate cloud resource provisioning and management to enhance reliability and efficiency. • Optimize monitoring, deployment, and repair systems to ensure high availability. • Improve and maintain CI/CD pipelines for smooth, reliable deployments. • Troubleshoot and resolve infrastructure and application issues promptly. • Drive innovation with solutions that enhance engineering productivity. • Continuously improve operational metrics like deployment speed and system reliability.
• 7+ years of experience in DevOps, SRE, or similar infrastructure-focused roles. • Deep expertise in Kubernetes and Terraform, with hands-on experience in GCP or AWS. • Strong problem-solving skills, with a focus on automation and scalability. • Fluent in English and comfortable working in a fully remote environment. • Experience managing and monitoring distributed systems and databases such as Kafka, PostgreSQL, and Elasticsearch. • Adaptability to a startup environment, where delivering impactful features quickly is a priority.
• Work from anywhere: We’re a remote-first company, so you can work from wherever feels like home. Plus, you’ll have access to coworking spaces worldwide if you ever need a change of scenery. • Health coverage: From medical support, dental, and vision, we've got you covered. • Home office setup on us: We’ll provide all the equipment you need to work comfortably.
Apply Now🕒 July 27
Senior DevOps Engineer at Publitas, a remote-first SaaS company. Responsible for AWS/GCP infrastructure and handling high-severity incidents with a focus on operational tasks.
AWS
Google Cloud Platform
🕒 July 7
Senior Site Reliability Engineer ensuring fault-tolerance and scale for services at Nebius. Involved in hardware automation and CI/CD improvement with a focus on infrastructure.
Linux
Python
🕒 June 8
Senior Site Reliability Engineer maintaining and growing systems for Nebius' AI cloud platform. Collaborating within a fast-paced SRE team to improve user experience.
Java
Kotlin
Python
Ruby
Spring
Unix
Go
🕒 April 2
Senior Site Reliability Engineer at Nebius ensuring fault-tolerance and uninterrupted service operations using cutting-edge cloud technology.
Ansible
Cloud
Docker
Kubernetes
Python
SaltStack
Terraform
Unix
Go
🕒 January 26
Senior Site Reliability Engineer for the Compute Node team at Nebius, ensuring operational reliability and performance of compute nodes across cloud regions.
Linux
Node.js