Platform Engineer

🔥 0 minutes ago

🌐 Spain, Portugal, +3 more countries – Remote

infoinfo

⏰ Full Time

🟡 Mid-level

🟠 Senior

🏗️ Platform Engineer

👻 Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of CloudLinux

CloudLinux

51 - 200 employees

Founded 2009

☁️ SaaS

🔐 Security

🌐 Web 3

SaaS • Security • Web 3

CloudLinux is a leading provider of operating systems designed specifically for web hosting environments. The company offers a line of products including CloudLinux OS Legacy, CloudLinux OS Shared Pro, and CloudLinux OS Solo, each tailored to improve server stability, security, and performance. With features such as kernel live patching, advanced automation and monitoring tools, and specialized WordPress optimization, CloudLinux helps hosting companies maximize security and profitability while ensuring stable server environments. Over 4,000 companies trust CloudLinux to power millions of websites worldwide, benefitting from increased stability and reduced churn rates. The company emphasizes compatibility with major hosting control panels and CentOS, offering solutions for shared hosts, agencies, and small businesses.

📋 Description

• Own an agreed set of platform services and keep them reliable within agreed service levels. • Operate the observability platform, onboard teams, monitor cost and capacity, and maintain alerting. • Operate GitLab and the CI runner fleet, including upgrades, capacity, access, backups, and restore drills. • Maintain other services with production monitoring and runbooks. • Deploy requested services from scratch using appropriate designs, infrastructure as code, monitoring, backups, and documentation. • Handle developer requests involving access, onboarding, pipeline problems, exporters, and dashboards; convert recurring requests into self-service. • Respond to incidents, diagnose and mitigate impact, restore services safely, conduct root-cause analyses and post-mortems, and implement prevention or detection improvements. • Ship changes as code reviewed in merge requests and plan and verify each change. • Write runbooks, onboarding guides, maintenance notices, and actionable status updates for engineers outside the team. • Work with AI agents by delegating collection and drafting, reviewing their output, and recording learnings for the team. • Collaborate with product and engineering teams on scope, priority, timing, and platform needs.

🎯 Requirements

• Senior-level experience in infrastructure, platform or site reliability engineering, including responsibility for keeping at least one production service operational. • Linux systems administration and debugging on bare metal and virtual machines. • Kubernetes in production delivered through GitOps, including personally performing cluster upgrades. • Infrastructure as code using Ansible and Terraform or OpenTofu, with changes reviewed in merge requests. • GitLab administration and GitLab CI in production, self-hosted or SaaS; equivalent depth with another CI system is acceptable. • Working knowledge of Prometheus and Grafana, including operating them for a team, writing alert rules and dashboards, and reading PromQL. • Ability to write technical explanations for engineers outside the team, including runbooks, notices, and request responses. • Strong communication and interpersonal skills. • Advanced use of AI engineering assistants such as Claude and Codex. • Ability to provide context, break down tasks, design agent loops, and delegate plans for unattended end-to-end execution within defined scope and permissions. • Ability to explain, debug, and test resulting automation and verify generated commands, scripts, and conclusions before production use. • English at upper-intermediate level or higher. • Nice to have: alerting design, SLOs, burn-rate alerts, and data-sized thresholds. • Nice to have: Kata Containers, Firecracker, or gVisor. • Nice to have: S3-compatible object storage operations such as Ceph RGW. • Nice to have: AWS with real cost work. • Nice to have: self-hosted Sentry or another Kafka, ClickHouse, and Redis-backed application operated under load. • Nice to have: Python or Go for exporters and small internal services.

🏖️ Benefits

• A focus on professional development. • Interesting and challenging projects. • Fully remote work with flexible working hours, allowing work from any location worldwide. • Paid 24 days of vacation per year. • 10 days of national holidays. • Unlimited sick leaves. • Compensation for private medical insurance. • Co-working and gym/sports reimbursement. • Budget for education. • Opportunity to receive a reward for the most innovative idea that the company can patent.

Apply Now

Similar Jobs

🕒 August 31

TheWhiteam

201 - 500

💼 Consulting

📦 Logistics

📣 Marketing

Platform Engineer administrando plataformas Azure, Databricks y Fabric para consultora tecnolĂłgica internacional. Liderando gobernanza, redes, seguridad y entrega DevOps de entornos cloud y de datos.

🗣️🇪🇸 Spanish Required

Azure

Cloud

DNS

Unity

🕒 August 28

VIAVI Solutions

1001 - 5000

🚘 Automotive

💼 Consulting

🎖️ Defense

Platform Engineer deploying scalable software and infrastructure for VIAVI, a telecom network equipment and systems provider. Designing highly available solutions across cloud and bare-metal environments.

Ansible

Azure

Docker

Java

Kubernetes

Linux

OpenStack

Python

Terraform

Go

🕒 August 27

Buxton Resources

1 - 10

💸 Finance

🛍️ eCommerce

🤝 B2B

SRE Platform Engineer building AWS, Kubernetes, and Terraform infrastructure for Audiense’s analytics-to-action platform. Automating internal tooling, improving reliability, and reducing developer cognitive load.

AWS

Cloud

Kubernetes

Terraform

🕒 August 19

Hotelbeds

1001 - 5000

🏨 Hospitality

📦 Logistics

✈️ Travel

Platform Engineer managing AWS, Kubernetes, and Terraform infrastructure for HBX Group’s global travel technology platform. Supporting high-volume car rental integrations, observability, security, and disaster recovery.

AWS

Grafana

Kubernetes

Node.js

Postgres

Prometheus

Python

Terraform

🕒 August 18

Ledgy

51 - 200

💸 Finance

💳 Fintech

☁️ SaaS

Engineering Manager building Ledgy’s AI agent and developer platform for equity management and financial reporting. Leading platform engineers, CI/CD, infrastructure, observability, and security strategy for regulated fintech software.

Cloud

Docker

Google Cloud Platform

JavaScript

Kubernetes

MongoDB

Node.js

React

Terraform

TypeScript