Lead Site Reliability Engineer – Cloud

🔥 3 minutes ago

🇫🇷 France – Remote

💵 €60k - €70k / year

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 0%

infoinfo

🗣️🇫🇷 French Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Scalingo

Scalingo

11 - 50 employees

Founded 2015

☁️ SaaS

💰 Debt Financing on 2024-07

SaaS • Cloud Computing • Data Hosting

Scalingo is a European Platform as a Service (PaaS) provider that focuses on delivering secure and reliable cloud hosting solutions. The company enables businesses to deploy, manage, and scale web applications without handling servers, using a wide array of supported programming languages and databases. Scalingo is ISO 27001 and HDS certified, primarily hosting data in Paris, France with strong emphasis on GDPR compliance and data sovereignty.

📋 Description

• Drive the team’s ways of working, including processes, rituals and documentation • Support the team, help prioritize work, and review technical decisions and implementations • Contribute to team members’ professional development and foster autonomy • Promote SRE best practices: reliability, observability, incident management and automation • Champion the SRE technical vision in strategic projects • Analyze performance, identify bottlenecks, and optimize resource usage and scalability • Define, implement and improve observability tools, including monitoring, metrics, logs and alerting • Write, maintain and evolve operational processes • Maintain continuous technology watch on infrastructure practices and technologies • Provide a portion of Level 3 customer support in accordance with SLAs • Participate in incident management and on-call rotations, approximately half a week every three weeks • Respond to critical incidents to minimize their impact and ensure service continuity • Lead incident retrospectives and define sustainable corrective actions • Write and publish post-mortem reports following major incidents • Coordinate crisis communication internally and with customers • Ensure compliance with SLA, RPO and RTO commitments • Establish service quality indicators and SLOs • Contribute to ISO 27001 and HDS compliance, as well as internal and external audits • Plan, execute and analyze business continuity and disaster recovery tests • Collaborate with development teams to integrate operability requirements from the design stage • Advise Product Engineering teams on reliability, customer experience and administration tools • Contribute to clear, up-to-date operational documentation • Provide technical and operational leadership without direct line-management responsibility, reporting to an Engineering Manager

🎯 Requirements

• Strong expertise in cloud environments and distributed infrastructure, with a strong culture of high availability and production reliability • Proficiency in observability practices, including logs, metrics and alerting, with a structured approach to diagnosing complex incidents • Solid understanding of containerized environments and their operational challenges • Proven production database skills, including reliability, backups, restoration, replication and scalability • Experience with Infrastructure as Code and environment automation • Awareness of operational security considerations • Comfortable using Artificial Intelligence tools to improve day-to-day efficiency • Ability to work rigorously and reliably in complex, changing or uncertain environments • Comfortable prioritizing work, including during incidents • Clear and structured communication skills, with a collaborative cross-functional mindset and a passion for knowledge sharing • A blameless mindset, technical curiosity, composure and a strong focus on user impact • Ability to provide technical leadership, share knowledge and advance collective practices • This position must be based exclusively in France

🏖️ Benefits

• Fully remote, with one trip per quarter to Strasbourg or another city • Company events: one annual offsite and regular after-work gatherings • Remote-work allowance (€57.60) • Meal vouchers (€11.52 each) and a Swile card with additional benefits • Flexible working hours under a forfait-hours arrangement, including RTT days • Linux laptop • Budget for additional equipment, with company contribution

Apply Now

Similar Jobs

🕒 August 7

Atos

10,000+ employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Ingénieur DevOps OpenShift administrant des plateformes de production critiques chez Atos, spécialiste des services numériques sécurisés. Industrialisant les déploiements GitOps, CI/CD et l’observabilité.

🗣️🇫🇷 French Required

Ansible

Grafana

Jenkins

Kafka

Kubernetes

Linux

Microservices

OpenShift

Prometheus

Terraform

VMware

🕒 August 1

Dailymotion

201 - 500

📱 Media

👥 B2C

Senior DevOps Engineer focusing on empowering ML Engineers with tools and infrastructure at Dailymotion. Collaborating closely with AI teams in an AI-first environment.

🗣️🇫🇷 French Required

BigQuery

Docker

Google Cloud Platform

Kubernetes

Prometheus

Python

Rust

Terraform

Go

🕒 July 31

BeReal.

51 - 200

👥 B2C

📱 Media

Senior SRE maintaining SRE practices and optimizing infrastructure for BeReal's growth platform. Contributing to observability and developer-friendly solutions in a collaborative environment.

🗣️🇫🇷 French Required

AWS

Cloud

Distributed Systems

Google Cloud Platform

Kubernetes

Terraform

🕒 July 27

FFF - Fédération Française de Football

201 - 500

⚽ Sports

📱 Media

📚 Education

DevOps Engineer at Cloudi-Fi responsible for scalable and secure systems automation. Collaborating in a multicultural team towards safer internet experiences.

🗣️🇫🇷 French Required

Ansible

DNS

Docker

ElasticSearch

HAProxy

Linux

MySQL

NGINX

PHP

RabbitMQ

TCP/IP

🕒 July 27

DiXiO

51 - 200

💳 Fintech

🏦 Banking

☁️ SaaS

DevOps Engineer at DiXiO maintaining cloud infrastructure reliability and security. Working remotely while optimizing various operational aspects of the IT environment.

🗣️🇫🇷 French Required

AWS

Azure

Cloud

Linux

Swift

Terraform

TypeScript