Engineering Manager – Site Reliability

Stelle nicht auf LinkedIn

🕒 vor 2 Monaten

🇬🇧 Vereinigtes Königreich – Remote

⏰ Vollzeit

🟠 Senior

🔴 Experte

⛑ DevOps- und Site Reliability Engineer (SRE)

👻 Geisterscore 26%

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of Pliant

Pliant

201 - 500 Mitarbeiter

Gegründet 2020

💳 Fintech

☁️ SaaS

🤝 B2B

💰 €40.000.000 Series B - Pliant im 2025-04

Fintech • SaaS • B2B

Pliant ist eine Kreditkartenplattform, die es Unternehmen, Banken und Fintechs ermöglicht, Kreditkarten auszugeben und zu verwalten, Zahlungen zu automatisieren sowie Ausgaben- und Buchhaltungsworkflows zu optimieren. Pliant bietet Payment Apps, eine Pro-API für Kartenausgabe und Automatisierung, Cards-as-a-Service (CaaS) und Banking-as-a-Service (BaaS), um integrierte oder White-Label-Kartenprogramme bereitzustellen, sowie globale Konten, Devisen und Überweisungen, Kredite & Finanzierung sowie Compliance- und Enablement-Dienste. Das Unternehmen bedient Firmenkunden, E-Commerce, Wiederverkäufer, SaaS-Unternehmen, Reise- und Marketingagenturen sowie Banken mit Funktionen wie Echtzeitüberwachung, Ausgabenkontrollen, Belegmanagement, Integrationen in Buchhaltungs- und Ausgabensysteme, einmalig verwendbaren virtuellen Karten und API-gesteuerter Automatisierung. Pliant besitzt eine E-Geld-Lizenz in der EU, gibt Karten im Vereinigten Königreich aus, ist PCI DSS und ISO/IEC 27001 zertifiziert und konzentriert sich auf die Optimierung von B2B-Zahlungen und Embedded-Finance-Lösungen.

Beschreibung

• Define the framework for teams to set SLOs and error budgets, educating and supporting product teams • Own blameless post-mortems and root-cause fixes • Implement production readiness reviews for new releases • Improve Datadog observability coverage, including alerts, dashboards, and on-call pages • Hire and build the Site Reliability team from the ground up • Build the on-call rotation and incident management process • Establish SLOs and an incident review process • Integrate reliability into the software development lifecycle and reduce repeat incidents

🎯 Anforderungen

• 7–10 years of engineering experience, including at least 3 years directly managing engineers • Track record of hiring and developing engineers, including levelling or promoting team members • Hands-on production or reliability engineering background • Prior experience carrying a pager; this is not a first management role • Strong AWS and Terraform experience • Comfortable working inside a managed infrastructure-as-code pipeline • Experience building or running an on-call rotation and incident management process • Strong platform observability experience • Clear communication for a technical, cross-team audience • Track record of introducing reliability practices into product engineering teams • Proficiency with AI-assisted development tools such as Claude Code and Cursor • Ability to rigorously review AI-written pull requests • Familiarity with PCI DSS, SOC 2, and ISO 27001 environments is relevant to the stack/context

🏖️ Vorteile

• Attractive remuneration • Choice of preferred OS: Windows or Mac • Flat hierarchy and transparent communication in a relaxed, professional atmosphere • Opportunity to develop your talent in a dynamic team with ambitious goals • Flexibility and possibility to work remotely • Pliant Card with monthly credit to explore the product and enjoy food with colleagues

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 2 Monaten

Cognativ Inc

51 - 200

🤝 B2B

💼 Beratung

🤖 Künstliche Intelligenz

Senior Site Reliability Engineer managing reliability for a distributed, camera-based video monitoring and AI alerting platform. Focusing on operational health, service objectives, and incident response.

🇬🇧 Vereinigtes Königreich – Remote

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 2 Monaten

GitLab

1001 - 5000

💼 Beratung

📣 Marketing

🤖 Künstliche Intelligenz

Senior Backend Engineer developing cloud-native and self-managed deployment environments for GitLab. Building consistency in deployment across development and production environments.

🇬🇧 Vereinigtes Königreich – Remote

💰 Secondary Market im 2020-11

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 2 Monaten

Runware

11 - 50

🤖 Künstliche Intelligenz

🔌 API

📱 Medien

Site Reliability Engineer ensuring reliability and performance of Runware's AI platforms. Collaborating across software, infrastructure, and operations to enhance observability and reduce incidents.

🇬🇧 Vereinigtes Königreich – Remote

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 3 Monaten

Omilia - Conversational Intelligence

201 - 500

💼 Beratung

🛡️ Versicherung

✈️ Reisen

Senior Site Reliability Engineer operating and maintaining production clusters while developing observability solutions. Collaborating with teams to enhance platform reliability through automation and monitoring.

🇬🇧 Vereinigtes Königreich – Remote

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 3 Monaten

DeepHealth

11 - 50

🏥 Gesundheitswesen

💼 Beratung

📦 Logistik

DevOps Engineer managing AWS infrastructure and enhancing platform reliability at DeepHealth. Collaborating with teams to automate processes and improve software delivery.

🇬🇧 Vereinigtes Königreich – Remote

💵 £60.000 - £70.000 / Jahr

💰 €225.000 Grant im 2019-08

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich