Engineering Manager – Site Reliability

Stelle nicht auf LinkedIn

🕒 vor 17 Tagen

🇬🇧 Vereinigtes Königreich – Remote

⏰ Vollzeit

🟠 Senior

🔴 Experte

⛑ DevOps- und Site Reliability Engineer (SRE)

👻 Geisterscore 12%

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of Pliant

Pliant

201 - 500 Mitarbeiter

Gegründet 2020

💳 Fintech

☁️ SaaS

🤝 B2B

💰 €40.000.000 Series B - Pliant im 2025-04

Fintech • SaaS • B2B

Pliant ist eine Kreditkartenplattform, die es Unternehmen, Banken und Fintechs ermöglicht, Kreditkarten auszugeben und zu verwalten, Zahlungen zu automatisieren sowie Ausgaben- und Buchhaltungsworkflows zu optimieren. Pliant bietet Payment Apps, eine Pro-API für Kartenausgabe und Automatisierung, Cards-as-a-Service (CaaS) und Banking-as-a-Service (BaaS), um integrierte oder White-Label-Kartenprogramme bereitzustellen, sowie globale Konten, Devisen und Überweisungen, Kredite & Finanzierung sowie Compliance- und Enablement-Dienste. Das Unternehmen bedient Firmenkunden, E-Commerce, Wiederverkäufer, SaaS-Unternehmen, Reise- und Marketingagenturen sowie Banken mit Funktionen wie Echtzeitüberwachung, Ausgabenkontrollen, Belegmanagement, Integrationen in Buchhaltungs- und Ausgabensysteme, einmalig verwendbaren virtuellen Karten und API-gesteuerter Automatisierung. Pliant besitzt eine E-Geld-Lizenz in der EU, gibt Karten im Vereinigten Königreich aus, ist PCI DSS und ISO/IEC 27001 zertifiziert und konzentriert sich auf die Optimierung von B2B-Zahlungen und Embedded-Finance-Lösungen.

Beschreibung

• Define the framework for teams to set SLOs and error budgets, educating and supporting product teams • Own blameless post-mortems and root-cause fixes • Implement production readiness reviews for new releases • Improve Datadog observability coverage, including alerts, dashboards, and on-call pages • Hire and build the Site Reliability team from the ground up • Build the on-call rotation and incident management process • Establish SLOs and an incident review process • Integrate reliability into the software development lifecycle and reduce repeat incidents

🎯 Anforderungen

• 7–10 years of engineering experience, including at least 3 years directly managing engineers • Track record of hiring and developing engineers, including levelling or promoting team members • Hands-on production or reliability engineering background • Prior experience carrying a pager; this is not a first management role • Strong AWS and Terraform experience • Comfortable working inside a managed infrastructure-as-code pipeline • Experience building or running an on-call rotation and incident management process • Strong platform observability experience • Clear communication for a technical, cross-team audience • Track record of introducing reliability practices into product engineering teams • Proficiency with AI-assisted development tools such as Claude Code and Cursor • Ability to rigorously review AI-written pull requests • Familiarity with PCI DSS, SOC 2, and ISO 27001 environments is relevant to the stack/context

🏖️ Vorteile

• Attractive remuneration • Choice of preferred OS: Windows or Mac • Flat hierarchy and transparent communication in a relaxed, professional atmosphere • Opportunity to develop your talent in a dynamic team with ambitious goals • Flexibility and possibility to work remotely • Pliant Card with monthly credit to explore the product and enjoy food with colleagues

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 24 Tagen

PatSnap

501 - 1000

💼 Beratung

🏥 Gesundheitswesen

📦 Logistik

Site Reliability Engineering Leader at PatSnap, leading the SRE team ensuring reliability for a global SaaS platform. Overseeing strategy, automation, and team development in cloud technologies.

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 24 Tagen

TwinStream

51 - 200

🎖️ Verteidigung

💼 Beratung

📦 Logistik

DevOps Engineer maintaining and deploying cross-domain systems using Docker and AMQP architecture for TwinStream clients. Collaborating with teams and ensuring system performance and availability.

🇬🇧 Vereinigtes Königreich – Remote

💵 £70.000 - £85.000 / Jahr

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 24 Tagen

RTX

10.000+ Mitarbeiter

🚀 Luft- und Raumfahrt

🎖️ Verteidigung

🏭 Fertigung

Principal Site Reliability Engineer managing AWS infrastructures for Collins Aerospace. Delivering B2B products and ensuring service availability with scalable solutions in aviation technology.

🇬🇧 Vereinigtes Königreich – Remote

💰 €200.000 Grant - RTX im 2024-11

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 25 Tagen

Cognativ

11 - 50

💼 Beratung

🥽 AR/VR

🤖 Künstliche Intelligenz

Senior Site Reliability Engineer managing reliability for a distributed, camera-based video monitoring and AI alerting platform. Focusing on operational health, service objectives, and incident response.

🇬🇧 Vereinigtes Königreich – Remote

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 25 Tagen

Peratera

11 - 50

💳 Fintech

🤝 B2B

🔌 API

DevOps/SRE Engineer responsible for platform reliability and automation at UK fintech. Building and evolving cloud infrastructure with automation and observability practices.

🇬🇧 Vereinigtes Königreich – Remote

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich