Infrastructure – DevOps Lead

Stelle nicht auf LinkedIn

🕒 vor 1 Monat

🇬🇧 Vereinigtes Königreich – Remote

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🇬🇧 UK-Skilled-Worker-Visum-Sponsor

infoinfo

👻 Geisterscore 24%

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of Brahma

Brahma

11 - 50 Mitarbeiter

Gegründet 2022

₿ Crypto

💳 Fintech

🔌 API

💰 €4.206.900 Seed Round - Brahma im 2022-02

Crypto • Fintech • API

Brahma ist eine entwicklerorientierte Orchestrierungsschicht, die die Onchain-Finanzlogik mit Offchain-Zahlungs- und Abwicklungssystemen verbindet. Sie ermöglicht programmierbare Kapitalflüsse über Blockchains, Strategien und traditionelle Zahlungswege und bietet Smart Accounts, automatisierte Agenten und programmierbare Kartenausgabe, damit Anwendungen Onchain-Positionen in reale Ausgaben umwandeln können. Brahma abstrahiert Infrastruktur und Abwicklung für Entwickler, sodass sie Onchain- oder hybride Anwendungen aufbauen können, ohne Backend-Operationen durchführen zu müssen.

Beschreibung

• Lead, mentor, and grow a team of 7 DevOps and Infrastructure engineers. • Drive agile delivery, sprint planning, and backlog prioritisation to align infrastructure deliverables with AI research and product roadmaps. • Establish best practices for Reliability Engineering, Infrastructure-as-Code (IaC), continuous integration, and incident post-mortems. • Manage high-density GPU clusters across a multi-cloud ecosystem optimised for large custom AI model training and real-time inference workflows. • Oversee infrastructure consumption, track cloud/hardware costs, negotiate vendor terms, and optimise GPU utilisation. • Serve as the senior technical escalation point for complex infrastructure incidents and architecture decisions. • Standardise platform deployments using Infrastructure as Code and modern container orchestration. • Partner with security stakeholders to ensure our AI training environments meet industry security standards.

🎯 Anforderungen

• Proven track record leading or managing a team of 5+ infrastructure, platform, or DevOps engineers. • Hands-on experience architecting and managing GPU-intensive workloads (NVIDIA clusters, cloud AI accelerators) for compute-heavy applications. • Expertise with Kubernetes, Docker, Terraform (or OpenTofu), and multi-cloud environments (with strong hands-on GCP experience). • Demonstrated experience designing, optimising, and maintaining high-performance storage architectures and caching layers for demanding compute workloads. • Strong experience with cloud cost governance (FinOps), capacity planning, and vendor interaction. • Exceptional stakeholder management skills with the ability to bridge business requirements and deep technical infrastructure details.

🏖️ Vorteile

• Health insurance • Professional development opportunities

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 1 Monat

Ensono

1001 - 5000

💼 Beratung

Site Reliability Engineer managing Cloud and Infrastructure as Code at Ensono. Leading client-facing discussions and driving service improvement initiatives.

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

Omilia - Conversational Intelligence

201 - 500

💼 Beratung

🛡️ Versicherung

✈️ Reisen

Senior Site Reliability Engineer operating and maintaining production clusters while developing observability solutions. Collaborating with teams to enhance platform reliability through automation and monitoring.

🇬🇧 Vereinigtes Königreich – Remote

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

DeepHealth

11 - 50

🏥 Gesundheitswesen

💼 Beratung

📦 Logistik

DevOps Engineer managing AWS infrastructure and enhancing platform reliability at DeepHealth. Collaborating with teams to automate processes and improve software delivery.

🇬🇧 Vereinigtes Königreich – Remote

💵 £60.000 - £70.000 / Jahr

💰 €225.000 Grant im 2019-08

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

DeepHealth

11 - 50

🏥 Gesundheitswesen

💼 Beratung

📦 Logistik

DevOps Engineer responsible for AWS and Kubernetes platform management at DeepHealth, a healthcare SaaS provider. Ensuring cloud infrastructure is secure and reliable for efficient software delivery.

🇬🇧 Vereinigtes Königreich – Remote

💵 £60.000 - £70.000 / Jahr

💰 €225.000 Grant im 2019-08

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

RTX

10.000+ Mitarbeiter

🏭 Fertigung

💼 Beratung

📦 Logistik

Principal Site Reliability Engineer managing complex AWS infrastructures for Collins Aerospace. Delivering B2B solutions and overseeing cloud migration projects in a collaborative team environment.

🇬🇧 Vereinigtes Königreich – Remote

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich