AI Infrastructure & Platform Operations Engineer

🕒 vor 18 Tagen

đŸ‡ȘđŸ‡ș Europa – Remote

đŸ’” $60.000 - $67.000 / Jahr

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

đŸ—ïž Plattformingenieur

đŸ—ŁïžđŸ‡ș🇾🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 ÜberprĂŒfen Sie Ihre Lebenslauf-Bewertung fĂŒr diese Stelle

Verbessern Sie Ihre Chancen auf ein VorstellungsgesprĂ€ch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung ĂŒberprĂŒfen.

Logo of Mirantis

Mirantis

501 - 1000 Mitarbeiter

đŸ’Œ Beratung

đŸ„ Gesundheitswesen

📩 Logistik

Consulting ‱ Healthcare ‱ Logistics

Mirantis ist ein Unternehmen, das sich auf Container-Management und Cloud-Infrastrukturlösungen spezialisiert hat. Das Portfolio umfasst unter anderem Mirantis Kubernetes Engine (MKE), Mirantis OpenStack for Kubernetes (MOSK) und Mirantis Container Cloud (MCC) – Plattformen fĂŒr Kubernetes und Container-Management auf Enterprise-Niveau. DarĂŒber hinaus entwickelt Mirantis Werkzeuge fĂŒr sichere Software-Lieferketten, etwa die Mirantis Container Runtime (MCR) und die Mirantis Secure Registry (MSR). Als Verfechter von Open-Source-Technologien unterstĂŒtzt Mirantis verschiedene Projekte und stellt Ressourcen wie Lens Desktop, eine beliebte Kubernetes-IDE, sowie technischen Support fĂŒr Unternehmen bereit, die Cloud-native Technologien einfĂŒhren. Die Lösungen von Mirantis richten sich an Bereiche wie den öffentlichen Sektor, Finanzdienstleistungen sowie SaaS- und Technologiedienstleistungen.

Beschreibung

‱ Monitor, operate, and support production AI infrastructure platforms. ‱ Investigate and resolve infrastructure, networking, hardware, and platform-related incidents. ‱ Support NVIDIA GPU infrastructure and associated platform services. ‱ Monitor and troubleshoot Kubernetes-based environments. ‱ Investigate performance, availability, and reliability issues across infrastructure and platform components. ‱ Collaborate with engineering teams, hardware vendors, datacenter personnel, and service delivery teams to resolve technical issues. ‱ Participate in incident response, root cause analysis, and operational improvement activities. ‱ Contribute to improvements in monitoring, observability, automation, and operational processes. ‱ Maintain operational documentation, runbooks, and knowledge articles.

🎯 Anforderungen

‱ 3+ years of experience in infrastructure operations, platform operations, network operations, site reliability engineering, cloud operations, datacenter operations, or related technical roles. ‱ Strong Linux administration and troubleshooting skills. ‱ Good understanding of networking concepts and experience diagnosing infrastructure-related issues. ‱ Working knowledge of Kubernetes in production environments. ‱ Experience supporting production infrastructure and services. ‱ Strong analytical and problem-solving skills. ‱ Experience working within structured operational and incident management processes. ‱ Excellent communication and collaboration skills. ‱ Ability to work within a shift-based operational environment. ‱ Experience in one or more of the following areas is highly desirable: NVIDIA GPU infrastructure and accelerated computing platforms. ‱ InfiniBand networking and NVIDIA UFM. ‱ Kubernetes platform operations. ‱ AI infrastructure or HPC environments. ‱ Site Reliability Engineering (SRE) or Platform Engineering. ‱ Observability platforms such as Grafana, Prometheus, ELK, or OpenTelemetry. ‱ Infrastructure automation technologies and Infrastructure-as-Code practices. ‱ Large-scale distributed systems and production platforms.

đŸ–ïž Vorteile

‱ Work with some of the most advanced AI infrastructure environments in production today. ‱ Gain exposure to NVIDIA GPU technologies, Kubernetes platforms, and high-performance networking environments. ‱ Help define how next-generation AI infrastructure is operated and supported. ‱ Be part of a team shaping the future of AI-powered operations through k0rdent AI. ‱ Join a growing organisation investing heavily in AI infrastructure and platform services.

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 1 Monat

Sourcegraph

51 - 200

đŸ€– KĂŒnstliche Intelligenz

☁ SaaS

🏱 Unternehmen

Software Engineer, verantwortlich fĂŒr die Entwicklung von Inter-Cloud-KonnektivitĂ€tslösungen und den Aufbau von API-Infrastruktur fĂŒr die Sourcegraph Cloud. Zusammenarbeit im Team bei komplexen technischen Problemen und Bereitstellung zuverlĂ€ssiger Services in einer Remote-Arbeitsumgebung.

đŸ‡ȘđŸ‡ș Europa – Remote

đŸ’” $148.000 / Jahr

💰 €150.000.000 Series D im 2021-07

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

đŸ—ïž Plattformingenieur

đŸ—ŁïžđŸ‡ș🇾🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

Spacelift

51 - 200

☁ SaaS

🏱 Unternehmen

Platform Engineer verantwortlich fĂŒr Aufbau und Betrieb der Infrastruktur-Orchestrierungsplattform von Spacelift. Umfasst CI/CD, Observability, Incident Response und Verbesserung der Developer Experience.

đŸ‡ȘđŸ‡ș Europa – Remote

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

đŸ—ïž Plattformingenieur

đŸ—ŁïžđŸ‡ș🇾🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

Zartis

201 - 500

đŸ’Œ Beratung

📩 Logistik

📣 Marketing

Senior Platform Engineer bei Zartis, Entwicklung moderner Plattformtechnologien. Evaluierung von API-Gateway-Lösungen und Konsolidierung von Pipelines fĂŒr Anwendungsteams.

đŸ‡ȘđŸ‡ș Europa – Remote

💰 Pre Seed Round im 2011-12

⏰ Vollzeit

🟠 Senior

đŸ—ïž Plattformingenieur

đŸ—ŁïžđŸ‡ș🇾🇬🇧 Englisch erforderlich

🕒 vor 3 Monaten

saas.group

51 - 200

đŸ’Œ Beratung

📣 Marketing

☁ SaaS

Senior Platform Engineer fĂŒr ScraperAPI, verantwortlich fĂŒr das Management und die Konsolidierung der Infrastruktur fĂŒr leistungsstarke Web-Scraping-Lösungen. Zusammenarbeit mit Engineering-Teams zur Umsetzung wesentlicher Verbesserungen der Plattform.

đŸ‡ȘđŸ‡ș Europa – Remote

⏰ Vollzeit

🟠 Senior

đŸ—ïž Plattformingenieur

đŸ—ŁïžđŸ‡ș🇾🇬🇧 Englisch erforderlich

🕒 vor 4 Monaten

Polar

1 - 10

💳 Fintech

☁ SaaS

🔌 API

Senior Plattformingenieur, der die Polar-Plattform fĂŒr hochdynamische Start-ups entwirft und weiterentwickelt. Konzeption von Systemen mit Schwerpunkt auf ZuverlĂ€ssigkeit und Skalierbarkeit in finanziellen Workflows ĂŒber verschiedene technische Ebenen hinweg.

đŸ‡ȘđŸ‡ș Europa – Remote

⏰ Vollzeit

🟠 Senior

đŸ—ïž Plattformingenieur

đŸ—ŁïžđŸ‡ș🇾🇬🇧 Englisch erforderlich