Technical Product Manager – Cluster Experience

Stelle nicht auf LinkedIn

Vermutlich ein Geisterjob

🕒 vor 8 Monaten

🌐 Niederlande, Deutschland, +1 weitere Länder – Remote

infoinfo

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⚒️ Technischer Produktmanager (TPM)

👻 Geisterscore 65%

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of Nebius Group

Nebius Group

1001 - 5000 Mitarbeiter

🤖 Künstliche Intelligenz

🏢 Unternehmen

☁️ SaaS

Artificial Intelligence • Enterprise • SaaS

Die Nebius Group baut eines der weltweit führenden Unternehmen für KI-Infrastruktur auf und konzentriert sich darauf, die notwendige Rechenleistung, Speicherkapazität und Tools für Entwickler im KI-Bereich bereitzustellen. Mit Sitz in Europa und an der Nasdaq notiert verfügt Nebius über eine globale Präsenz mit F&E-Zentren in Europa, Nordamerika und Israel. Das zentrale Angebot des Unternehmens ist eine KI-zentrierte Cloud-Plattform, die für rechenintensive KI-Workloads ausgelegt ist, ergänzt durch verschiedene weitere Geschäftsbereiche in den Bereichen Generative KI, Edtech und autonome Technologien.

Beschreibung

• Own key tracks in Cluster Experience: reliability, performance, and user experience for distributed ML workloads. • Define product direction from problem discovery → design → delivery → adoption, working closely with engineering and research teams. • Drive cross-functional execution across compute, networking, storage, observability, and platform teams. • Perform deep customer research: interviews, analytics, and workload studies to identify bottlenecks across hardware, network, scheduler, and runtime. • Translate state-of-the-art ML papers ideas into practical, scalable product features for large GPU clusters. • Shape how users interact with clusters - from dashboards and notifications to partitioning, node management, and training observability.

🎯 Anforderungen

• 3–5+ years of experience in product management, ML infrastructure/MLOps, distributed systems engineering, or cloud architecture. • Strong technical foundation in computer science, distributed systems, or ML infrastructure. • Hands-on familiarity with ML training, ideally using orchestrators like Slurm, Kubernetes, Ray, or similar systems. • Proven ability to ship technically complex features with multiple engineering teams. • Excellent communicator capable of influencing engineering, research, and customer stakeholders. • Experience with product analytics, data-driven prioritization, and experiment design. • Strong willingness and ability to learn quickly in a fast-evolving ML and infrastructure environment.

🏖️ Vorteile

• Competitive salary and comprehensive benefits package. • Opportunities for professional growth within Nebius. • Flexible working arrangements. • A dynamic and collaborative work environment that values initiative and innovation.

Jetzt Bewerben