
51 - 200 Mitarbeiter
Gegründet 2020
💼 Beratung
📦 Logistik
🏭 Fertigung
Consulting • Logistics • Manufacturing
NexGen Cloud ist ein Anbieter von großangelegter, souveräner und nachhaltiger AI-Cloud-Infrastruktur und GPU-as-a-Service. Das Unternehmen bietet die Plattformen AI Supercloud und Hyperstack mit On-Demand NVIDIA HGX/H100/GB200 GPUs, verwaltetem Kubernetes, kundenspezifischen Hardware-/Software-Konfigurationen und Beratungs- sowie F&E-Partnerschaften über NexGen Labs an, um Unternehmenstraining von grundlegenden Modellen und GenAI-Workloads zu unterstützen. Das Unternehmen konzentriert sich auf skalierbare, hybride und souveräne Bereitstellungen in Europa und Nordamerika und bietet POCs, Onboarding und fortlaufenden Support an.
🕒 vor 1 Monat
🗣️🇺🇸🇬🇧 Englisch erforderlich
Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

51 - 200 Mitarbeiter
Gegründet 2020
💼 Beratung
📦 Logistik
🏭 Fertigung
Consulting • Logistics • Manufacturing
NexGen Cloud ist ein Anbieter von großangelegter, souveräner und nachhaltiger AI-Cloud-Infrastruktur und GPU-as-a-Service. Das Unternehmen bietet die Plattformen AI Supercloud und Hyperstack mit On-Demand NVIDIA HGX/H100/GB200 GPUs, verwaltetem Kubernetes, kundenspezifischen Hardware-/Software-Konfigurationen und Beratungs- sowie F&E-Partnerschaften über NexGen Labs an, um Unternehmenstraining von grundlegenden Modellen und GenAI-Workloads zu unterstützen. Das Unternehmen konzentriert sich auf skalierbare, hybride und souveräne Bereitstellungen in Europa und Nordamerika und bietet POCs, Onboarding und fortlaufenden Support an.
• Own end-to-end cluster architecture for large-scale NVIDIA GPU deployments, from customer requirements through rack layouts, BOM, power and cooling design, to production handover • Design high-performance network fabrics across compute (InfiniBand, RDMA, NVLink/NVSwitch), storage, and WAN, defining topology, oversubscription models, and scaling strategies • Engage directly with OEMs and vendors to validate hardware configurations, review quotes, and ensure technically sound and commercially optimized designs • Provide technical oversight during deployment and bring-up, supporting hardware validation and performance testing and serving as escalation point for complex integration issues • Act as a senior technical leader across Solutions Architecture, Cloud Engineering, and data centre partners • Contribute to standardized reference designs and build out the HPC engineering function • Report to the Head of Infrastructure
• Proven experience in HPC or AI software stack design and delivery at scale, including workload profiling, scheduler configuration (SLURM, PBS, or equivalent), MPI/NCCL tuning, and distributed training frameworks such as PyTorch, JAX, or DeepSpeed • Deep understanding of GPU software environments: CUDA, cuDNN, NCCL, driver stacks, and production AI training and inference tooling • Hands-on experience optimizing AI and HPC workloads across multi-GPU and multi-node configurations, including profiling, bottleneck identification, and performance tuning at application and infrastructure layers • Strong working knowledge of containerization and orchestration in HPC/AI contexts: Docker, Kubernetes, NVIDIA GPU Operator, and container-native workload management • Background in an OEM, hyperscaler, neo-cloud, or enterprise/research HPC environment, with exposure to the full design-to-deployment lifecycle for GPU-accelerated workloads • Ability to produce clear, professional technical documentation and architecture diagrams for engineering and board-level audiences • Ability to engage confidently with customers, vendors, and internal engineering teams as a technical authority and translate complex software and performance trade-offs into actionable decisions • Nice to have: experience with large-scale cluster performance benchmarking, including NCCL tests or MLPerf • Nice to have: exposure to MLOps tooling and AI platform layers, including MLflow, W&B, Triton, vLLM, Kubeflow, and Airflow • Nice to have: familiarity with InfiniBand and high-performance networking for distributed training performance
• Competitive salary and annual discretionary bonus scheme • Employee wellbeing benefits • 25 days of holiday, plus public holidays • Flexible working arrangements (remote or hybrid, depending on role and location) • Real ownership and autonomy, with the trust to take initiative and experiment • The opportunity to make a visible, meaningful impact as we scale • Clear career progression and growth opportunities in a fast-growing company • A collaborative, international culture built on trust, transparency, and ownership • The chance to help shape NexGen Cloud's team, culture, and future alongside ambitious, mission-driven colleagues
Jetzt Bewerben🕒 vor 1 Monat
Key role in transforming customer service with AI Agents at Parloa. Supporting implementation of solutions while enhancing customer experiences for clients and partners.
🗣️🇺🇸🇬🇧 Englisch erforderlich
🕒 vor 2 Monaten
AWS Business Architect delivering AWS-based solutions and supporting consulting activities across sales and delivery for Kyndryl's customers. Leading engagements to identify requirements and articulate the value of AWS capabilities while collaborating with internal teams.
🇬🇧 Vereinigtes Königreich – Remote
⏰ Vollzeit
🟡 Mittelstufe
🟠 Senior
🏛️ Architekt
🇬🇧 UK-Skilled-Worker-Visum-Sponsor
🗣️🇺🇸🇬🇧 Englisch erforderlich
🕒 vor 2 Monaten
Principal strategy and architecture lead advising UK public sector organisations on technology, cyber security, and transformation. Shaping strategies, roadmaps, governance, and architecture for Phoenix, a managed IT services provider.
🗣️🇺🇸🇬🇧 Englisch erforderlich
🕒 vor 3 Monaten
Visionary leader defining technical strategy for global products at G-P. Overseeing architecture evolution incorporating AI/ML capabilities and ensuring scalability and security.
🗣️🇺🇸🇬🇧 Englisch erforderlich
🕒 vor 3 Monaten
Application Architect facilitating AI-driven solutions for medical imaging in clinical trials. Collaborating with product teams to refine and document technical specifications while ensuring compliance and scalability.
🗣️🇺🇸🇬🇧 Englisch erforderlich