
51 - 200 employees
Founded 2020
đź’Ľ Consulting
📦 Logistics
🏠Manufacturing
Consulting • Logistics • Manufacturing
NexGen Cloud is a provider of large-scale, sovereign and sustainable AI cloud infrastructure and GPU-as-a-service. It offers the AI Supercloud and Hyperstack platforms with on-demand NVIDIA HGX/H100/GB200 GPUs, managed Kubernetes, custom hardware/software configurations, and consulting and R&D partnerships through NexGen Labs to support enterprise training of foundational models and GenAI workloads. The company focuses on scalable, hybrid and sovereign deployments in Europe and North America and provides POCs, onboarding, and ongoing support.
đź•’ July 16
🇬🇧 United Kingdom – Remote
⏰ Full Time
🟡 Mid-level
đźź Senior
🏛️ Architect
đź‘» Ghost score 23%
Improve your chances of getting an interview by checking your resume score before you apply.

51 - 200 employees
Founded 2020
đź’Ľ Consulting
📦 Logistics
🏠Manufacturing
Consulting • Logistics • Manufacturing
NexGen Cloud is a provider of large-scale, sovereign and sustainable AI cloud infrastructure and GPU-as-a-service. It offers the AI Supercloud and Hyperstack platforms with on-demand NVIDIA HGX/H100/GB200 GPUs, managed Kubernetes, custom hardware/software configurations, and consulting and R&D partnerships through NexGen Labs to support enterprise training of foundational models and GenAI workloads. The company focuses on scalable, hybrid and sovereign deployments in Europe and North America and provides POCs, onboarding, and ongoing support.
• Own end-to-end cluster architecture for large-scale NVIDIA GPU deployments, from customer requirements through rack layouts, BOM, power and cooling design, to production handover • Design high-performance network fabrics across compute (InfiniBand, RDMA, NVLink/NVSwitch), storage, and WAN, defining topology, oversubscription models, and scaling strategies • Engage directly with OEMs and vendors to validate hardware configurations, review quotes, and ensure technically sound and commercially optimized designs • Provide technical oversight during deployment and bring-up, supporting hardware validation and performance testing and serving as escalation point for complex integration issues • Act as a senior technical leader across Solutions Architecture, Cloud Engineering, and data centre partners • Contribute to standardized reference designs and build out the HPC engineering function • Report to the Head of Infrastructure
• Proven experience in HPC or AI software stack design and delivery at scale, including workload profiling, scheduler configuration (SLURM, PBS, or equivalent), MPI/NCCL tuning, and distributed training frameworks such as PyTorch, JAX, or DeepSpeed • Deep understanding of GPU software environments: CUDA, cuDNN, NCCL, driver stacks, and production AI training and inference tooling • Hands-on experience optimizing AI and HPC workloads across multi-GPU and multi-node configurations, including profiling, bottleneck identification, and performance tuning at application and infrastructure layers • Strong working knowledge of containerization and orchestration in HPC/AI contexts: Docker, Kubernetes, NVIDIA GPU Operator, and container-native workload management • Background in an OEM, hyperscaler, neo-cloud, or enterprise/research HPC environment, with exposure to the full design-to-deployment lifecycle for GPU-accelerated workloads • Ability to produce clear, professional technical documentation and architecture diagrams for engineering and board-level audiences • Ability to engage confidently with customers, vendors, and internal engineering teams as a technical authority and translate complex software and performance trade-offs into actionable decisions • Nice to have: experience with large-scale cluster performance benchmarking, including NCCL tests or MLPerf • Nice to have: exposure to MLOps tooling and AI platform layers, including MLflow, W&B, Triton, vLLM, Kubeflow, and Airflow • Nice to have: familiarity with InfiniBand and high-performance networking for distributed training performance
• Competitive salary and annual discretionary bonus scheme • Employee wellbeing benefits • 25 days of holiday, plus public holidays • Flexible working arrangements (remote or hybrid, depending on role and location) • Real ownership and autonomy, with the trust to take initiative and experiment • The opportunity to make a visible, meaningful impact as we scale • Clear career progression and growth opportunities in a fast-growing company • A collaborative, international culture built on trust, transparency, and ownership • The chance to help shape NexGen Cloud's team, culture, and future alongside ambitious, mission-driven colleagues
Apply Nowđź•’ July 15
Key role in transforming customer service with AI Agents at Parloa. Supporting implementation of solutions while enhancing customer experiences for clients and partners.
đź•’ July 6
AWS Business Architect delivering AWS-based solutions and supporting consulting activities across sales and delivery for Kyndryl's customers. Leading engagements to identify requirements and articulate the value of AWS capabilities while collaborating with internal teams.
đź•’ July 1
Principal strategy and architecture lead advising UK public sector organisations on technology, cyber security, and transformation. Shaping strategies, roadmaps, governance, and architecture for Phoenix, a managed IT services provider.
đź•’ June 9
Visionary leader defining technical strategy for global products at G-P. Overseeing architecture evolution incorporating AI/ML capabilities and ensuring scalability and security.
🇬🇧 United Kingdom – Remote
đź’µ ÂŁ159.2k - ÂŁ199k / year
⏰ Full Time
đźź Senior
🏛️ Architect
đź•’ May 28
Application Architect facilitating AI-driven solutions for medical imaging in clinical trials. Collaborating with product teams to refine and document technical specifications while ensuring compliance and scalability.