
11 - 50 employés
💼 Conseil
📦 Logistique
⚡ Énergie
💰 Venture Round en 2022-09
Consulting • Logistics • Energy
Fuse Energy est une entreprise énergétique qui simplifie le processus de changement de fournisseur d'électricité et fournit des informations de facturation en temps réel. Elle propose des tarifs compétitifs au Royaume-Uni et exploite des projets solaires et éoliens, réinvestissant les bénéfices dans des initiatives d'énergie renouvelable à l'échelle mondiale. Fuse Energy vise à offrir à ses clients une facturation transparente et un changement facile grâce à leur application, tout en promouvant la durabilité.
🕒 il y a 1 mois
🌐 Royaume-Uni, États-Unis, +1 autres pays – Distant
⏰ Temps Plein
🟡 Intermédiaire
🟠 Senior
🤖 Intelligence Artificielle
🇬🇧 Parrain de Visa de Travailleur Qualifié UK
👻 Score fantôme 44%
🗣️🇺🇸🇬🇧 Anglais requis
Améliorez vos chances d'obtenir un entretien en vérifiant votre score de CV avant de postuler.

11 - 50 employés
💼 Conseil
📦 Logistique
⚡ Énergie
💰 Venture Round en 2022-09
Consulting • Logistics • Energy
Fuse Energy est une entreprise énergétique qui simplifie le processus de changement de fournisseur d'électricité et fournit des informations de facturation en temps réel. Elle propose des tarifs compétitifs au Royaume-Uni et exploite des projets solaires et éoliens, réinvestissant les bénéfices dans des initiatives d'énergie renouvelable à l'échelle mondiale. Fuse Energy vise à offrir à ses clients une facturation transparente et un changement facile grâce à leur application, tout en promouvant la durabilité.
• Define Fuse's inference serving strategy and architecture from first principles. • Design and build the serving stack: request routing, batching, scheduling, and autoscaling for high-throughput, latency-sensitive inference workloads. • Own model-level optimisation strategy for serving - deciding where and how to apply quantisation, distillation, speculative decoding, and similar techniques to improve throughput and cost per token, partnering with the CUDA/GPU engineers. • Make the core software architecture calls on serving frameworks and orchestration (e.g. vLLM, TensorRT-LLM, SGLang, Triton Inference Server, or equivalents). • Translate throughput, latency, and uptime commitments into concrete technical specifications and serving capacity plans. • Act as a direct technical owner of inference performance and reliability. • Work closely with the CUDA and GPU engineering teams to ensure custom kernels and hardware performance work are integrated cleanly into the serving layer. • Set the standards, tooling, and benchmarks this function will run on as it grows.
• 4+ years of experience building or operating large-scale inference serving systems, or equivalent strong project/industry experience. • Deep, hands-on experience with inference serving frameworks and the techniques used to optimise them (batching, KV-cache management, quantisation, speculative decoding). • Strong systems thinking - able to reason about the full path from incoming request to served response across a large cluster. • Comfortable working directly with GPU/CUDA engineers to integrate low-level performance work into a serving system. • A track record of making high-stakes architecture calls and owning the outcome. • Comfort operating without a playbook - this is a founding role shaping a new function around architecture that's still early-stage, not joining an established one. • **Nice to Have** • Experience with Triton or custom ML inference/training frameworks. • Experience with autoscaling or capacity planning for large-scale inference workloads. • Exposure to multi-tenant serving or SLA-driven infrastructure. • Background at a hyperscaler, frontier AI lab, or large-scale distributed inference system. • Familiarity with Kubernetes/Slurm for cluster orchestration. • Interest or experience in energy markets, grid systems, or sustainability-focused compute.
• Competitive salary and an equity sign-on bonus. • Biannual bonus scheme. • Fully expensed tech to match your needs. • Breakfast and dinner allowance for office based employees.
Postuler Maintenant🕒 il y a 1 mois
AI Strategist leading strategic enterprise accounts for Ema's AI platform, driving transformative AI adoption across major enterprises.
🗣️🇺🇸🇬🇧 Anglais requis
🕒 il y a 2 mois
AI Educator teaching thousands how to put AI to work at The Rundown. Engaging content creation and collaboration with a passionate team in a fast-growing media company.
🇬🇧 Royaume-Uni – Télétravail
💰 €85 000 000 Private Equity Round - Electrify Video Partners en 2023-09
⏰ Temps Plein
🟡 Intermédiaire
🟠 Senior
🤖 Intelligence Artificielle
🗣️🇺🇸🇬🇧 Anglais requis
🕒 il y a 2 mois
Senior Director driving innovative insights in AI and infrastructure strategy at Gartner. Engaging clients and leading teams to shape effective technology adoptions.
🇬🇧 Royaume-Uni – Télétravail
⏰ Temps Plein
🟠 Senior
🤖 Intelligence Artificielle
🇬🇧 Parrain de Visa de Travailleur Qualifié UK
🗣️🇺🇸🇬🇧 Anglais requis
🕒 il y a 2 mois
Senior AI Consultant at Telefónica Tech designing and delivering AI solutions using Azure technology stack. Focus on AI, data science, and advanced analytics solutions.
🗣️🇺🇸🇬🇧 Anglais requis
🕒 il y a 2 mois
AI Business Transformation & Design Strategist at Coupa enhancing procurement through AI integration and transformation strategies. Lead workshops and design metrics for successful AI deployment.
🇬🇧 Royaume-Uni – Télétravail
⏰ Temps Plein
🟠 Senior
🔴 Expert
🤖 Intelligence Artificielle
🇬🇧 Parrain de Visa de Travailleur Qualifié UK
🗣️🇺🇸🇬🇧 Anglais requis