Senior AI Research Scientist, Model-based RL

Job not on LinkedIn

🔥 11 minutes ago

🇬🇧 United Kingdom – Remote

💵 £87.7k - £165.4k / year

⏰ Full Time

🟠 Senior

🧠 AI Research Scientist

🇬🇧 UK Skilled Worker Visa Sponsor

info
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Phaidra

Phaidra

51 - 200 employees

🏭 Manufacturing

📦 Logistics

🏗️ Construction

Manufacturing • Logistics • Construction

Phaidra is a company providing artificial intelligence controls to optimize mission-critical facilities such as data centers and industrial plants. Their closed-loop AI control service enhances plant stability, energy efficiency, and sustainability by reducing downtime, increasing productivity, and lowering CO2 emissions. Unlike traditional control systems, Phaidra's AI-driven controls continuously learn and improve over time without the need for new hardware. The system provides real-time optimization and integrates with existing control systems, enhancing safety and operational stability while providing full transparency of performance data. Phaidra uses cutting-edge deep reinforcement learning techniques to deliver exceptional results in some of the world's toughest challenges, including significant energy savings in Google's data centers.

📋 Description

• Design, implement, and evaluate model-based reinforcement learning agents — including planning-based controllers (MPC, MPPI) — and the software prototypes needed to deploy them on real industrial control systems. • Develop learned dynamics and world models (learned surrogates) that generalize across systems, including the training pipelines — pretraining, curriculum learning, active/adversarial learning, and fine-tuning — needed to make them reliable for planning and control. • Research and implement methods for e.g. safe RL, constrained control, scenario planning and Bayesian RL, to develop agents that satisfy safety constraints during deployment. • Report and present research findings and developments including status and results clearly and efficiently both internally and externally, verbally and in writing. • Participate in and organize ambitious collaborative research projects, and work with external collaborators and partners to translate research into production outcomes. • Mentor and guide Research Engineers to apply research findings and developments to industrial domains. • Independently defines new research directions • Translates research into practical outcomes • Owns the development and rollout for an entire research area or large project

🎯 Requirements

• PhD in a technical field or equivalent practical experience, with a strong background in model-based reinforcement learning and demonstrated knowledge in one or more of the following: • Planning algorithms • World models / learned dynamics surrogates • Reinforcement Learning and Deep Learning • Control Theory • Safe / constrained RL • 2+ years of research experience in academia or industry after PhD graduation. • Extensive research in the fields of {ModelBased, ModelFree, Safe}RL and Control Theory, with particular depth in model-based methods. • Hands-on experience building and evaluating agents against simulators (e.g. differentiable simulators or world models) and closing the sim-to-real gap. • Share our company values: Collaboration, Transparency, Operational Excellence, Ownership, and Empathy.

🏖️ Benefits

• Fast-paced, team-oriented environment where your work directly shapes the company’s direction. • We are a 100% remote company. • Competitive compensation & meaningful equity. • Outsized responsibilities & professional development. • Training is foundational; functional, customer immersion, and development training. • Medical, dental, and vision insurance (exact benefits vary by region). • Unlimited paid time off, with a required minimum of 20 days per year. • Paid parental leave (exact benefits vary by region). • Flexible stipends to support your workspace, well-being, and continued professional development. • Company MacBook.

Apply Now

Similar Jobs

🕒 June 17

Tether.to

11 - 50

₿ Crypto

💳 Fintech

💸 Finance

AI Research Engineer handling innovative vision-language models development at Tether. Join a global team working on cutting-edge fintech solutions.

🕒 June 11

Tether.to

11 - 50

₿ Crypto

💳 Fintech

💸 Finance

AI Research Engineer driving innovation in architecture development for LLMs and Multi-Modal systems. Collaborate with global teams to enhance intelligence and efficiency in AI.

🕒 May 22

Brahma

11 - 50

₿ Crypto

💳 Fintech

🔌 API

Machine Learning Researcher focusing on voice synthesis models for audio at Brahma AI. Researching and building deep learning systems to generate expressive, natural-sounding speech.

🕒 February 25

Brahma

11 - 50

₿ Crypto

💳 Fintech

🔌 API

Machine Learning Researcher collaborating with a world-class team at BRAHMA AI to advance generative video models. Focusing on controllable expressions and audio-driven lip synchronisation.