76 to 100 of 107 Reinforcement Learning Jobs in the UK

Head of Applied AI Research

Hiring Organisation
Higher - AI recruitment
Location
London Area, United Kingdom
publications role. It is a role for someone who wants to do serious science and see it matter. WHAT YOU BRING PhD in Machine Learning from a leading research institution 7+ years leading applied AI research teams in a commercial environment with access to large-scale, unique data Deep … expertise in at least two of: LLMs and foundation models (pretraining, fine-tuning, distillation), physics-informed neural networks, or reinforcement learning Experience with agent frameworks, tool-use planning and workflow orchestration Systems-level fluency across data pipelines, training infrastructure and inference (PyTorch, JAX, TensorFlow) Strong MLOps discipline ...

ML RL & Optimization Scientist for Scalable LLMs

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Huawei Technologies Research & Development (UK) Ltd is looking for a skilled professional to research and develop advanced machine learning systems focused on scaling reinforcement learning and optimization infrastructure. You will design and execute RL workflows to enhance machine learning capabilities while managing large-scale distributed training ...

Applied Scientist (Integrity)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
What We Are Looking For Deep expertise in one applied ML, statistics or data science speciality. Examples could be risk modelling, LLM detection, or reinforcement learning – any area of genuine depth counts. Judgement about when to reach for simple statistics, classical ML, LLMs, or agentic approaches. 3+ years ...

Senior AI Director: NLP/LLM, Graphs & Regulated AI

Hiring Organisation
Jobleads-UK
Location
England, United Kingdom
Chief Data & Analytics Office (CDAO) invites applications for an Applied AI ML Director – NLP/LLM and Graphs. This role will apply sophisticated machine learning methods to NLP, graph analytics, speech analytics, time series, reinforcement learning and recommendation systems. You will collaborate with diverse teams to deploy ...

Lead Software Engineer - Java, AI, AWS

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
ensure our systems remain reliable and future‐ready. Job Responsibilities Design, develop, and implement AI and data engineering solutions, including data pipelines, machine learning models, analytics applications, and Agentic AI systems. Write secure, high‐quality code in Java, applying best practices for AI and data engineering. Develop and maintain … Agentic AI applications. Implement and monitor autonomous AI agents, ensuring reliability, safety, and alignment with organizational goals. Stay current with advancements in Agentic AI, reinforcement learning, and related frameworks. Required Qualifications, Capabilities, and Skills Formal training or certification in software engineering concepts and recent applied experience in software ...

Senior AI Software Engineer

Hiring Organisation
Jobleads-UK
Location
Cambridge, England, United Kingdom
novel model architectures and iterate rapidly toward measurable outcomes Solid experience with automated training pipelines, training orchestration, and scalable model evaluation Practical experience applying reinforcement learning in real product or research settings Excellent Python skills and strong software engineering fundamentals for building reliable, maintainable AI systems Strong communication ...

Senior Software Engineer, Data Platform London, UK Apply →

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
enterprises figure out how to add it to their products. To make them safe, aligned and actually useful, these models need human evaluation and reinforcement learning through human feedback (RLHF) during pre-training, fine-tuning, and production evaluations. This is the main innovation that’s enabled ChatGPT ...

Senior AI Engineer — Design & Manufacturing (Generative AI)

Hiring Organisation
Jobleads-UK
Location
Cambridge, England, United Kingdom
agents for enterprise workflows, and create robust training pipelines that enable rapid experimentation and deployment at scale. We value hands-on expertise with LLMs, reinforcement learning, and agentic systems, along with strong Python and software engineering fundamentals. #J-18808-Ljbffr ...

Research Engineer

Hiring Organisation
Jobleads-UK
Location
Tipton, England, United Kingdom
product‐minded research engineer who enjoys building real systems people depend on. You’ll likely have: Strong technical background in software engineering, machine learning, or applied AI, demonstrated through an advanced degree and/or equivalent experience building production AI systems Strong software engineering fundamentals and good judgment … Practical understanding of modern model adaptation and post‐training methods, including LoRA/QLoRA, SFT, distillation, preference optimization, reward modeling, DPO/GRPO, and reinforcement learning from verifiable feedback Ownership mindset: you drive projects end‐to‐end, move quickly from real usage, and care about shipping measurable improvements ...

AI Research Engineer

Hiring Organisation
Zealous Agency
Location
England, United Kingdom
well-funded AI research lab about to blast out of stealth mode, doing some novel work at the frontier of reinforcement learning. They’re moving into their next phase of growth after closing funding and are already working alongside major frontier AI labs. As a Founding RL Researcher … real influence over technical direction from day one and work directly with experienced second-time founders. Core skills and experience required Strong background in reinforcement learning Hands-on post-training experience: RLHF, GRPO etc Experience designing or building RL environments grounded in real-world workflows Distributed training ...

Chief Product Officer

Hiring Organisation
Jobleads-UK
Location
United Kingdom
real industrial environments. Shape our product architecture and build the engineering team. We're looking for someone who has Strong expertise in Machine Learning , especially time-series data, physics-informed AI, reinforcement learning, or agentic systems. Experience building and deploying ML products with real-world sensor data. ...

Senior AI Technologist

Hiring Organisation
17918
Location
Chelmsford, Essex, United Kingdom
join a rapidly expanding Data and Decision Support Capability team. This role focuses on cutting-edge research and development across various AI domains, including reinforcement learning, NLP/LLMs, knowle... WKCL1_UKTJ ...

Senior Model-Based RL AI Research Scientist (Remote)

Hiring Organisation
Jobleads-UK
Location
United Kingdom
Jobgether is seeking a Senior AI Research Scientist (Model-based RL) based in the United Kingdom. You will design and implement model-based reinforcement learning agents and production-ready AI systems for industrial automation, collaborating with a global team. You will explore safe RL, learned dynamics and world ...

Senior AI Technologist

Hiring Organisation
Anson Mccade
Location
Chelmsford, Essex, United Kingdom
Employment Type
Permanent
Salary
GBP 85,000 Annual
join a rapidly expanding Data and Decision Support Capability team. This role focuses on cutting-edge research and development across various AI domains, including reinforcement learning, NLP/LLMs, knowle click apply for full job details ...

Onsite RL Engineer – Robotics & Embodied AI

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Randstad Technologies Recruitment is searching for a Reinforcement Learning (RL) Engineer to join a leading robotics company in London. This permanent position involves working collaboratively to define the company's 2026 technical roadmap while solving complex challenges in autonomous systems. The role requires expertise in AI, MLOps, Software ...

Senior RL Data Engineer: End-to-End Pipelines & QA

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Senior Engineer in Greater London to develop and manage data pipelines for AI systems. This role involves significant responsibilities, including ensuring the quality of reinforcement learning data, collaborating with various teams, and innovating operational frameworks. The ideal candidate will possess strong software engineering skills, have a background ...

AI Deployment Engineer - Startups

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
plus. Have experience as a technical founder, or engineer at an early stage startup. Have familiarity with, or interest in, model training pipelines and reinforcement learning. Have experience building AI applications, agents, or evaluation systems, and can reason clearly about model behavior in complex workflows. Are comfortable working directly ...

Deep Learning Research Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
improve multimodal LLMs to achieve new functionality for our customers and optimize their deployments (cloud and edge). Some of our deep learning models are truly tiny - the memory footprint of our smallest computer vision model is just 1MB. You will train and design more accurate models, while also … LLMs. Trained neural networks that moved into production. Nice To Have Industry experience with efficient inference deployments (cloud or edge). Experience with Deep Reinforcement Learning. We only consider applicants who are currently based in, or willing to relocate to, London or Amsterdam. We have flexible working hours ...

Senior Research Scientist - Multimodal Vision Translation

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
DeepL is seeking a Senior Research Scientist to lead fine-tuning, post-training, and reinforcement learning for its document translation multimodal and vision models. You will develop models that reason about document layout, fuse data sources, and drive breakthroughs from prototype to deployment. The role requires hands ...

Senior Research Scientist: Steerable LLMs for Translation

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
DeepL is seeking a Senior Research Scientist to lead fine-tuning, post-training, model steerability, and reinforcement learning for our LLM-based translation models. You will own a major research direction, prototype rapidly, run large-scale experiments, and drive breakthroughs into production. You will fuse high-value human ...

Senior Model-Steering Scientist for LLM Translation

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
DeepL is seeking a Senior Research Scientist to lead fine-tuning, post-training, model steerability, and reinforcement learning for the next generation of translation models. You will fuse high-value human data with synthetic data and drive breakthroughs from research to production at scale. You will work with ...

RL Scaling Research Engineer: Large-Scale Experiments

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
United States Digital Space LLC is seeking a Research Engineer to design and run large-scale experiments in Reinforcement Learning. The role involves debugging complex issues, maintaining benchmarks, and working closely with research and engineering teams. Successful candidates will have strong empirical research skills and proficiency in Python. ...

RL Research Scientist for Humanoid Robots

Hiring Organisation
Jobleads-UK
Location
England, United Kingdom
Boston Dynamics is seeking a Research Scientist to advance reinforcement learning for Atlas humanoid platforms. You will design, implement, and train RL policies, building production-grade Python and C++ code and validating them in high-fidelity simulators before hardware testing. You will collaborate with controls and platform teams ...

Senior Research Scientist, RL & Post-Training (Hybrid)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
DeepL is seeking a Senior Research Scientist to design, implement, and deploy cutting-edge reinforcement learning and post-training research at scale. You will drive innovations that translate to production, collaborating with engineering, ML platforms, and HPC teams. The role emphasizes building scalable RL pipelines, aligning models with ...

Robotics Research Scientist: Real-World Impact

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
will work with foundation models and real robots, translating lab research into real-world systems, and contributing to diverse areas such as simulation, reinforcement learning, and vision-language-action models. #J-18808-Ljbffr ...