76 to 99 of 99 Reinforcement Learning Jobs in England

Senior Software Engineer, Data Platform London, UK Apply →

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
enterprises figure out how to add it to their products. To make them safe, aligned and actually useful, these models need human evaluation and reinforcement learning through human feedback (RLHF) during pre-training, fine-tuning, and production evaluations. This is the main innovation that’s enabled ChatGPT ...

Senior AI Engineer — Design & Manufacturing (Generative AI)

Hiring Organisation
Jobleads-UK
Location
Cambridge, England, United Kingdom
agents for enterprise workflows, and create robust training pipelines that enable rapid experimentation and deployment at scale. We value hands-on expertise with LLMs, reinforcement learning, and agentic systems, along with strong Python and software engineering fundamentals. #J-18808-Ljbffr ...

Research Engineer

Hiring Organisation
Jobleads-UK
Location
Tipton, England, United Kingdom
product‐minded research engineer who enjoys building real systems people depend on. You’ll likely have: Strong technical background in software engineering, machine learning, or applied AI, demonstrated through an advanced degree and/or equivalent experience building production AI systems Strong software engineering fundamentals and good judgment … Practical understanding of modern model adaptation and post‐training methods, including LoRA/QLoRA, SFT, distillation, preference optimization, reward modeling, DPO/GRPO, and reinforcement learning from verifiable feedback Ownership mindset: you drive projects end‐to‐end, move quickly from real usage, and care about shipping measurable improvements ...

AI Research Engineer

Hiring Organisation
Zealous Agency
Location
England, United Kingdom
well-funded AI research lab about to blast out of stealth mode, doing some novel work at the frontier of reinforcement learning. They’re moving into their next phase of growth after closing funding and are already working alongside major frontier AI labs. As a Founding RL Researcher … real influence over technical direction from day one and work directly with experienced second-time founders. Core skills and experience required Strong background in reinforcement learning Hands-on post-training experience: RLHF, GRPO etc Experience designing or building RL environments grounded in real-world workflows Distributed training ...

Senior AI Technologist

Hiring Organisation
17918
Location
Chelmsford, Essex, United Kingdom
join a rapidly expanding Data and Decision Support Capability team. This role focuses on cutting-edge research and development across various AI domains, including reinforcement learning, NLP/LLMs, knowle... WKCL1_UKTJ ...

Senior AI Technologist

Hiring Organisation
Anson Mccade
Location
Chelmsford, Essex, United Kingdom
Employment Type
Permanent
Salary
GBP 85,000 Annual
join a rapidly expanding Data and Decision Support Capability team. This role focuses on cutting-edge research and development across various AI domains, including reinforcement learning, NLP/LLMs, knowle click apply for full job details ...

Onsite RL Engineer – Robotics & Embodied AI

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Randstad Technologies Recruitment is searching for a Reinforcement Learning (RL) Engineer to join a leading robotics company in London. This permanent position involves working collaboratively to define the company's 2026 technical roadmap while solving complex challenges in autonomous systems. The role requires expertise in AI, MLOps, Software ...

Senior RL Data Engineer: End-to-End Pipelines & QA

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Senior Engineer in Greater London to develop and manage data pipelines for AI systems. This role involves significant responsibilities, including ensuring the quality of reinforcement learning data, collaborating with various teams, and innovating operational frameworks. The ideal candidate will possess strong software engineering skills, have a background ...

AI Deployment Engineer - Startups

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
plus. Have experience as a technical founder, or engineer at an early stage startup. Have familiarity with, or interest in, model training pipelines and reinforcement learning. Have experience building AI applications, agents, or evaluation systems, and can reason clearly about model behavior in complex workflows. Are comfortable working directly ...

Deep Learning Research Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
improve multimodal LLMs to achieve new functionality for our customers and optimize their deployments (cloud and edge). Some of our deep learning models are truly tiny - the memory footprint of our smallest computer vision model is just 1MB. You will train and design more accurate models, while also … LLMs. Trained neural networks that moved into production. Nice To Have Industry experience with efficient inference deployments (cloud or edge). Experience with Deep Reinforcement Learning. We only consider applicants who are currently based in, or willing to relocate to, London or Amsterdam. We have flexible working hours ...

Senior Research Scientist - Multimodal Vision Translation

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
DeepL is seeking a Senior Research Scientist to lead fine-tuning, post-training, and reinforcement learning for its document translation multimodal and vision models. You will develop models that reason about document layout, fuse data sources, and drive breakthroughs from prototype to deployment. The role requires hands ...

Senior Research Scientist: Steerable LLMs for Translation

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
DeepL is seeking a Senior Research Scientist to lead fine-tuning, post-training, model steerability, and reinforcement learning for our LLM-based translation models. You will own a major research direction, prototype rapidly, run large-scale experiments, and drive breakthroughs into production. You will fuse high-value human ...

Senior Model-Steering Scientist for LLM Translation

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
DeepL is seeking a Senior Research Scientist to lead fine-tuning, post-training, model steerability, and reinforcement learning for the next generation of translation models. You will fuse high-value human data with synthetic data and drive breakthroughs from research to production at scale. You will work with ...

RL Scaling Research Engineer: Large-Scale Experiments

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
United States Digital Space LLC is seeking a Research Engineer to design and run large-scale experiments in Reinforcement Learning. The role involves debugging complex issues, maintaining benchmarks, and working closely with research and engineering teams. Successful candidates will have strong empirical research skills and proficiency in Python. ...

RL Research Scientist for Humanoid Robots

Hiring Organisation
Jobleads-UK
Location
England, United Kingdom
Boston Dynamics is seeking a Research Scientist to advance reinforcement learning for Atlas humanoid platforms. You will design, implement, and train RL policies, building production-grade Python and C++ code and validating them in high-fidelity simulators before hardware testing. You will collaborate with controls and platform teams ...

Senior Research Scientist, RL & Post-Training (Hybrid)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
DeepL is seeking a Senior Research Scientist to design, implement, and deploy cutting-edge reinforcement learning and post-training research at scale. You will drive innovations that translate to production, collaborating with engineering, ML platforms, and HPC teams. The role emphasizes building scalable RL pipelines, aligning models with ...

Robotics Research Scientist: Real-World Impact

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
will work with foundation models and real robots, translating lab research into real-world systems, and contributing to diverse areas such as simulation, reinforcement learning, and vision-language-action models. #J-18808-Ljbffr ...

Robotics Research Scientist: AI Agents & Real-World Robots

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
will design, implement and evaluate large models for robotic agents, while collaborating across teams and contributing to open research. This role emphasizes scalable ML, reinforcement learning, and real-world robot integration. You will work with state-of-the-art robotics platforms, prototype applications, and cutting-edge datasets ...

Senior Research Scientist — Foundation Model Adaptation

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
DeepL is seeking a Senior Research Scientist to lead cutting-edge reinforcement learning research and scale post-training for multi-modal models. You will drive innovations from conception to production, ensuring safety, efficiency, and robust performance across our global translations platform. You will work with a diverse, international ...

Senior Research Scientist: Scale RL & Post-Training for Models

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
DeepL in the United Kingdom is hiring a Senior Research Scientist to design, implement, and deploy cutting-edge reinforcement learning research at scale, driving innovations that align pre-trained models with real user goals. You'll collaborate with engineering, ML platform, and HPC teams to translate findings into ...

Posted 1 day ago London, United Kingdom On-site Amazon Principal Technical Product Manager, Pri[...]

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
will build coordinated intelligence so offers respect customer signals and journey stage. Additionally, you will partner with Applied Science and ML Engineering on reinforcement learning and multi-objective optimization to prove that respecting customers drives better outcomes. #J-18808-Ljbffr ...

Senior AI/ML Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
background in Eval-driven-development. Expert-level proficiency in Python and Typescript/Javascript. Proven experience building an LLM agent harness . Experience utilizing reinforcement learning. Strong architectural skills with experience designing scalable, production-grade systems for machine learning applications. A solid understanding of statistical principles specifically ...

Research Scientist, Strategic Bets, DeepMind

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
designing new architectures, our research scientists work on real-world problems that span the breadth of computer science, such as machine (and deep) learning, data mining, natural language processing, hardware and software performance analysis, improving compilers for mobile platforms, as well as core search and much more. … scientific discovery, ensuring safety and ethics are always our highest priority. We are pushing the boundaries across multiple domains. Our global teams offer diverse learning opportunities and varied career pathways for those driven to achieve exceptional results through collective effort. Minimum qualifications: PhD degree in Machine learning ...

Founding Applied Scientist

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Customers: Work directly with the engineering and product teams of our most strategic customers. You'll be their trusted advisor for all things machine learning, helping them adopt agentic architectures. End-to-End Model Development: Design, build, and deploy production-grade machine learning models for our customers using … Significant ownership in a company tackling the next layer of the AI stack. Hard Problems: Work on unsolved problems in agentic reasoning, memory, and reinforcement learning. World‐Class Team: Collaborate with a dense talent cluster of researchers and engineers who have shipped products serving hundreds of millions of users. ...