1 to 25 of 101 Permanent Reinforcement Learning Jobs in England

Research Scientist/Engineer - General Decision & Control Agent

Hiring Organisation
Adecco
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£100,000 - £180,000 per annum
Research Scientist/Engineer - Agent Systems & Reinforcement Learning Location: London Salary: £ per annum + permanent benefits + bonus Job Type: Permanent, Full-Time, On-site About the Opportunity We are partnering with a leading AI research organisation focused on developing sustainable, generalisable and evolvable Agent systems that represent … periods of execution. This is an exceptional opportunity to join a world-class research environment working at the intersection of Agents, Large Language Models, Reinforcement Learning and Autonomous Systems , contributing to cutting-edge research that could play a significant role in advancing the path towards Artificial General Intelligence ...

Research Scientist/Engineer - General Decision & Control Agent

Hiring Organisation
Adecco
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£100000 - £180000/annum + perm benefits +bonus
Research Scientist/Engineer - Agent Systems & Reinforcement Learning Location: London Salary: £(phone number removed) per annum + permanent benefits + bonus Job Type: Permanent, Full-Time, On-site About the Opportunity We are partnering with a leading AI research organisation focused on developing sustainable, generalisable and evolvable Agent … periods of execution. This is an exceptional opportunity to join a world-class research environment working at the intersection of Agents, Large Language Models, Reinforcement Learning and Autonomous Systems , contributing to cutting-edge research that could play a significant role in advancing the path towards Artificial General Intelligence ...

Applied AI ML Engineer Director – NLP / LLM and Graphs

Location
Greater London, England, United Kingdom
making. The CDAO is also responsible for developing and implementing solutions that support the firm's commercial goals by harnessing artificial intelligence and machine learning technologies to develop new products, improve productivity, and enhance risk management effectively and responsibly. As an Applied AI ML Director - NLP/… Graphs within the Chief Data & Analytics Office, Machine Learning Centre of Excellence, you will have the opportunity to apply sophisticated machine learning methods to complex tasks including natural language processing, graph analytics, speech analytics, time series, reinforcement learning and recommendation systems. You will collaborate with various ...

Research Scientist, Robotics, DeepMind

Location
Greater London, England, United Kingdom
designing new architectures, our research scientists work on real-world problems that span the breadth of computer science, such as machine (and deep) learning, data mining, natural language processing, hardware and software performance analysis, improving compilers for mobile platforms, as well as core search and much more. … applications and work with real robots inside and outside the lab to manage real-world use cases. A strong algorithmic background in scalable machine learning (e.g., reinforcement learning/imitation learning; multimodal foundation models) and experience with real robots/robot simulation and training setups ...

Applied AI ML Lead Engineer- (NLP/LLM/Graph)

Location
Greater London, England, United Kingdom
Description NLP/LLM Scientist – Applied AI ML Lead – Machine Learning Centre of Excellence The Machine Learning Center of Excellence invites the successful candidate to apply sophisticated machine learning methods to a wide variety of complex tasks including natural language processing, large language models, and recommendation systems. … environment together with the business, technologists and control partners to deploy solutions into production. The candidate must also have a strong passion for machine learning and invest independent time towards learning, researching and experimenting with new innovations in the field. The candidate must have solid expertise in Deep ...

Applied AI ML Director - NLP / LLM and Graphs

Location
Greater London, England, United Kingdom
making. The CDAO is also responsible for developing and implementing solutions that support the firm’s commercial goals by harnessing artificial intelligence and machine learning technologies to develop new products, improve productivity, and enhance risk management effectively and responsibly. As an Applied AI ML Director - NLP/… Graphs within the Chief Data & Analytics Office, Machine Learning Centre of Excellence, you will have the opportunity to apply sophisticated machine learning methods to complex tasks including natural language processing, graph analytics, speech analytics, time series, reinforcement learning and recommendation systems. You will collaborate with various ...

Principal Machine Learning Engineer, AI & Data Platforms (AiDP)

Location
Greater London, England, United Kingdom
Principal Machine Learning Engineer, AI & Data Platforms (AiDP) London, England, United Kingdom Corporate Functions At Apple, we build AI systems that define experiences for billions of people and we do it with an unwavering commitment to privacy, performance, and craft. The AI & Data Platforms (AiDP) team is seeking … Principlal Machine Learning Engineer to lead the design, fine‐tuning, evaluation, and productionisation of large language models and generative internal AI systems at global scale. This is a deeply hands‐on, high‐impact role: you will work across the full model lifecycle, from reinforcement learning and upstream ...

Research Engineer

Location
Oxford, England, United Kingdom
Role Research Engineer Salary Competitive Contract Perm Location Oxford The Role As our Research Engineer, you will sit at the boundary between frontier machine‐learning research and shipped product. You will take ideas from the research edge, reinforcement learning, RLHF, multi‐objective policy optimisation and LLM post … customers depend on, while connecting that research to OD’s product vision and to what our customers actually need. Responsibilities Apply and adapt machine‐learning and reinforcementlearning research to OD’s product roadmap, translating research direction into shippable capability. Design, train and evaluate models using reinforcement ...

Research Engineer / Scientist, Post-training - London

Location
Greater London, England, United Kingdom
seeking those who are dedicated as much to building safely and responsibly as to advancing disruptive agentic capabilities. We promote a mindset of openness, learning, and collaboration, where everyone has something to contribute. About the Research & Models Team The Models team builds the foundational models that power our cutting … perceive, understand, and act within complex environments. We own the entire pipeline including synthetic data generation, environment design, mid-training, supervised fine-tuning, offline reinforcement learning, online reinforcement learning, reward modelling, transition modelling, etc. Our team also has dedicated MLOps, Infra and Inference support at scale. ...

Senior Reinforcement Learning Engineer

Location
Oxford, England, United Kingdom
Gravis Robotics is a startup turning heavy construction machines into intelligent and autonomous robots. Our unique combination of learning-based automation and augmented remote control enables a single operator to safely manage a fleet of earthmoving machines in a gamified environment. With over a decade of academic experience … working with physical robots, navigating the challenges of sim-to-real (sim2real) transfer, and deploying robotic systems into production environments. What you will do Learning-Based Planning and Control for Real Systems Develop data driven planning and control systems for autonomous excavation that generalize across machine models and soil ...

Research Scientist, Agent Post-Training

Location
Greater London, England, United Kingdom
backbone of upcoming releases. You will help bridge research and engineering by designing scalable experiments and building reliable infrastructure for tool use and reinforcement learning. You will value various experience and backgrounds to create extraordinary impact.Artificial intelligence will be one of humanity’s most transformative inventions. At Google DeepMind … scientific discovery, ensuring safety and ethics are always our highest priority. We are pushing the boundaries across multiple domains. Our global teams offer various learning opportunities and varied career pathways for those driven to achieve exceptional results through collective effort.ResponsibilitiesLead the full research process, from forming hypotheses to delivering ...

Senior Research Scientist FMTA

Location
Greater London, England, United Kingdom
large language models. We focus on developing algorithms and systems that align pre-trained models with tasks and performance goals through techniques like reinforcement learning. As a research-driven team, we stay up to date with current literature to integrate cutting-edge ideas into our core stack. As part … controllability, and safer, more effective user experiences. Your responsibilities As a Senior Research Scientist, you'll design, implement, and deploy cutting-edge research in reinforcement learning and post-training at scale, driving innovations that make it into production. Build and deploy state-of-the-art reinforcement learning ...

Senior Research Scientist FMTA

Location
Greater London, England, United Kingdom
large language models. We focus on developing algorithms and systems that align pre-trained models with tasks and performance goals through techniques like reinforcement learning. As a research-driven team, we stay up to date with current literature to integrate cutting-edge ideas into our core stack. As part … controllability, and safer, more effective user experiences. Your responsibilities As a Senior Research Scientist, you’ll design, implement, and deploy cutting-edge research in reinforcement learning and post-training at scale, driving innovations that make it into production. You will: Build and deploy state-of-the-art reinforcement ...

Machine Learning Scientist - Reinforcement Learning

Hiring Organisation
Roc Search
Location
London, UK
Employment Type
Full-time
Description We’re working with an innovative technology business that is looking to hire a Machine Learning Scientist with strong Reinforcement Learning experience to join their growing team. This is an exciting opportunity for someone who wants to work on real-world AI problems, developing models that … move beyond research and simulation into practical applications. The Role You’ll be working on challenging machine learning and control problems, with a particular focus on Reinforcement Learning. The role will involve: Developing and improving Reinforcement Learning models and agents Working with real-world data ...

Machine Learning Engineer - Agentic AI & Reinforcement Learning

Location
Crawley, England, United Kingdom
solutions that efficiently and responsibly resolve complex natural resource, digital, energy transition and infrastructure challenges. We are looking for a hands‐on Machine Learning Engineer to design, develop and evaluate intelligent agent systems, turning emerging AI methods into reliable, scalable and reusable solutions for truly real world data. About … Team Learn more about our work and the team on our website, recent blog post featuring one of our senior machine learning engineers and our latest Voices of Viridien. Key responsibilities Design and develop single-agent and multi-agent systems for complex workflows. Build reasoning, planning, task decomposition, memory ...

Research Scientist, Human Data, Robotics, DeepMind

Location
London, United Kingdom
designing new architectures, our research scientists work on real-world problems that span the breadth of computer science, such as machine (and deep) learning, data mining, natural language processing, hardware and software performance analysis, improving compilers for mobile platforms, as well as core search and much more. … scientific discovery, ensuring safety and ethics are always our highest priority. We are pushing the boundaries across multiple domains. Our global teams offer varied learning opportunities and career pathways for those driven to achieve exceptional results through collective effort. Minimum qualifications: PhD in Computer Science, a related field ...

Research Scientist, Human Data, Robotics, DeepMind

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
designing new architectures, our research scientists work on real-world problems that span the breadth of computer science, such as machine (and deep) learning, data mining, natural language processing, hardware and software performance analysis, improving compilers for mobile platforms, as well as core search and much more. … scientific discovery, ensuring safety and ethics are always our highest priority. We are pushing the boundaries across multiple domains. Our global teams offer varied learning opportunities and career pathways for those driven to achieve exceptional results through collective effort. Minimum qualifications: PhD in Computer Science, a related field ...

Research Scientist, Human Data, Robotics, DeepMind

Location
Greater London, England, United Kingdom
Scientist, Human Data, Robotics, DeepMind DeepMind London, UK PhD in Computer Science, a related field, or equivalent practical experience. 2 years of experience with reinforcement learning and imitation learning. Experience with vision, vision-language, video, and other multimodal models. Experience with multimodal generative modeling, training and inference. Preferred … designing new architectures, our research scientists work on real-world problems that span the breadth of computer science, such as machine (and deep) learning, data mining, natural language processing, hardware and software performance analysis, improving compilers for mobile platforms, as well as core search and much more. Artificial intelligence ...

Sr Data Scientist, AD/ADAS

Location
Greater London, England, United Kingdom
systems. To deliver reliable data-driven technologies to millions of Toyota vehicles, we are solving complex real-world problems using large-scale data, machine learning, and state-of-the-art architectures for Perception, Prediction, and Motion Planning. WHO ARE WE LOOKING FOR? The team is looking for a strong … Autonomy team. You will combine modern data approaches with safety standards while also considering cost efficient options. You will be an expert in deep learning and software engineering. Furthermore, you are proactive towards handling the processes required for production development and approach them by asking "What ...

Machine Learning Scientist — Large Multimodal Models (Post-Training)

Location
West of England, England, United Kingdom
SUMMARY We are seeking a Machine Learning Scientist to join the Enchant team at Iambic Therapeutics. Our mission is to deliver better medicines through innovation in AI-based discovery technologies. In this role, you will research and develop post-training methods for Enchant - our multimodal transformer model trained … discovery. The role centers on designing and evaluating post-training approaches for large multimodal language models including supervised fine-tuning, parameter-efficient fine-tuning, reinforcement learning, preference or reward-based optimization, and other emerging post-training methods. You will develop rigorous evaluations and training infrastructure that make ...

Senior Machine Learning Engineer - Messaging Platform

Location
Greater London, England, United Kingdom
moment. We’re evolving how messaging works at Spotify — moving from short-term optimization toward systems that understand long-term user journeys. By combining reinforcement learning approaches with deeper domain signals, we’re expanding how machine learning shapes the entire messaging funnel. What You’ll Do Design … build, and ship machine learning models that optimize messaging across push, email, and in-app channels Plan and run A/B experiments in a multi-objective environment, balancing conversion, engagement, retention, and reachability Contribute to reinforcement learning systems that optimize for long-term user outcomes rather ...

Senior AI Platform Engineer

Hiring Organisation
Capgemini
Location
Greater London, United Kingdom
Employment Type
Full Time
small senior team with heavy AI leverage. The ambition runs past serving frontier models: we close the loop from production feedback through reinforcement learning and fine-tuning, and train our own LLMs and SLMs where evaluations and economics justify it. What you will own One or more platform … guardrail engine, tool and agent registry, or shared product services The training and adaptation loop: pipelines that turn production traces and evaluation verdicts into reinforcement learning and fine-tuning datasets, and the infrastructure to train, evaluate, and serve NewCo-tuned LLMs and SLMs behind the same gates ...

Senior Research Scientist | Model Steering

Location
Greater London, England, United Kingdom
deliver perfect translations for the most demanding use cases. To that end, we take responsibility for the entire life cycle of the machine learning models that power our language AI products. This includes data, training, quality assurance, and operational aspects. In our highly collaborative teams, each person … drive impact across the company. Your responsibilities We are looking for a Senior Research Scientist to lead fine-tuning, post-training, model-steerability, and reinforcement learning for the next generation of DeepL's LLM-based translation models. This is a high-impact, hands‐on role for a researcher ...

Senior Research Scientist | Multimodal Systems

Location
Greater London, England, United Kingdom
deliver perfect translations for the most demanding use cases. To that end, we take responsibility for the entire life cycle of the machine learning models that power our language AI products. This includes data, training, quality assurance, and operational aspects. In our highly collaborative teams, each person … scope to drive impact across the company. Your responsibilities We are looking for a Senior Research Scientist to lead fine-tuning, post-training, and reinforcement learning for the next generation of DeepL's document translation multimodal and vision models. This is a high-impact, hands-on role ...

Senior Research Scientist | Multimodal Systems

Location
Greater London, England, United Kingdom
Your responsibilities We are looking for a Senior Research Scientist to lead fine-tuning, post-training, and reinforcement learning for the next generation of DeepL's document translation multimodal and vision models. This is a high-impact, hands-on role for a researcher who can own a major … hands‐on research and development on post‐training for our vision and/or multimodal models: supervised fine‐tuning, knowledge distillation, preference optimization, and reinforcement learning tuned to translation quality. Build evaluator models for document and design quality, including rubric‐ and reference‐based grading, and investigate and mitigate ...