151 to 165 of 165 Reinforcement Learning Jobs in England

Field Engineer Deployed Engineering London

Location
Greater London, England, United Kingdom
meaningful impact in the world. Our work frequently takes us right up to the state of the art in technical innovation, be it reinforcement learning, distributed systems, generative AI, or deployment infrastructure. The defence industry is entering the most exciting phase of the technological development curve. Advances ...

RL Environment Architect & Data Research Engineer

Location
Greater London, England, United Kingdom
Eigent AI is seeking an RL Environment Data Engineer/Researcher to design and refine reinforcement learning training environments. This role emphasizes data collection, task definition, and the implementation of anti-reward-hacking mechanisms. The ideal candidate will have strong Python coding skills and a solid understanding … reinforcement learning. Responsibilities include collaborating with multiple teams to improve RL tasks and validation environments. #J-18808-Ljbffr ...

Prototyping Engineer Hardware Engineering London; Oxford

Location
Oxford, England, United Kingdom
testing, identifying integration bottlenecks, developing custom hardware solutions, and keeping Helsing's developmental fleet operational and ready for flight. Your work directly enables faster learning cycles and more capable platforms, with tangible impact on the pace and quality of Helsing's hardware development programme. The day‐to‐day Design … safety, and ethical considerations are vital. Our work frequently takes us right up to the state‐of‐the‐art in technical innovation, be it reinforcement learning, distributed systems, generative AI, or deployment infrastructure. The defence industry is entering the most exciting phase of the technological development curve. Advances ...

Prototyping Engineer Hardware Engineering London; Oxford

Location
Greater London, England, United Kingdom
testing, identifying integration bottlenecks, developing custom hardware solutions, and keeping Helsing's developmental fleet operational and ready for flight. Your work directly enables faster learning cycles and more capable platforms, with tangible impact on the pace and quality of Helsing's hardware development programme. The day‐to‐day Design … safety, and ethical considerations are vital. Our work frequently takes us right up to the state‐of‐the‐art in technical innovation, be it reinforcement learning, distributed systems, generative AI, or deployment infrastructure. The defence industry is entering the most exciting phase of the technological development curve. Advances ...

Senior Model-Steering Scientist for LLM Translation

Location
Greater London, England, United Kingdom
DeepL is seeking a Senior Research Scientist to lead fine-tuning, post-training, model steerability, and reinforcement learning for the next generation of translation models. You will fuse high-value human data with synthetic data and drive breakthroughs from research to production at scale. You will work with ...

Senior RL Robotics Controls Engineer - Remote/Hybrid

Location
Greater London, England, United Kingdom
Elysium Robotics is seeking a Senior Robotic Controls & Reinforcement Learning Simulation Engineer to join their innovative team in the UK. This role focuses on developing and optimizing control algorithms for advanced robotic systems using proprietary elastic actuators. Your responsibilities will include training robust control policies, managing simulation infrastructures ...

RL Engineer — Robotic Manipulation & Real-World Training

Location
Greater London, England, United Kingdom
Humanoid in London is seeking a Reinforcement Learning Engineer to join our Autonomy team. You will leverage RL in simulation and physical environments to develop robust manipulation policies for our humanoid platform. You will build and test RL pipelines, collaborate with teleoperations, testing, and operations, and explore ...

Remote RLHF Program Manager: AI Alignment

Location
Cheltenham, England, United Kingdom
Coaley Peak is seeking an RLHF Manager to lead reinforcement learning from human feedback across its internal AI engines and external client projects. You will design reward models, manage human evaluator programmes, and evaluate model alignment against business objectives and UK regulatory expectations. The role is dual-facing ...

Robotics Research Scientist: AI Agents & Real-World Robots

Location
Greater London, England, United Kingdom
will design, implement and evaluate large models for robotic agents, while collaborating across teams and contributing to open research. This role emphasizes scalable ML, reinforcement learning, and real-world robot integration. You will work with state-of-the-art robotics platforms, prototype applications, and cutting-edge datasets ...

Strategic Project Lead

Hiring Organisation
Huzzle.com
Location
City of London, London, United Kingdom
Strategic Project Lead Huzzle Labs · Full-time · Palo Alto, London, Berlin, Remote About Huzzle Labs Huzzle Labs is an applied research lab building reinforcement learning environments for frontier AI labs. Our work covers computer use, enterprise workflows and code: the long, multi-application tasks that make up most ...

Clinical Engineer: AI-Driven Healthcare Impact

Location
City Of London, England, United Kingdom
Clinical Engineer to leverage medical expertise and technical skills. This role involves building innovative products that enhance patient care using AI and reinforcement learning. No coding experience required, but a strong aptitude is necessary. Ideal candidates are eager to solve complex problems and work efficiently in a collaborative environment. ...

Robotics Software Engineer

Hiring Organisation
Som3
Location
London, United Kingdom
Employment Type
Permanent
Salary
£50000 - £60000/annum
delivering high-quality robotics software within deadlines. Desirable Experience Experience in any of the following would be highly beneficial: Path planning, motion control or reinforcement learning-based control. Simulation tools such as Gazebo, Isaac Sim or MATLAB/Simulink. Robotics frameworks and interfaces including ROS/ROS2, Serial ...

Product Program Manager

Location
Greater London, England, United Kingdom
complex robotic or mechatronic hardware , including supplier qualification and manufacturing scale‐up. Familiarity with simulation-to-real transfer, digital twin practices, or reinforcementlearning-based control as used in modern humanoid development. What We Offer Competitive equity: stock options with meaningful upside as we scale. 30+ paid days ...

Research Scientist/Engineer (Science of Scheming)

Location
Greater London, England, United Kingdom
cognition and actively spend time trying to understand how they think. Experience RL‐training LLMs: You have hands‐on experience in training LLMs via reinforcement learning. You have encountered and resolved countless painful issues from GPU failures to debugging learning instabilities. Strong analytical skills: You bring rigorous quantitative ...

Robotics Research Scientist, Human Data & Multimodal AI

Location
Greater London, England, United Kingdom
robotic models, integrating human data and simulations to advance real-world robot capabilities. You will publish findings and contribute to multi-disciplinary teams, applying reinforcement and imitation learning, vision-language models, and generative techniques. A passion for translating lab research into production robotics is essential. #J-18808-Ljbffr ...