9 of 9 Reinforcement Learning Jobs in the Thames Valley

Artificial Intelligence Researcher

Hiring Organisation
microTECH Global LTD
Location
Slough, Berkshire, UK
Employment Type
Full-time
permanent position with candidates required to do hybrid working in either Cambridge or London. Our client are looking for AI Researchers specialising in Reinforcement Learning with Human Feedback (RLHF) and Generative AI. In this role, you will design and optimise the algorithms that align large-scale generative models … build the next generation of foundation models Responsibilities: Develop and refine RLHF algorithms for large language and generative models. Research and implement deep reinforcement learning methods (policy gradients, actor-critic, off-policy learning) for model alignment. Train, fine-tune, and evaluate LLMs and diffusion models at scale. ...

Machine Learning Engineer

Hiring Organisation
Brio Digital
Location
High Wycombe, Buckinghamshire, UK
Employment Type
Full-time
Senior Machine Learning Engineer (Generative AI/LLMs) Location: Fully Remote (UK-based) Salary: £75,000 – £100,000 (depending on experience) The Role We're hiring a Senior Machine Learning Engineer to lead the design and productionisation of Generative AI and Large Language Model (LLM) applications. This role … strong influence over architecture, tooling, and the future direction of LLM-powered products. What You'll Be Doing Design, develop, and deploy advanced machine learning and deep learning models into production. Architect scalable LLMOps pipelines on GCP/Vertex AI, including fine-tuning, vector search, and low-latency ...

Machine Learning Engineer

Hiring Organisation
Higher - AI recruitment
Location
Slough, Berkshire, UK
Employment Type
Full-time
partnering with an early-stage, mission-driven company at the intersection of AI and national defence to appoint exceptional Machine Learning Engineers. This fast-growing organisation is transforming mission-critical combat planning and operational decision-making by building next-generation AI software tools for Western forces. Founded … sector knowledge with cutting-edge design and software engineering expertise to deliver state-of-the-art SaaS capabilities leveraging NLP, GenAI, Computer Vision, and Reinforcement Learning technologies. Position location (hybrid): London (Shoreditch) or Paris (Le Marais) We are seeking Machine Learning Engineers who are passionate about using ...

AI Engineering Lead

Hiring Organisation
Akixi
Location
Slough, Berkshire, UK
Employment Type
Full-time
similar conversational-AI platforms. Deep understanding of prompt engineering and fine-tuning of large language models. Strong grounding in ML concepts — supervised, unsupervised, and reinforcement learning. Familiarity with cloud AI/ML services (Azure Cognitive Services, AWS SageMaker, GCP Vertex AI). Experience deploying and monitoring AI workloads ...

AI Engineering Lead

Hiring Organisation
Akixi
Location
Oxford, Oxfordshire, UK
Employment Type
Full-time
similar conversational-AI platforms. Deep understanding of prompt engineering and fine-tuning of large language models. Strong grounding in ML concepts — supervised, unsupervised, and reinforcement learning. Familiarity with cloud AI/ML services (Azure Cognitive Services, AWS SageMaker, GCP Vertex AI). Experience deploying and monitoring AI workloads ...

AI Engineering Lead

Hiring Organisation
Akixi
Location
High Wycombe, Buckinghamshire, UK
Employment Type
Full-time
similar conversational-AI platforms. Deep understanding of prompt engineering and fine-tuning of large language models. Strong grounding in ML concepts — supervised, unsupervised, and reinforcement learning. Familiarity with cloud AI/ML services (Azure Cognitive Services, AWS SageMaker, GCP Vertex AI). Experience deploying and monitoring AI workloads ...

Applied Scientist

Hiring Organisation
Cubiq Recruitment
Location
Slough, Berkshire, UK
Employment Type
Full-time
internal research papers, technical memos, and (where appropriate) external publications. Who You Are Currently completing or recently completed a PhD in Physics, Mathematics, Machine Learning, Computer Science, or a related field from a top university. Strong publication record at leading venues such as NeurIPS, ICML, ICLR, ACL, CVPR, ICCV … EMNLP. Solid understanding of modern ML architectures (transformers, diffusion, retrieval-augmented systems, reinforcement learning, etc.). Strong coding skills in Python and experience with at least one major ML framework (PyTorch, JAX, TensorFlow). Ability to bridge high-level research with practical implementation. Curious, humble, and excited ...

Lead ML Engineer (London)

Hiring Organisation
Glite Tech
Location
Slough, Berkshire, UK
Employment Type
Full-time
English to intermediate and advanced learners. We're on the verge of solving one of the biggest challenges in education – making high-quality, personalised learning accessible to everyone. We are building a fundamental model for education - one that can accurately predict student knowledge and orchestrate lessons, adapting … models to production, to own the ML team in our growing company. What you will do Build fundamental models for education - solving the ultimate learning task of predicting student knowledge and optimal 'next task' Work with a vast amount of unique data - we have data from over 1M language ...

Agentic Developer - Building guardrails for autonomous AI

Hiring Organisation
governr
Location
Slough, Berkshire, UK
Employment Type
Full-time
requirements through first principles • You communicate technical concepts clearly to non-technical stakeholders Highly Valued (Differentiated Candidates) • Publications or research in multi-agent systems, reinforcement learning, AI safety, or agent architectures • Experience at AI labs (Anthropic, OpenAI, DeepMind) or leading AI research groups • Production experience with agents: LangChain … Dr. Ayman Hindy, Marcel Cassard, and leading figures in AI, high frequency risk management and financial regulation. Early team of sharp, mission-driven builders. Learning Curve: You'll gain expertise in cutting-edge AI architectures, enterprise software, regulatory frameworks, and category creation simultaneously. This is one of those roles ...