25 of 25 Reinforcement Learning Jobs in the South East

Artificial Intelligence Researcher

Hiring Organisation
microTECH Global LTD
Location
Slough, Berkshire, UK
Employment Type
Full-time
permanent position with candidates required to do hybrid working in either Cambridge or London. Our client are looking for AI Researchers specialising in Reinforcement Learning with Human Feedback (RLHF) and Generative AI. In this role, you will design and optimise the algorithms that align large-scale generative models … build the next generation of foundation models Responsibilities: Develop and refine RLHF algorithms for large language and generative models. Research and implement deep reinforcement learning methods (policy gradients, actor-critic, off-policy learning) for model alignment. Train, fine-tune, and evaluate LLMs and diffusion models at scale. ...

Artificial Intelligence / Machine Learning Engineer

Hiring Organisation
European Tech Recruit
Location
Staines-Upon-Thames, England, United Kingdom
Artificial Intelligence/Machine Learning Engineer Technology Incubation & Innovation Lab Location: Staines-upon-Thames, Surrey, UK Working model: Hybrid 3 days onsite, 2 days remote Fixterm 12 month Contract About Our Client Our client is a global technology leader with a dedicated innovation and incubation lab focused on transforming … with strong emphasis on clinical relevance, regulatory compliance, and measurable impact. Role Overview Our client is seeking an experienced Artificial Intelligence/Machine Learning Engineer with a strong healthcare or biomedical background to join their Technology Incubation & Innovation Lab. The role focuses on developing AI driven predictive healthcare ...

Machine Learning - (Healthcare) - Fixed Term 12 Months

Hiring Organisation
microTECH Global LTD
Location
Egham, England, United Kingdom
these AI solutions to ensure they meet user needs and drive meaningful impact in both healthcare and accessibility domains. Responsibilities: Develop and optimize machine learning models for disease prediction, early diagnosis and personalised healthcare solutions. Process and analyze structured and unstructured health data (EHR, Wearables, HL7/FHIR … implement deep learning algorithms for predictive healthcare applications. Contribute to research on AI-driven personalization strategies to empower users in managing their health effectively. Develop AI-powered accessibility solutions for our products, leveraging multi-modal AI (text, image, audio). Adhere to data privacy regulations (GDPR, MDR, HIPPA, EHDS ...

Machine Learning Engineer (0–3 Years Experience).

Hiring Organisation
IT Graduate Recruitment
Location
London, South East, England, United Kingdom
Employment Type
Full-Time
Salary
£45,000 - £75,000 per annum, OTE
Machine Learning Engineer (LLM/AI Systems) London/Hybrid | 0–3 Years Experience | Competitive Salary Are you obsessed with AI and large language models? We’re an early-stage startup building real-world products powered by LLMs — from intelligent copilots to adaptive automation tools — and we’re looking … research — we give you time and resources to explore, learn, and publish. What We’re Looking For 0–3 years of experience in Machine Learning, Data Science, or NLP/LLM. Strong Python skills; exposure to PyTorch/TensorFlow/Hugging Face. (Bonus) understand fundamentals of deep learning ...

AI/ML Engineer

Hiring Organisation
Brio Digital
Location
High Wycombe, Buckinghamshire, UK
Employment Type
Full-time
/ML Engineer (Generative AI/LLMs) Location: Fully Remote (UK-based) The Role We're hiring a Senior Machine Learning Engineer to lead the design and production of Generative AI and Large Language Model (LLM) applications. This role sits at the heart of an AI-focused engineering team … strong influence over architecture, tooling, and the future direction of LLM-powered products. What You'll Be Doing Design, develop, and deploy advanced machine learning and deep learning models into production. Architect scalable LLMOps pipelines on GCP/Vertex AI, including fine-tuning, vector search, and low-latency ...

AI/ML Engineer

Hiring Organisation
Brio Digital
Location
Woking, Surrey, UK
Employment Type
Full-time
/ML Engineer (Generative AI/LLMs) Location: Fully Remote (UK-based) The Role We're hiring a Senior Machine Learning Engineer to lead the design and production of Generative AI and Large Language Model (LLM) applications. This role sits at the heart of an AI-focused engineering team … strong influence over architecture, tooling, and the future direction of LLM-powered products. What You'll Be Doing Design, develop, and deploy advanced machine learning and deep learning models into production. Architect scalable LLMOps pipelines on GCP/Vertex AI, including fine-tuning, vector search, and low-latency ...

AI/ML Engineer

Hiring Organisation
Brio Digital
Location
Portsmouth, Hampshire, UK
Employment Type
Full-time
/ML Engineer (Generative AI/LLMs) Location: Fully Remote (UK-based) The Role We're hiring a Senior Machine Learning Engineer to lead the design and production of Generative AI and Large Language Model (LLM) applications. This role sits at the heart of an AI-focused engineering team … strong influence over architecture, tooling, and the future direction of LLM-powered products. What You'll Be Doing Design, develop, and deploy advanced machine learning and deep learning models into production. Architect scalable LLMOps pipelines on GCP/Vertex AI, including fine-tuning, vector search, and low-latency ...

AI/ML Engineer

Hiring Organisation
Brio Digital
Location
Dartford, Kent, UK
Employment Type
Full-time
/ML Engineer (Generative AI/LLMs) Location: Fully Remote (UK-based) The Role We're hiring a Senior Machine Learning Engineer to lead the design and production of Generative AI and Large Language Model (LLM) applications. This role sits at the heart of an AI-focused engineering team … strong influence over architecture, tooling, and the future direction of LLM-powered products. What You'll Be Doing Design, develop, and deploy advanced machine learning and deep learning models into production. Architect scalable LLMOps pipelines on GCP/Vertex AI, including fine-tuning, vector search, and low-latency ...

AI/ML Engineer

Hiring Organisation
Brio Digital
Location
Newport, Isle of Wight, UK
Employment Type
Full-time
/ML Engineer (Generative AI/LLMs) Location: Fully Remote (UK-based) The Role We're hiring a Senior Machine Learning Engineer to lead the design and production of Generative AI and Large Language Model (LLM) applications. This role sits at the heart of an AI-focused engineering team … strong influence over architecture, tooling, and the future direction of LLM-powered products. What You'll Be Doing Design, develop, and deploy advanced machine learning and deep learning models into production. Architect scalable LLMOps pipelines on GCP/Vertex AI, including fine-tuning, vector search, and low-latency ...

AI/ML Engineer

Hiring Organisation
Brio Digital
Location
Crawley, West Sussex, UK
Employment Type
Full-time
/ML Engineer (Generative AI/LLMs) Location: Fully Remote (UK-based) The Role We're hiring a Senior Machine Learning Engineer to lead the design and production of Generative AI and Large Language Model (LLM) applications. This role sits at the heart of an AI-focused engineering team … strong influence over architecture, tooling, and the future direction of LLM-powered products. What You'll Be Doing Design, develop, and deploy advanced machine learning and deep learning models into production. Architect scalable LLMOps pipelines on GCP/Vertex AI, including fine-tuning, vector search, and low-latency ...

AI/ML Engineer

Hiring Organisation
Brio Digital
Location
Brighton, East Sussex, UK
Employment Type
Full-time
/ML Engineer (Generative AI/LLMs) Location: Fully Remote (UK-based) The Role We're hiring a Senior Machine Learning Engineer to lead the design and production of Generative AI and Large Language Model (LLM) applications. This role sits at the heart of an AI-focused engineering team … strong influence over architecture, tooling, and the future direction of LLM-powered products. What You'll Be Doing Design, develop, and deploy advanced machine learning and deep learning models into production. Architect scalable LLMOps pipelines on GCP/Vertex AI, including fine-tuning, vector search, and low-latency ...

Machine Learning Engineer

Hiring Organisation
Higher - AI recruitment
Location
Slough, Berkshire, UK
Employment Type
Full-time
partnering with an early-stage, mission-driven company at the intersection of AI and national defence to appoint exceptional Machine Learning Engineers. This fast-growing organisation is transforming mission-critical combat planning and operational decision-making by building next-generation AI software tools for Western forces. Founded … sector knowledge with cutting-edge design and software engineering expertise to deliver state-of-the-art SaaS capabilities leveraging NLP, GenAI, Computer Vision, and Reinforcement Learning technologies. Position location (hybrid): London (Shoreditch) or Paris (Le Marais) We are seeking Machine Learning Engineers who are passionate about using ...

AI Engineering Lead

Hiring Organisation
Akixi
Location
Dartford, Kent, UK
Employment Type
Full-time
similar conversational-AI platforms. Deep understanding of prompt engineering and fine-tuning of large language models. Strong grounding in ML concepts — supervised, unsupervised, and reinforcement learning. Familiarity with cloud AI/ML services (Azure Cognitive Services, AWS SageMaker, GCP Vertex AI). Experience deploying and monitoring AI workloads ...

AI Engineering Lead

Hiring Organisation
Akixi
Location
Basingstoke, Hampshire, UK
Employment Type
Full-time
similar conversational-AI platforms. Deep understanding of prompt engineering and fine-tuning of large language models. Strong grounding in ML concepts — supervised, unsupervised, and reinforcement learning. Familiarity with cloud AI/ML services (Azure Cognitive Services, AWS SageMaker, GCP Vertex AI). Experience deploying and monitoring AI workloads ...

AI Engineering Lead

Hiring Organisation
Akixi
Location
Slough, Berkshire, UK
Employment Type
Full-time
similar conversational-AI platforms. Deep understanding of prompt engineering and fine-tuning of large language models. Strong grounding in ML concepts — supervised, unsupervised, and reinforcement learning. Familiarity with cloud AI/ML services (Azure Cognitive Services, AWS SageMaker, GCP Vertex AI). Experience deploying and monitoring AI workloads ...

AI Engineering Lead

Hiring Organisation
Akixi
Location
Oxford, Oxfordshire, UK
Employment Type
Full-time
similar conversational-AI platforms. Deep understanding of prompt engineering and fine-tuning of large language models. Strong grounding in ML concepts — supervised, unsupervised, and reinforcement learning. Familiarity with cloud AI/ML services (Azure Cognitive Services, AWS SageMaker, GCP Vertex AI). Experience deploying and monitoring AI workloads ...

AI Engineering Lead

Hiring Organisation
Akixi
Location
Guildford, Surrey, UK
Employment Type
Full-time
similar conversational-AI platforms. Deep understanding of prompt engineering and fine-tuning of large language models. Strong grounding in ML concepts — supervised, unsupervised, and reinforcement learning. Familiarity with cloud AI/ML services (Azure Cognitive Services, AWS SageMaker, GCP Vertex AI). Experience deploying and monitoring AI workloads ...

AI Engineering Lead

Hiring Organisation
Akixi
Location
High Wycombe, Buckinghamshire, UK
Employment Type
Full-time
similar conversational-AI platforms. Deep understanding of prompt engineering and fine-tuning of large language models. Strong grounding in ML concepts — supervised, unsupervised, and reinforcement learning. Familiarity with cloud AI/ML services (Azure Cognitive Services, AWS SageMaker, GCP Vertex AI). Experience deploying and monitoring AI workloads ...

AI Engineering Lead

Hiring Organisation
Akixi
Location
Newport, Isle of Wight, UK
Employment Type
Full-time
similar conversational-AI platforms. Deep understanding of prompt engineering and fine-tuning of large language models. Strong grounding in ML concepts — supervised, unsupervised, and reinforcement learning. Familiarity with cloud AI/ML services (Azure Cognitive Services, AWS SageMaker, GCP Vertex AI). Experience deploying and monitoring AI workloads ...

AI Engineering Lead

Hiring Organisation
Akixi
Location
Crawley, West Sussex, UK
Employment Type
Full-time
similar conversational-AI platforms. Deep understanding of prompt engineering and fine-tuning of large language models. Strong grounding in ML concepts — supervised, unsupervised, and reinforcement learning. Familiarity with cloud AI/ML services (Azure Cognitive Services, AWS SageMaker, GCP Vertex AI). Experience deploying and monitoring AI workloads ...

AI Engineering Lead

Hiring Organisation
Akixi
Location
Brighton, East Sussex, UK
Employment Type
Full-time
similar conversational-AI platforms. Deep understanding of prompt engineering and fine-tuning of large language models. Strong grounding in ML concepts — supervised, unsupervised, and reinforcement learning. Familiarity with cloud AI/ML services (Azure Cognitive Services, AWS SageMaker, GCP Vertex AI). Experience deploying and monitoring AI workloads ...

Applied Scientist

Hiring Organisation
Cubiq Recruitment
Location
Slough, Berkshire, UK
Employment Type
Full-time
internal research papers, technical memos, and (where appropriate) external publications. Who You Are Currently completing or recently completed a PhD in Physics, Mathematics, Machine Learning, Computer Science, or a related field from a top university. Strong publication record at leading venues such as NeurIPS, ICML, ICLR, ACL, CVPR, ICCV … EMNLP. Solid understanding of modern ML architectures (transformers, diffusion, retrieval-augmented systems, reinforcement learning, etc.). Strong coding skills in Python and experience with at least one major ML framework (PyTorch, JAX, TensorFlow). Ability to bridge high-level research with practical implementation. Curious, humble, and excited ...

Senior ML Infrastructure Engineer Robotics

Hiring Organisation
Harnham - Data & Analytics Recruitment
Location
London, South East, England, United Kingdom
Employment Type
Contractor
Contract Rate
£600 - £1,000 per day
heart of their technology. This is a hands-on engineering role focused on scale, performance, and reliability. You will work across the full machine learning lifecycle, from distributed training pipelines to highly optimised inference systems deployed into production robotics environments. The Role You will join a highly technical team … working at the intersection of software engineering, machine learning infrastructure, and robotics. Your focus will be on turning cutting-edge models into robust, production-ready systems that run efficiently across cloud and constrained hardware environments. You will collaborate closely with researchers and ML engineers, help shape architectural decisions ...

Lead ML Engineer (London)

Hiring Organisation
Glite Tech
Location
Slough, Berkshire, UK
Employment Type
Full-time
English to intermediate and advanced learners. We're on the verge of solving one of the biggest challenges in education – making high-quality, personalised learning accessible to everyone. We are building a fundamental model for education - one that can accurately predict student knowledge and orchestrate lessons, adapting … models to production, to own the ML team in our growing company. What you will do Build fundamental models for education - solving the ultimate learning task of predicting student knowledge and optimal 'next task' Work with a vast amount of unique data - we have data from over 1M language ...

Agentic Developer - Building guardrails for autonomous AI

Hiring Organisation
governr
Location
Slough, Berkshire, UK
Employment Type
Full-time
requirements through first principles • You communicate technical concepts clearly to non-technical stakeholders Highly Valued (Differentiated Candidates) • Publications or research in multi-agent systems, reinforcement learning, AI safety, or agent architectures • Experience at AI labs (Anthropic, OpenAI, DeepMind) or leading AI research groups • Production experience with agents: LangChain … Dr. Ayman Hindy, Marcel Cassard, and leading figures in AI, high frequency risk management and financial regulation. Early team of sharp, mission-driven builders. Learning Curve: You'll gain expertise in cutting-edge AI architectures, enterprise software, regulatory frameworks, and category creation simultaneously. This is one of those roles ...