Research Scientist, Reinforcement Learning, DeepMind
- Hiring Organisation
- Location
- London, UK
with advanced reinforcement learning topics, such as RL for sequence models, post-training, preference-based learning, or agentic systems.Familiarity with modern research stacks (e.g., JAX/Flax or PyTorch) and experience scaling experiments.Ability to be comfortable with scaling methodologies, evaluation techniques, and diagnosing complex failure modes.Ability to push projects … initiative.Strong experimental judgment, including selecting appropriate baselines and designing insightful ablations.Excellent communication skills, with a focus on clear presentation of research results.As an organization, Google maintains a portfolio of research projects driven by fundamental research, new product innovation, product contribution and infrastructure goals, while providing individuals and teams ...