DeepMind logo

Research Scientist/Engineer, Frontier Reasoning, DeepMind

DeepMind·London·Hybrid·Posted 7 days ago

Google DeepMind is a London-headquartered AI research lab behind AlphaGo and AlphaFold. They're hiring a Research Scientist or Engineer in Frontier Reasoning to advance the reasoning abilities of large models, based in London. This AI engineering role sits within the UK's frontier AI sector.

Minimum qualifications:

  • Bachelor's or Master's degree in Computer Science, Mathematics, Physics, a related quantitative field, or equivalent practical experience.
  • 4 years of experience building, scaling, and debugging machine learning models using deep learning frameworks (e.g., JAX, PyTorch, or TensorFlow).
  • Experience in at least one core area: Reinforcement Learning (RL), Post-Training (SFT/RLHF/RLAIF), Agentic Tool-Use, or Inference-Time Search.

Preferred qualifications:

  • PhD in Computer Science, Machine Learning, Physics, or a related quantitative field.
  • Experience designing asynchronous agent-environment simulation loops or large distributed post-training pipelines.
  • Experience prototyping new hypotheses quickly while keeping shared codebases clean, robust, and production-grade.

About the job

At Google DeepMind, the PRISM (Planning, Reasoning, Inference & Structured Models) team brings together researchers and engineers to advance the frontiers of AI reasoning and autonomous agentic systems. We reject the false tradeoff between research and execution, pursuing breakthroughs on open AI challenges while embedding directly into core teams to land those capabilities in production.
Our work powers Gemini & Gemma—developing core reasoning, multi-agent capabilities, RL and inference scaling in Gemini, and spearheading Gemma 270M. We deliver critical contributions to AI Grand Challenges (such as our gold medal-winning IMO 2025 effort), drive product innovations like 'Deep Think' mode and agentic inference scaling in antigravity, and contribute to Alphabet-wide initiatives including AI for Science and Project Big Sleep.
You will operate across the full research-and-engineering lifecycle, developing distributed post-training infrastructure and algorithms that enable Gemini models to solve complex, multi-step problems autonomously.

Artificial intelligence will be one of humanity’s most transformative inventions. At Google DeepMind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing AI development to solve complex global challenges and accelerate high-quality product innovation for billions of users. We use our technologies for widespread public benefit and scientific discovery, ensuring safety and ethics are always our highest priority.

We are pushing the boundaries across multiple domains. Our global teams offer diverse learning opportunities and varied career pathways for those driven to achieve exceptional results through collective effort. Individual pay is determined by factors including job-related skills, experience, and relevant education or training.

US: $207000 - $300000 (USD) + 20% bonus target + equity + benefits

Learn more about benefits at Google.

Responsibilities

  • You will operate across the full research-and-engineering lifecycle of frontier reasoning and agentic systems.
  • Bridge research to production: Tackle unsolved problems in agentic reasoning, turning early exploratory prototypes into hardened production features for Gemini releases.
  • Build and scale systems: Architect and optimize distributed post-training pipelines and agent-environment simulation loops across thousands of accelerators.
  • Run scientific ablations: Design rigorous experiments and failure analyses to isolate performance bottlenecks and communicate findings through clear write-ups.
  • Drive technical excellence: Maintain high code quality and architectural health across shared Reinforcement Learning and modeling codebases.
DeepMind logo

About DeepMind

Google DeepMind is an AI research laboratory focused on building artificial general intelligence. Founded in London in 2010 by Demis Hassabis, Shane Legg, and Mustafa Suleyman, the company was acquired by Google in 2014 for approximately $660M. DeepMind pioneered breakthroughs including AlphaGo, the first program to defeat a world champion at Go, and AlphaFold, which solved the 50-year-old protein folding problem and has been used by over two million researchers worldwide. The lab remains headquartered in London with a large UK engineering presence and continues to push the boundaries of reinforcement learning, large language models, and scientific AI.

Stage: AcquiredLondon
View DeepMind profile →