This role has closed 1w ago. It is no longer on Google's board, so there is nothing left to apply to. The posting is kept here because you saved it or opened it; it is a record, not an offer.

Verified from career page · Posted 1w ago

Google

Research Scientist/Engineer, Frontier Reasoning, DeepMind

Google · DeepMind

London

Mid-level

Machine learningPyTorchTensorFlow

Last seen 1w ago

Posted
1w ago

Posted on 21 September 2026

Workplace
Not specified

Work model not stated

Salary
Not disclosed

Salary range not shared by the company

Visa sponsorship
Not specified

Visa sponsorship details unknown

This role has closed 1w ago ago. It's kept as a record — see Google's open roles or the similar live roles below.

See Google's open roles

At DeepMind, the Planning, Reasoning, Inference & Structured Models (PRISM) team brings together researchers and engineers to advance the frontiers of AI reasoning and autonomous agentic systems. We reject the false tradeoff between research and execution, pursuing breakthroughs on open AI challenges while embedding directly into core teams to land those capabilities in production.

Our work powers Gemini & Gemma—developing core reasoning capabilities and RL scaling for Gemini 3, and leading Gemma 3 270M, including multi-agent Gemini capabilities. We deliver critical contributions to AI Grand Challenges (such as our gold medal-winning IMO 2025 effort), drive product innovations like 'deep think' mode and agentic inference scaling in antigravity, and lead Alphabet-wide initiatives including AI for Science and Project Big Sleep.

In this role, you will operate across the full research-and-engineering lifecycle, developing distributed post-training infrastructure and algorithms that enable Gemini models to solve complex, multi-step problems autonomously.

Artificial intelligence will be one of humanity’s most transformative inventions. At Google DeepMind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing AI development to solve complex global challenges and accelerate high-quality product innovation for billions of users. We use our technologies for widespread public benefit and scientific discovery, ensuring safety and ethics are always our highest priority.

We are pushing the boundaries across multiple domains. Our global teams offer diverse learning opportunities and varied career pathways for those driven to achieve exceptional results through collective effort.

Individual pay is determined by factors including job-related skills, experience, and relevant education or training.

US: $207000 - $300000 (USD) + 20% bonus target + equity + benefits

Learn more about benefits at Google.

Operate across the full research-and-engineering lifecycle of frontier reasoning and agentic systems.

Work on unsolved problems in agentic reasoning, turning early exploratory prototypes into hardened production features for Gemini releases.

Architect and optimize distributed post-training pipelines and agent-environment simulation loops across thousands of accelerators.

Design rigorous experiments and failure analyses to isolate performance bottlenecks and communicate findings through clear write-ups.

Maintain high code quality and architectural health across shared reinforcement learning and modeling codebases.

Minimum qualifications:

Bachelor's degree in Computer Science, Mathematics, Physics, a related quantitative field, or equivalent practical experience.

4 years of experience building, scaling, and debugging machine learning models using deep learning frameworks (e.g., JAX, PyTorch, or TensorFlow).

Experience in one core area: Reinforcement Learning (RL), Post-Training (SFT/RLHF/RLAIF), Agentic Tool-Use, or Inference-Time Search.

Preferred qualifications:

PhD in Computer Science, Machine Learning, Physics, or a related quantitative field.

Experience training and managing models on large-scale distributed accelerator clusters (e.g., TPUs or GPUs).

Experience designing asynchronous agent-environment simulation loops or large distributed post-training pipelines.

Experience prototyping new hypotheses quickly while keeping shared codebases clean, robust, and production-grade.

About Google

Google runs core product engineering out of London, Dublin, Zurich, Warsaw and Munich — not support functions. Zurich is one of its largest engineering sites anywhere, and Warsaw has grown substantially. The bar is high and the process is long, but these are genuine product teams.

Similar roles

  1. PRO members only Pro Posted today
  2. PRO members only Pro Posted today
  3. PRO members only Pro Posted today
  4. PRO members only Pro Posted today
  5. PRO members only Pro Posted today
  6. PRO members only Pro Posted today
4,532 more roles like this. Open the board →