PrometheusRoot
Blog Links Prometheans 100+ AI Books AI Companies Why are you here?
← Prometheans 100+
×
Sholto Douglas
rising
Researcher
X / Twitter
googledeepmindgeminiscaling

Related

pioneer Demis Hassabis
← Prometheans 100+ Sholto Douglas

Anthropic member of technical staff, scaling reinforcement learning

Sholto Douglas

Member of Technical Staff, Scaling RL — Anthropic Researcher, Gemini team — Google DeepMind
Listen — profile
0:00 / 2:52

Profile

Sholto Douglas is one of the most listened-to insider voices explaining how frontier AI models are actually built — not from the outside as a commentator, but from inside the rooms where Gemini and Claude were trained. He is a Member of Technical Staff at Anthropic, where he works on scaling reinforcement learning: the training regime that has, over the past two years, turned language models from autocomplete engines into systems that can reason through competition math, win at competitive programming, and act autonomously across long agentic tasks. If you want to understand why models got dramatically better at coding and reasoning since 2024, Douglas is close to the source of that shift.

His path is unusual. Douglas studied Mechatronic (Space) Engineering at the University of Sydney and was a near-Olympic-level fencer — he placed 21st at the 2017 World Championships — before pivoting into AI. He joined Google DeepMind in 2021 as a research engineer, worked on large-scale language modeling, and became a central figure on the Gemini team, co-leading inference infrastructure and helping shape how those models run efficiently on TPUs at massive scale. He was the person who kicked off and wrote the first version of How to Scale Your Model, DeepMind’s now-widely-cited open textbook on the systems engineering of training and serving large models. He later moved to Anthropic to focus on scaling RL.

What makes Douglas essential for developers isn’t just his résumé — it’s that he can explain the machinery clearly. His long-form conversations with fellow Anthropic researcher Trenton Bricken on the Dwarkesh Podcast, hosted by Dwarkesh Patel, are widely regarded as some of the best available “context dumps” on how modern models are trained, what reinforcement learning actually buys you, how far the current paradigm can scale, and where the bottlenecks are. He is candid about uncertainty in a field full of hype, and specific where most public commentary stays vague.

For someone learning AI today, Douglas is worth following because he sits at the intersection of two things most explanations separate: the low-level systems reality (parallelism, inference cost, hardware constraints) and the high-level capability story (why RL on verifiable tasks is unlocking reasoning and agency). He is bullish on AI progress and on AI coding in particular — he talks openly about the path toward AI “coworkers” — but he grounds that optimism in mechanics rather than vibes. That combination of engineering depth and clear explanation is exactly what makes him a rising, must-listen figure.

Key Articles & Papers

How to Scale Your Model: A Systems View of LLMs on TPUs 2025 — The open textbook Douglas kicked off at DeepMind — how large models actually parallelize and run on real hardware during training and inference. Essential systems-level reading for anyone scaling models.

Videos

YouTube video
YouTube video
YouTube video

Spotify Podcasts

Sam Altman on Codex 5.3 Launch, Anthropic's Sholto Douglas, Alphabet Beats Q4 Estimates | Sam Altman, Sholto Douglas, Daniel Barcelo, Mandy Fields, Ivan Burazin, Scott Rogowsky
Sam Altman on Codex 5.3 Launch, Anthropic's Sholto Douglas, Alphabet Beats Q4 Estimates | Sam Altman, Sholto Douglas, Daniel Barcelo, Mandy Fields, Ivan Burazin, Scott Rogowsky
TBPN
2026
Reviewing the Best AI Apps, Anthropic Unveils Claude 4.5 Opus, Doug DeMuro | Sholto Douglas, Quinn Slack, Alex Stauffer & Alex Shevchenko
Reviewing the Best AI Apps, Anthropic Unveils Claude 4.5 Opus, Doug DeMuro | Sholto Douglas, Quinn Slack, Alex Stauffer & Alex Shevchenko
TBPN
2025
Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
The MAD Podcast with Matt Turck
2025
Elon Musk vs. Donald Trump, AI Day | Shaun Maguire, Mark Chen, Sholto Douglas, Jack Whitaker, Aarush Selvan, Michael Mignano, Oliver Cameron, Delian Asparouhov
Elon Musk vs. Donald Trump, AI Day | Shaun Maguire, Mark Chen, Sholto Douglas, Jack Whitaker, Aarush Selvan, Michael Mignano, Oliver Cameron, Delian Asparouhov
TBPN
2025
Ep 66: Member of Technical Staff at Anthropic Sholto Douglas on Claude 4, Next Phase for AI Coding, and the Path to AI Coworkers
Ep 66: Member of Technical Staff at Anthropic Sholto Douglas on Claude 4, Next Phase for AI Coding, and the Path to AI Coworkers
Unsupervised Learning with Jacob Effron
2025
Is RL + LLMs enough for AGI? — Sholto Douglas & Trenton Bricken
Is RL + LLMs enough for AGI? — Sholto Douglas & Trenton Bricken
Dwarkesh Podcast
2025
AMA: career advice given AGI, how I research ft. Sholto & Trenton
AMA: career advice given AGI, how I research ft. Sholto & Trenton
Dwarkesh Podcast
2025
076 - Sholto Douglas and Trenton Bricken on AI
076 - Sholto Douglas and Trenton Bricken on AI
AI: Unplugged
2024
LW - Notes on Dwarkesh Patel's Podcast with Sholto Douglas and Trenton Bricken by Zvi
LW - Notes on Dwarkesh Patel's Podcast with Sholto Douglas and Trenton Bricken by Zvi
The Nonlinear Library
2024
Sholto Douglas & Trenton Bricken — How LLMs actually think
Sholto Douglas & Trenton Bricken — How LLMs actually think
Dwarkesh Podcast
2024

YouTube

YouTube video
2025
YouTube video
2024

Related People

pioneer Demis Hassabis
© 2026 PrometheusRoot