Sholto Douglas

AI Researcher, Anthropic · 1 appearance on the record.

computed by AI from the episodes · how this works → · full disclaimer →

scientistengineerathlete@_sholtodouglas ↗LinkedIn ↗sholtodouglas.github.io ↗

Sholto Douglas is an artificial intelligence researcher at Anthropic specializing in reinforcement learning, test-time compute, and autonomous coding agents behind models like Claude 4.5 Sonnet. He previously worked at Google DeepMind on the Gemini inference stack and co-authored foundational work on scaling transformer inference.

49statements → 26claims → 8claims resolved → 63%fully supported → 3.71/5average certainty → 2.39/5average debate potential → 4.5/5argument clarity · the sources → 11said about them ↓

5 supported 2 partly supported 1 contradicted 1 not yet assessed 17 not checkable as stated how the 26 claims stand · each chip opens the sources

7 predictions · 19 assertions · 3 opinions · 14 insights · 5 disclosures · 1 what if · every statement was checked. The predictions and assertions are the 26 claims: statements the public record can support or contradict. 8 are resolved, 1 is not yet assessed, and 17 name no date, number or outcome precise enough to check. Everything else (opinions, insights, what ifs, disclosures) can never be settled by the record, so it carries no assessment.

The record, in short

What the tape says about how Sholto argues and how the claims held up. Everything they said, and everything said about them, is in the tabs below.

Their most notable supported claim

Assertion Supported
Douglas: AI autonomous task execution time horizons double every six months
“And so I think it's like every couple of months, the time horizon that the AIs are capable of doing is doubling or something, something crazy. Maybe, maybe every six months the time horizon doubles”
Sholto Douglas Oct 2, 2025 ▶ 40:13 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)

Their most notable contradicted claim

Assertion Contradicted
Douglas: Anthropic models autonomously replicated the Claude.ai website in hours
“And in this case, the model replicated Claude.ai with artifacts, with everything else I can't quite remember how long that one took. Maybe a couple hours to do.”
Sholto Douglas Oct 2, 2025 ▶ 42:36 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)

Expressed certainty vs assessment result

none yet certainty 1
none yet certainty 2
75% certainty 3
75% certainty 4
none yet certainty 5

weighted support: a fully supported claim counts one, a partly supported claim counts half. Each filled bar is clickable and opens exactly those claims; "none yet" means nothing said at that certainty level has resolved yet

Argument clarity: do they answer the question? how? →

4.5 / 5 directness 4.7 · coherence 4.8 · precision 4.4 · compression 4

redirected or did not address 1 of 13 assessed questions (8%). Watch them ▸

This is a score against a rubric. It is not a rank. Every host question → answer exchange is scored with names hidden on directness, coherence, precision and compression, 1–5 each, on meaning alone: disfluencies are ignored, and only raw unedited episodes count. This is the score that measures thought. Every scored exchange, scores shown → · The rubric and its checks →

How they sound: speaking style how? →

280 words/min while actually speaking · 43 um and uh per 1k words

Measured by listening to the audio itself: 11,044 words across 1 episode of raw-level tape, transcribed verbatim with every um and uh kept, each one attributed only where the alignment onto our timed stream is unambiguous. These are measurements of speaking style. We do not rank them: across this corpus, fluency and argument quality are nearly uncorrelated (ρ≈0.2), and smooth talking does not signal clear thinking. How it's measured →

Everything Sholto Douglas said on the MAD Podcast that made the record, most notable first. Filter by type, assessment or year in the ledger →

Prediction Not checkable as stated
Douglas: Anthropic believes AGI is reachable in a couple of years
“We think that, you know, AGI is within reach in the next couple of years.”
Sholto Douglas Oct 2, 2025 ▶ 27:33 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Prediction Open · timeframe Oct 2028
Douglas: AI industry will reach human-level computer capabilities in 2-3 years
“Which is that in the next two or three years, given the right feedback loops, given the right compute, given the right, you know, elbow grease and this kind of stuff, we think that we as the AI industry are all on track to create something that is at least as …”
Sholto Douglas Oct 2, 2025 ▶ 59:06 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Assertion Not checkable as stated
Douglas: Transformers successfully model any domain given sufficient data and compute
“I don't think that's true. I think we haven't yet really found anything that transformers haven't been able to model provided sufficient data and sufficient compute.”
Sholto Douglas Oct 2, 2025 ▶ 1:00:25 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Insight
Douglas: Independent technical blogs are the highest AI hiring signals
“The fastest route, or like, the most immediate one is whenever we see a really good blog post where people have, like, done incredible amount of work in an independent fashion, it's one of the highest signal things there is.”
Sholto Douglas Oct 2, 2025 ▶ 10:54 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Prediction Not checkable as stated
Douglas predicts DeepMind will lead the world in AI science discoveries
“DeepMind, if you wanted to solve science, is the best place in the world. Like, I think that DeepMind will directly contribute to more scientific discoveries from AI than anything else, right?”
Sholto Douglas Oct 2, 2025 ▶ 16:54 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Disclosure
Douglas: Anthropic's ethos is that scaling current techniques achieves AGI
“Like really for the last five or six years, Anthropics ethos has been scaling compute with broadly the current set of techniques is like AGI is tractable within those bounds.”
Sholto Douglas Oct 2, 2025 ▶ 27:52 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Insight
Douglas: AI takeoff speed depends on AI assisting AI research
“We think that one of the most important signals of whether or not we are basically the speed of takeoff, the speed of progress is driven by how much AI is able to assist AI research.”
Sholto Douglas Oct 2, 2025 ▶ 29:18 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Opinion
Douglas: Claude 3.5 Sonnet drove product-market fit for code editor Cursor
“In many ways, this model is what caused PMF for Cursor. Cursor took off like a rocket, right, with because they were in the right place, and they were able to capitalize on that model as offering a coding experience that didn't previously exist.”
Sholto Douglas Oct 2, 2025 ▶ 34:37 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Assertion Not checkable as stated
Douglas: An Anthropic AI agent operated autonomously for 30 hours building apps
“We asked it to build something that looks roughly like a chat app, you know, something like Slack or, you know. And it was, it, the model just worked for 30 hours. Like, it was just spinning there on a computer for 30 hours, and came out with a really good wor…”
Sholto Douglas Oct 2, 2025 ▶ 36:03 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Assertion Supported
Douglas: AI autonomous task execution time horizons double every six months
“And so I think it's like every couple of months, the time horizon that the AIs are capable of doing is doubling or something, something crazy. Maybe, maybe every six months the time horizon doubles”
Sholto Douglas Oct 2, 2025 ▶ 40:13 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Prediction Not checkable as stated
Douglas: AI application development will see another massive leap next year
“Over the next six months, over the next year, expect dramatic progress here. And like look at where we are now versus where we were a year ago. And the difference is I expect the same jump basically.”
Sholto Douglas Oct 2, 2025 ▶ 42:57 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Assertion Not checkable as stated
Douglas: AI coding interventions stem from taste, not raw programming capability
“Right now you need to intervene quite frequently, but it's usually on questions of taste rather than it is questions of, like, raw programming ability.”
Sholto Douglas Oct 2, 2025 ▶ 43:54 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Insight
Douglas: Solving AI hallucinations intrinsically requires reinforcement learning
“Saying, I don't know, or solving, you know, hallucinations is, ah, intrinsically requires reinforcement learning in many ways.”
Sholto Douglas Oct 2, 2025 ▶ 49:35 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Insight
Douglas: AI reasoning strategies emerge naturally with enough compute and RL feedback
“Give it math questions, tell it whether it got them right or wrong, and the model will learn. This is, it comes down to a bit of lesson in scale and search, is just allow the model to search, have enough compute to run the experiments, and the model actually e…”
Sholto Douglas Oct 2, 2025 ▶ 55:19 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Opinion
Douglas: LLM pipelines are two and a half years of desperate effort
“When I look at an LLM training pipeline, it is two and a half years of best effort, last minute, desperate effort. And there's just so much room to go on every part of it.”
Sholto Douglas Oct 2, 2025 ▶ 1:03:26 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Prediction Not checkable as stated
Douglas: AI beating GDP benchmarks won't immediately alter the broader economy
“We'll probably reach like better than human on the GDP eval, and it won't change anything economically because It'll be all the connective tissue, and all the, like, you know, the context, and actually, like, the task won't be representative.”
Sholto Douglas Oct 2, 2025 ▶ 1:04:53 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Prediction Not checkable as stated
Douglas: Individuals will manage 24/7 AI agent teams within two years
“If coding agents progress in the way I've been saying, in a year or two, you'll be able to manage a team, basically, that works 24 seven for you doing work.”
Sholto Douglas Oct 2, 2025 ▶ 1:06:11 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Assertion Not checkable as stated
Douglas: Robotic locomotion is essentially solved using basic reinforcement learning
“Locomotion's kind of solved, to be honest, with basic RL.”
Sholto Douglas Oct 2, 2025 ▶ 1:08:09 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Assertion Not checkable as stated
Douglas: The post-ChatGPT AI compute supercycle begins properly in 2025
“So finally, this year is where the compute, like, super cycle is, like, beginning properly in effect.”
Sholto Douglas Oct 2, 2025 ▶ 2:40 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Assertion Not checkable as stated
Douglas: Google LLM inference stack saved hundreds of millions in months
“This ended up saving several hundred million dollars, I think, like, even over the first six months”
Sholto Douglas Oct 2, 2025 ▶ 14:18 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Disclosure
Douglas: Anthropic focuses on alignment and near-term economic coding utility
“Anthropic has been laser focused on on two things. One is, like, long-term AI alignment, and two is near-term economic impact. So Anthropik has been laser-focused on coding and computer use, and things that we think will make a direct impact to the economy, li…”
Sholto Douglas Oct 2, 2025 ▶ 17:25 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Disclosure
Douglas: Anthropic intentionally deprioritized math reasoning unlike OpenAI and DeepMind
“You know, one thing that Anthropik, like, You noticeably hasn't focused on compared to DeepMind and to OpenAI is is mathematical reasoning, right? DeepMind and OpenAI have been pursuing mathematical reasoning because of the implications for science and for sci…”
Sholto Douglas Oct 2, 2025 ▶ 17:47 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Insight
Douglas: Compute scale consistently wipes out clever human priors in AI
“Generations of people have developed clever methods of encoding priors about how they think an artificial intelligence should reason. And encoding it into the model, and all of this gets wiped out by scale and, you know, planning, like, basically, like, search…”
Sholto Douglas Oct 2, 2025 ▶ 20:40 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Assertion Not checkable as stated
Douglas: Even AI pioneer Noam Shazeer only sees 10% of ideas work
“I once asked this question of Noam Chazier. And he was like, yeah, maybe like 10% of my ideas work, and that's not, right? You know, one of the, you know, an absolute genius, one of the best in the field. So if only 10% of his ideas work, then I think that, yo…”
Sholto Douglas Oct 2, 2025 ▶ 24:47 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)

Show 24statements(25 left)

The other half of the tape: Sholto Douglas's own voice is left out of every number here. Other people bring the name up 11 times in 1 episode on the MAD Podcast. every mention, with the transcript →

Who brings them up most Matt Turck 11

Every mention by year

tap a year for its mentions
00811512026episodesmentions
0112026episodes it came up in
007.50.51512026episodesmentions per episode
2026 11 mentions in 1 episode

Appearances (1)

EpisodeDateSpeaking time
Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic) Oct 2, 2025 52m
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.