Claude 4, every mention
23 scenes, the whole family · ← back to Claude 4
tap a year for its mentions
every year anyone Will Brown 5Mike Krieger 4Shawn Wang 3Dylan Davis 2Alessio Fanelli 2Thorsten Ball 1Sarah Sachs 1Sam D'Amico 1Joel Becker 1Dax Reed 1
Verbatim, from the transcripts: the passages where Claude 4 comes up
The Future of AI Infra: from Kubernetes to Agent Sandboxes — Akshat Bubna, Modal CTO
- ▶ 38:00 Akshat Bubna Uh, I think like pre-cloud four they were not, and then now they're able to one-shot.
Notion’s Sarah Sachs & Simon Last on Custom Agents, Evals, and the Future of Work
- ▶ 23:32 Sarah Sachs Let's say that like we want to update, like, uh, you know, they deprecated Sonnet, um, four or whatever it is.
The Stove Guy: Sam D'Amico Shows New AI Cooking Features on America's Most Powerful Stove at Impulse
- ▶ 20:11 Sam D'Amico So I started using, I think this was Sonnet IV, I started using Sonnet IV with Cloud Code, and I was basically went and built an end-to-end fleet telemetry system from, like, Terraform, like, like, all, like, like, infrastructure in…
Why Every Agent Needs a Box — Aaron Levie, Box
- ▶ 32:28 Aaron Levie Um, and you're just seeing, you know, these incredible jumps in almost every single model in its own family of, you know, Opus four, um, you know, Sonnet four, six versus Sonnet four, five.
Measuring Exponential Trends Rising (in AI) — Joel Becker, METR
- ▶ 10:01 Joel Becker Opus-IV.
⚡️Claude Sonnet 4.5 and Anthropic's roadmap for Agents and Developers — Mike Krieger, Anthropic
- ▶ 1:15 Mike Krieger We have more traffic on Sonnet 4.5 than we had on Sonnet four. 2 times in the scene
- ▶ 2:40 Mike Krieger Um, even within cloud code, you know, we had listened to a lot of user feedback on, you know, if Sonnet 3.7 was, uh, too eager, maybe four was lazy in some places and, and really drove down laziness or, you know, when the model's like,…
- ▶ 20:51 Mike Krieger so this, you know, we had a customer and internally, we also got like a 30 hour plus kind of execution versus I think Opus four was seven hours.
Amp: The Emperor Has No Clothes
- ▶ 26:25 Thorsten Ball There's models who might not be as smart, um, as Sonnet four as the main agentic driver, but it might be 10 times as fast.
⚡️Launching Ona: Coding Agent with Fully Sandboxed Cloud Environment
- ▶ 18:51 Shawn Wang Any discoveries from, like, using GPT-V versus Cloud-IV, you know, anything like that?
⚡️OpenCode: Claude Code but Open Source, with Any Model, and frontier TUI - with Dax Reed (@thdxr)
⚡️Using RFT to Build Clinical Superintelligence
- ▶ 5:59 Brendan Fortuna And it's the same technique that's used to train these state-of-the-art reasoning models, like O-three, you know, R-one, uh, Claude-four.
Cline: The Collaborative AI Coder
- ▶ 1:06:42 unnamed speaker But like we've seen with, with like the cloud four release is these frontier model shops, they tend to train on their own application layer, and you might come up with like a very clever tool that in theory would work, work really well,… 2 times in the scene
⚡️Anthropic vs Cognition on Multi-Agents: A Breakdown with Dylan Davis
- ▶ 7:19 Dylan Davis So we're going to have the large model at the beginning, so it's going to be a big, beefy model, like O three, Gemini two dot five pro, Cloud Sonnet Opus, Sonnet, or yeah, Cloud Sonnet four, Cloud Sonnet four Opus, et cetera.
- ▶ 7:19 Dylan Davis So we're going to have the large model at the beginning, so it's going to be a big, beefy model, like O three, Gemini two dot five pro, Cloud Sonnet Opus, Sonnet, or yeah, Cloud Sonnet four, Cloud Sonnet four Opus, et cetera.
⚡️Launching AI Diplomacy: the hardest LLM Game Benchmark yet - Alex Duffy
- ▶ 19:17 Alex Duffy Cloud four coming out.
⚡️Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect
- ▶ 0:15 Shawn Wang And, uh, yeah, honestly, we knew that Cloud Four was coming and we just didn't, we just too busy to, like, have a dedicated episode.
- ▶ 2:50 unnamed speaker Um, and it's not, I mean, I didn't realize that, but extended thinking could not use tools before the way they worded it, and now they can, and in Oppos four, so that's great.
- ▶ 7:21 Will Brown And they had some internal benchmark for this that went from, like, 45% down to 15% for both for Sonnet and for Opus as opposed to three seconds. 2 times in the scene
- ▶ 7:21 Will Brown And they had some internal benchmark for this that went from, like, 45% down to 15% for both for Sonnet and for Opus as opposed to three seconds. 2 times in the scene
- ▶ 16:40 Will Brown And I was like, here are the top 10 things that builders are using in their agentic rag applications with the new groundbreaking clod four.
- ▶ 37:43 Shawn Wang Figure out Cloud Four.
The Winds of AI Winter (Q2 Four Wars of the AI Stack Recap)
- ▶ 1:11:07 Alessio Fanelli Do you think GPT-NEXT and Cloud-IV push it back down because they're coming out with higher intelligence, higher cost? 2 times in the scene