Claude 3.7 Sonnet, every mention
31 scenes · ← back to Claude 3.7 Sonnet
tap a year for its mentions
every year anyone Pratik Bhavsar 7Alessio Fanelli 6Will Brown 5Kat Wu 3David Hershey 3Thorsten Ball 2Sujay Jayakar 2Sarah Sachs 2Charles Packer 2Axel Backlund 2
Verbatim, from the transcripts: the passages where Claude 3.7 Sonnet comes up
Why AI Labs With Unlimited GPUs Still Fail — Anjney Midha, AMP
- ▶ 46:53 Shawn Wang Whatever, uh, three seven?
When AI Agents Run Businesses — Lukas Petersson and Axel Backlund of Andon Labs
- ▶ 42:41 Axel Backlund Like, why is Opus 4.7 here, like, way better than everyone else? 2 times in the scene
- ▶ 49:04 unnamed speaker So you would say it's a step up from four, six to four, seven. 2 times in the scene
Devin’s 80% Moment: Background Agents, 7x PRs, & End of Hand-Held Coding — Walden Yan & Cole Murray
- ▶ 3:18 unnamed speaker Like, um, there was this, the moment, which was, I think, Sonnet three seven, where, like, you guys rewrote Devin in one night or something? 2 times in the scene
Notion’s Sarah Sachs & Simon Last on Custom Agents, Evals, and the Future of Work
- ▶ 23:44 Sarah Sachs Three seven just got deprecated. 2 times in the scene
Steve Yegge's Vibe Coding Manifesto: Why Claude Code Isn't It & What Comes After the IDE
- ▶ 31:26 Steve Yegge Well, so, uh, look, as soon as open source models get to the point where they're as good as Cloud Sonnet three seven was, then you turn on Klein or something, and you've got something that's as good as Cloud Code was in March, which wasn't…
⚡️Claude Sonnet 4.5 and Anthropic's roadmap for Agents and Developers — Mike Krieger, Anthropic
- ▶ 2:40 Mike Krieger Um, even within cloud code, you know, we had listened to a lot of user feedback on, you know, if Sonnet 3.7 was, uh, too eager, maybe four was lazy in some places and, and really drove down laziness or, you know, when the model's like,…
Amp: The Emperor Has No Clothes
- ▶ 0:59 Thorsten Ball I mean, I'll start, you can jump in, but basically I came back to SoftSquare February, and then this was when cloud three five, three seven happened too. 2 times in the scene
⚡️OpenCode: Claude Code but Open Source, with Any Model, and frontier TUI - with Dax Reed (@thdxr)
🕰️ The Oral History of Windsurf (ft. Varun Mohan, Scott Wu, Jeff Wang, Kevin Hou, Anshul R)
⚡️Ranking Agentic LLMs — Pratik Bhavsar, Galileo
- ▶ 10:58 Pratik Bhavsar But then when we released the leaderboard and just in a week that launched 3.7, 3 times in the scene
- ▶ 23:05 Pratik Bhavsar We are trying to evaluate, uh, 3.7 with GPT four or mini, right? 3 times in the scene
- ▶ 32:53 Pratik Bhavsar Like let's say I generate by 3.7 and then check on 4.1
The Utility of Interpretability — Emmanuel Amiesen
- ▶ 45:19 Vibhu (Viboo) There was, like, a phase where people were basically saying, you know, 3.5 and 3.7 are just now
The AI Coding Factory
- ▶ 23:13 Matan Grinberg And then we, let's say when we upgraded from Sonnet, 3.5 to 3.7, we suddenly had a lot of developers being like, Hey, wait, it now does this less, or it does this more what's happening.
- ▶ 27:48 Eno Reyes So for example, Sonnet 3.7 clearly has, uh, it smells like cloud code, right?
⚡️Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect
- ▶ 6:21 Will Brown In terms of complexity of agents, I think the one thing that to me was really nice to see, I haven't, like, done too much testing myself yet, but in their reported benchmarks, the reward hacking issues, like, Sonnet three seven loves to,… 4 times in the scene
- ▶ 17:15 Will Brown I mean, I think coding with these models, especially like quad three, I did a fair amount, like for a few weeks, I was doing a lot of quad code with three seven, mostly
Claude Code: Anthropic's CLI Agent
- ▶ 54:58 Kat Wu The latest Sonnet three seven is, it's a very persistent model. 3 times in the scene
- ▶ 1:00:02 Alessio Fanelli And then you have 3.7.
Zed Agents — with Zed Cofounders Nathan Sobo & Antonio Scandurra
- ▶ 17:39 Antonio Scandurra Like, we've seen this pattern in Cloud 3.7, which is kind of different from the other models where, like, it, it really loves to gather up context first, and then it starts doing edits. 2 times in the scene
Why Every Agent needs Open Source Cloud Sandboxes
- ▶ 14:24 Alessio Fanelli And yeah, I think you can kind of see the slope, especially from like the Sonnet Ruben seven release.
Sleep-Time Compute — Letta AI (Charles Packer, Charlie Snell, Kevin Lin)
- ▶ 26:27 Charles Packer It's like pretty consistent across like both 3.7 deep seek, three mini, which all like the way you actually scale the x-axis here is fundamentally quite different in each case with 3.7 extended thinking mode. 2 times in the scene
Claude Plays Pokémon Hackathon: Escape from Mt. Moon!
- ▶ 2:40 David Hershey Uh, and so then, uh, a couple months ago, when we were, uh, finalizing 3.7 on it, I played with it, and you got, like, the first squint of, like, signs of life of something happening.
- ▶ 14:01 David Hershey Uh, I've tested, like, all sorts of the extended thinking mode with, ah, 3.7 on it, and, like, it doesn't really help.
- ▶ 38:11 unnamed speaker Um, I also think that if I was to, like, plug this into cloud 3.7, like, just taking my code, it would probably do better. 3 times in the scene
The #1 SWE-Bench Verified Agent
- ▶ 3:37 Alessio Fanelli And when you say the sequential thinking part is interesting because it's not a reasoning mode, do you have, like, do you have a good way to, like, just summarize why the sequential thinking MCP perform better than just like reasoning mode… 2 times in the scene
Fullstack-Bench: The Eval for Coding Agents — with Sujay Jayakar, Chief Scientist, Convex
- ▶ 17:23 Sujay Jayakar Like, for example, we just tried clod three seven and it performs worse than clod three five on convex evals with the same prompting. 2 times in the scene
How Claude Plays Pokémon was made
- ▶ 0:59 Alessio Fanelli So Sonnet Truebund Seven came out a couple of weeks ago. 2 times in the scene
- ▶ 22:05 unnamed speaker I'm curious, um, as you switched from 3.5 to 3.7 and sort of reasoning models, were there any degradations there? 3 times in the scene
- ▶ 26:27 David Hershey 3.7 sonnet that I've seen is like, it will have like meta commentary on what it's good at and bad at and it's knowledge base.