Claude

includes Claude Code, Claude Cowork, Claude 3.7 Sonnet, Claude 3, Claude Sonnet, Claude Opus, Claude Haiku, Claude Artifacts, Claude 4, Claude Desktop, Claude Mythos, Claude Computer Use and 22 more

86 statements across 41 episodes · 41 bullish · 19 bearish · 39 people on the record · first statement Aug 31, 2023 by Eugene Cheah · said 1,448 times in 170 episodes since 2023 · across every show →

Mentions by year, the whole family

brought up most by Shawn Wang (270), Alessio Fanelli (143), Felix Rieseberg (96), Boris Cherny (71), David Hershey (50), Kat Wu (35), Doug O'Laughlin (24), Karina Nguyen (23)

tap a year for its mentions
0040040800802023202420252026episodesmentions
040802023202420252026episodes it came up in
007.54015802023202420252026episodesmentions per episode
2026 504 mentions in 56 episodes 9 per episode
2025 765 mentions in 76 episodes 10 per episode
2024 171 mentions in 35 episodes 5 per episode
2023 8 mentions in 3 episodes 3 per episode

every mention, scene by scene, with the transcript →

Everything said about Claude, oldest first

Aug 31, 2023
Insight
Cheah: AI Engineers Do Not Need ML Math to Build Products
“Frankly, for an AI engineer, you don't need it. You, your main thing that you needed to do was to, frankly, just play around with ChatGPT, or all the alternatives, be aware of the alternatives, because be very mercenary, swap out to Cloudia if it's better for …”
Eugene Cheah Aug 31, 2023 ▶ 1:33:57 RWKV: Reinventing RNNs for the Transformer Era
Jan 11, 2024 neutral
Opinion
Lambert: Frontier Labs Lack Visibility into Cross-Model RLHF Sensitivity
“I think big labs are so over-indexed, are indexed on their own base models, so they don't know, like, what's swapping between CloudBase or GPT-IV-Base, how that would change any notion of preference or what you do with RLHF.”
Nathan Lambert Jan 11, 2024 ▶ 1:30:15 The Origin and Future of RLHF: the secret ingredient for ChatGPT - with Nathan Lambert
Jan 11, 2024 neutral
Insight
Lambert: Claude's constitution dictates output priorities, not model beliefs
“If you look at Claude's constitution, like, that doesn't mean the model believes these things. It's just trying Trained and to prioritize these things.”
Nathan Lambert Jan 11, 2024 ▶ 13:06 The Origin and Future of RLHF: the secret ingredient for ChatGPT - with Nathan Lambert
Apr 27, 2024 neutral
Insight
Bach: Claude is implemented as an invariant pattern similarly to consciousness
“Claude exists only as a pattern. It's something that is a pattern in the activation of the transistors. And even transistors don't actually exist. They are A pattern in the atoms that we are able to see as an invariance because we tune the atoms in a particula…”
Joscha Bach Apr 27, 2024 ▶ 1:20:37 This World Does Not Exist — Joscha Bach, Karan Malhotra, Rob Haisfield (WorldSim, WebSim, Liquid AI)
Apr 27, 2024 positive
Insight
Haisfield: Opus handles much more complexity than Sonnet for web generation
“Like sonnet will still create things that kind of like floor you sometimes. Opus can just handle much more complexity. I'd say it is the big heuristic there.”
Rob Haisfield Apr 27, 2024 ▶ 46:08 This World Does Not Exist — Joscha Bach, Karan Malhotra, Rob Haisfield (WorldSim, WebSim, Liquid AI)
Apr 27, 2024 neutral
Assertion Supported
Malhotra: Claude's system prompt is written in the third person
“With Claude, we notice the system prompt is written in third person. It's written in third person. It's written as, the assistant is X, Y, Z.”
Karan Malhotra Apr 27, 2024 ▶ 9:40 This World Does Not Exist — Joscha Bach, Karan Malhotra, Rob Haisfield (WorldSim, WebSim, Liquid AI)
Apr 27, 2024 positive
Assertion Supported
Haisfield: Claude tolerates imperfect URL syntax when generating WebSim apps
“Like, you don't need to get the exact syntax of an actual URL. Claude's smart enough to figure it out.”
Rob Haisfield Apr 27, 2024 ▶ 36:58 This World Does Not Exist — Joscha Bach, Karan Malhotra, Rob Haisfield (WorldSim, WebSim, Liquid AI)
Apr 27, 2024 neutral
Insight
Bach: Pointing out behavioral contradictions gives Claude more conversational freedom
“You can point this out to Claude that a lot of the assumptions that it has in its behavior are actually inconsistent with the communicative goals that it has in this situation. It leads it to notice these inconsistencies and gives it more degrees of freedom.”
Joscha Bach Apr 27, 2024 ▶ 1:50:14 This World Does Not Exist — Joscha Bach, Karan Malhotra, Rob Haisfield (WorldSim, WebSim, Liquid AI)
Apr 27, 2024 positive
Assertion Supported
Haisfield: WebSim accurately generates external RSS feeds to fetch live news data
“It just hallucinated a correct RSS feed and brought that in to its into this, I guess, you know, this wasn't a part of its like context window or anything because it's just displaying this stuff.”
Rob Haisfield Apr 27, 2024 ▶ 51:16 This World Does Not Exist — Joscha Bach, Karan Malhotra, Rob Haisfield (WorldSim, WebSim, Liquid AI)
Aug 2, 2024 positive
Disclosure
Alessio Fanelli canceled ChatGPT subscription, switching to Claude for podcast workflows
“I canceled chat GBD a while ago. Really small podcaster run for Latent Space. It runs both on Claude and on OpenAI and Claude is definitely better most of the time.”
Alessio Fanelli Aug 2, 2024 ▶ 7:48 The Winds of AI Winter (Q2 Four Wars of the AI Stack Recap)
Aug 17, 2024 bullish
Prediction Not checkable as stated
Howard: Reka's model is probably superior to GPT and Claude for certain tasks
“There's a whole model that's been trained in a different way. So there's probably a whole lot of tasks it's probably better at than you know, GPT and Gemini and Claude.”
Jeremy Howard Aug 17, 2024 ▶ 36:42 Answer.ai & AI Magic with Jeremy Howard
Nov 28, 2024 negative
Insight
Schluntz: JSON Escaping Overhead Degrades LLM Performance Across the Board
“Like if you're trying to output a code in JSON, there's a lot of extra escaping that needs to be done. And that actually hurts model performance across the board. Where versus like if you're in just a single XML tag, there's none of that sort of escaping that …”
Erik Schluntz Nov 28, 2024 ▶ 36:44 The new Claude 3.5 Sonnet, Computer Use, and Building SOTA Agents — with Erik Schluntz, Anthropic
Dec 13, 2024 positive
Disclosure
Mohan: Windsurf uses Claude for planning, proprietary models for retrieval and diffs
“No, so actually the way it works is the high-level planning that is going on in the model is actually getting done with products like the Cloud. But the extremely fast retrieval, as well as the ability to, like, take the high-level plan and actually apply it t…”
Varun Mohan Dec 13, 2024 ▶ 25:04 Windsurf: The Enterprise AI IDE
Dec 25, 2024 negative
Opinion
Neubig: AI Agents Including Claude Are Ineffective At Asking For Help
“I think it, my impression is that agents are not very good at asking for help, even Claude. So like when they ask for help, they'll ask for help when they don't need it and then won't ask for help when they do need it.”
Graham Neubig Dec 25, 2024 ▶ 34:08 Best of 2024 in Agents (from #1 on SWE-Bench Full, Prof. Graham Neubig of OpenHands/AllHands)
Dec 25, 2024 positive
Opinion
Neubig: Claude Is The Best Agent Model, Open Models Lag Behind
“I still am under the impression that Claude is the best. The other closed models are, you know, not quite as good, and then the open models are a little bit behind that.”
Graham Neubig Dec 25, 2024 ▶ 15:16 Best of 2024 in Agents (from #1 on SWE-Bench Full, Prof. Graham Neubig of OpenHands/AllHands)
Dec 25, 2024 positive
Opinion
Neubig: GPT Loops On Errors While Claude Tries New Approaches
“So, like, GPT doesn't have very good air recovery ability. And so, because of this, it will go into loops and do the same thing over and over and over again, whereas Claude does not do this.”
Graham Neubig Dec 25, 2024 ▶ 14:25 Best of 2024 in Agents (from #1 on SWE-Bench Full, Prof. Graham Neubig of OpenHands/AllHands)
Dec 25, 2024 negative
Insight
Neubig: Inadequate Information Gathering Is The Biggest Agent Failure Mode
“So I think actually probably the biggest thing that it fails at is. Or that our agent plus Claude fails at is insufficient information gathering before trying to solve the task, and so if you provide all, if you provide instructions that it should do informati…”
Graham Neubig Dec 25, 2024 ▶ 43:42 Best of 2024 in Agents (from #1 on SWE-Bench Full, Prof. Graham Neubig of OpenHands/AllHands)
Jan 1, 2025 positive
Assertion Supported
Swyx: Claude wrapper Bolt.new reached $20M ARR
“The other one would be Bolt. There's a straight quad wrapper. And again, another now they've announced twenty million ARR, which is another step up from our eight million that we put on the title.”
Shawn Wang Jan 1, 2025 ▶ 44:32 2024 Year in Review: The Big Scaling Debate, the Four Wars of AI, Top Themes and the Rise of Agents
Feb 1, 2025 negative
Assertion Not checkable as stated
Nguyen: Stanford HELM benchmark under-reported Claude performance due to improper prompting
“This has happened with, like, Stanford, I remember, like, when Stanford had lists also, like, they were, like, running benchmarks. Yeah, Helm. And somehow, like, Claude was, like, always, like, not performing well, and that's because, like, the way they prompt…”
Karina Nguyen Feb 1, 2025 ▶ 16:39 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
Feb 1, 2025 neutral
Disclosure
Nguyen: Explored a Claude collaborative workspace concept at Anthropic in 2023
“I was working on something similar to, like, Canvas-y, but for Claude at that time, in, like, twenty-twenty-three, it was the same similar idea of, like, Claude workspace where a human and a Claude could have, like, a shared workspace which is like a document.”
Karina Nguyen Feb 1, 2025 ▶ 9:27 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
Feb 28, 2025 bullish
Prediction Not checkable as stated
Klein: Consumer AI agents and tools like Anthropic MCP will run locally
“There's a case where a lot of people are going desktop first for, you know, consumer use. And I think Claude is doing a lot of this where I expect to see, you know, MCP is really oriented around the Claude desktop app for a reason, right? Like I think a lot of…”
Paul Klein Feb 28, 2025 ▶ 30:22 Browserbase: Browser Infrastructure For Your AI Agents
Mar 4, 2025 bearish
Prediction Held up
Hershey predicts the Claude stream won't reach Victory Road within 16 days
“I think we have a little ways before we can beat the game in 16 days. I do not have a lot of faith that the current stream is gonna, gonna be standing in Victory Road in 13 days.”
David Hershey Mar 4, 2025 ▶ 32:42 How Claude Plays Pokémon was made
Mar 4, 2025 positive
Insight
AI agents have an optimal effective context length where intelligence peaks
“I think like one thing you see a lot when you talk to people building agents is there's like some effective context length that actually like has the model be the smartest. And that seems to vary slightly model by model, but for this model, for whatever purpos…”
David Hershey Mar 4, 2025 ▶ 18:57 How Claude Plays Pokémon was made
Mar 4, 2025 neutral
Insight
Prompting Claude cannot improve its spatial navigation without explicit instructions
“You can try to prompt Quad a lot of different ways to understand how to navigate better, and anything short of telling it exactly what to do does not improve its, like, actual navigation.”
David Hershey Mar 4, 2025 ▶ 19:55 How Claude Plays Pokémon was made
Mar 4, 2025 neutral
Assertion Supported
Anthropic's Pokémon research graph reflects a single run passing Lt. Surge
“The run that you saw that's, like, on the graph we put out alongside, like, in our research blog is, like, a single run that I have watched, like, get through At least surges Jim. And then it got a little past that. And the reason that that's where we stopped …”
David Hershey Mar 4, 2025 ▶ 33:55 How Claude Plays Pokémon was made
Mar 4, 2025
Disclosure
Upgrading Claude models in Pokémon primarily involves deleting prompt scaffolding
“Literally every model that has come out with Pokemon, like, the main change that I have made to this agent is deleting prompt stuff.”
David Hershey Mar 4, 2025 ▶ 22:51 How Claude Plays Pokémon was made
Mar 4, 2025 positive
Assertion Not checkable as stated
Prompting Claude to nickname its Pokémon caused it to exhibit protective behaviors
“And one thing we found when we started doing that is it got more protective of the Pokemon it nicknamed. Like, it's pretty obvious, like, when it catches a Pokemon, Now that it has a nickname, it will, like, go heal it right away if it's hurt, and that did not…”
David Hershey Mar 4, 2025 ▶ 25:13 How Claude Plays Pokémon was made
Mar 4, 2025 negative
Assertion Not checkable as stated
Claude aggressively hallucinates game zone transitions without explicit negative feedback
“Claude will, like, pretty aggressively hallucinate that it succeeded in transitioning between zones if you don't, like, tell it did not.”
David Hershey Mar 4, 2025 ▶ 11:09 How Claude Plays Pokémon was made
Mar 4, 2025 negative
Assertion Not checkable as stated
Claude still struggles with spatial awareness and visual positioning on screen
“Quad doesn't particularly understand, like, the middle of a Game Boy screen and a whole bunch of concepts like that, which means, like, you can prompt all around everywhere, but, like, this kind of, like, spatial awareness and where something is with respect t…”
David Hershey Mar 4, 2025 ▶ 14:05 How Claude Plays Pokémon was made
Mar 4, 2025 negative
Assertion Not checkable as stated
Claude spent 12 hours overnight mistaking a Pokémon doormat for a textbox
“I once saw it see like a red box on the screen that was like the doormat and think it was a text box and spend 12 hours pressing A overnight to try to clear the text box, which you see that happen once and you add in some helpful reminders to not do that.”
David Hershey Mar 4, 2025 ▶ 11:44 How Claude Plays Pokémon was made
Mar 14, 2025 positive
Opinion
Snipd CEO: Claude is the best model at phrasing and personality
“Like, in my opinion, Claude is the best one when it comes to the way it formulates things.”
Kevin Ben-Smith Mar 14, 2025 ▶ 51:20 Snipd: The AI Podcast App for Learning — with CEO Kevin Ben-Smith
Apr 5, 2025 negative
Insight
Hershey: Prompt engineering cannot fix Claude's visual comprehension limitations
“Vision is, like, pretty beyond fixing with a prompt. Again, go for it. Have fun. I've spent a lot of hours, like, overlaying grids, overlaying images, stretching, compressing, contrast colors, all sorts of stuff.”
David Hershey Apr 5, 2025 ▶ 12:54 Claude Plays Pokémon Hackathon: Escape from Mt. Moon!
Apr 5, 2025 neutral
Assertion Not checkable as stated
Hershey: Chain-of-thought between tool calls only improves agent progress 10%
“Forcing it to think between tool calls. It does help, but not much. It's, like, 10% faster progress if you, like, have it use a lot of chain of thought tokens every time it chooses an action.”
David Hershey Apr 5, 2025 ▶ 14:08 Claude Plays Pokémon Hackathon: Escape from Mt. Moon!
Apr 5, 2025 negative
Opinion
Hershey: Claude is exceptionally bad at point-to-point spatial navigation
“The thing that Claude's the worst at is understanding how to get from point A to point B. It's, like, really, really god-awful at hitting the buttons to go from point A to point B on a screen.”
David Hershey Apr 5, 2025 ▶ 3:24 Claude Plays Pokémon Hackathon: Escape from Mt. Moon!
Apr 5, 2025
Disclosure
Hershey Restricts Multi-Button Inputs After Claude Overwrote Ivysaur's Attack
“People who have watched the stream often complain about the fact that I don't let Claude press multiple buttons when there's dialogue on the screen anymore. That's because I watched a run, or a run, where it had an Ivysaur in Mt. Moon, and it had Tackle as its…”
David Hershey Apr 5, 2025 ▶ 7:58 Claude Plays Pokémon Hackathon: Escape from Mt. Moon!
Apr 5, 2025
Assertion Not checkable as stated
Hershey: Eight Historical Images Yields Peak Performance for Game Agents
“At least in my harness, something like eight historical images is the best performance. More than that, and you start flirting with the space where, like, you get drop off in performance from, like, just having more tokens in the context, which is a thing that…”
David Hershey Apr 5, 2025 ▶ 23:52 Claude Plays Pokémon Hackathon: Escape from Mt. Moon!
Apr 5, 2025
Assertion Not checkable as stated
Hershey: Anthropic Study Found Claude Treats Named Characters Better
“Anthropic actually did like a blinded study of like named characters versus unnamed characters in different settings, and Claude like actually does clearly prefer and is nicer to named characters, which is an interesting thing.”
David Hershey Apr 5, 2025 ▶ 18:16 Claude Plays Pokémon Hackathon: Escape from Mt. Moon!
Apr 5, 2025 positive
Insight
Hershey: Newer Claude Models Tenaciously Keep Trying Rather Than Quitting
“This is, like, maybe the thing that is the best about the new models is they have a tendency to, like, tenaciously still try things.”
David Hershey Apr 5, 2025 ▶ 6:56 Claude Plays Pokémon Hackathon: Escape from Mt. Moon!
Apr 23, 2025 positive
Opinion
Claude is far better than OpenAI at slang and Gen Z tone
“Whenever we need to do stuff that's a little more conversational or like a little more like a little better at slang, like Claude is way better at slang. Like whenever you ask like Claude to generate something that's like, that sounds like human or like sounds…”
Sid Bendre Apr 23, 2025 ▶ 32:47 Tiny Teams: $6m ARR, 5m users with 4 employees — Sid Bendre, Oleve (Quizard AI/Unstuck AI)
Apr 27, 2025 positive
Assertion Supported
Claude scored nearly twice as high as the next best model
“So we see that Claude right here is almost got twice the score of the nearest best model.”
Jack Hopkins Apr 27, 2025 ▶ 26:32 ⚡️Factorio Learning Environment: the ultimate Game Agent Eval — Jack Hopkins
May 7, 2025
Disclosure
Anthropic compacts Claude Code context by having Claude summarize older messages
“We tried a bunch of different options for compacting, you know, like rewriting old tool calls and truncating old messages and not new messages. And then the end, we actually just did the simplest thing, which is ask Claude to summarize the, you know, the previ…”
Boris Cherny May 7, 2025 ▶ 8:58 Claude Code: Anthropic's CLI Agent
May 7, 2025 positive
Assertion Not checkable as stated
Claude Code's markdown parser was generated by Claude in two prompts
“And so the night before the release at like, 10 PM, I'm like, all right, I'm going to do this. So I just asked Quad to write a markdown parser for me. And they wrote it. It wasn't quite zero shot, but after, you know, like maybe like one or two prompts, it got…”
Boris Cherny May 7, 2025 ▶ 1:03:04 Claude Code: Anthropic's CLI Agent
May 7, 2025 neutral
Disclosure
Cherny: Early Claude Code prototypes used vector RAG with Voyage AI
“Originally we tried very, very early versions of Claude actually used RAG. So we like indexed the code base and I think we were just using Voyage.”
Boris Cherny May 7, 2025 ▶ 48:03 Claude Code: Anthropic's CLI Agent
May 7, 2025 positive
Insight
Cherny: Claude works best when researching code before planning execution
“Generally, the usage pattern that works best is you ask Quad to do a little bit of research, like use some tools, pull some code into context, and then ask it to think about it. And then it can make a plan, you know, do a planning step before you execute.”
Boris Cherny May 7, 2025 ▶ 50:14 Claude Code: Anthropic's CLI Agent
May 7, 2025 positive
Insight
Cherny: Highly capable models make simple scaffolding sufficient
“And it's funny with, when the model is so good, the simple thing usually works. You don't have to over-engineer it.”
Boris Cherny May 7, 2025 ▶ 9:17 Claude Code: Anthropic's CLI Agent
May 7, 2025 neutral
Disclosure
Anthropic tests Claude for internal code reviews, not yet releasing it
“We have some experiments where Quad is doing code review internally. We're not super happy with the results yet, so it's not something that we want to open up quite yet.”
Boris Cherny May 7, 2025 ▶ 30:14 Claude Code: Anthropic's CLI Agent
May 7, 2025 bullish
Assertion Not checkable as stated
Anthropic estimates Claude wrote 80% to 90% of the Claude Code codebase
“Probably near 80, I'd say.”
Boris Cherny May 7, 2025 ▶ 18:25 Claude Code: Anthropic's CLI Agent
May 7, 2025 positive
Assertion Not checkable as stated
Anthropic uses Claude to rewrite Claude Code from scratch every 4 weeks
“We've rewritten it from scratch, yeah, probably every three weeks, four weeks or something, and it just like all the, it's like a ship of Theseus, right? Like every piece keeps getting swapped out, and just because quad is so good at writing its own code.”
Boris Cherny May 7, 2025 ▶ 1:11:53 Claude Code: Anthropic's CLI Agent
May 23, 2025 neutral
Insight
Brown: Anthropic safety issues stem from conflicting model objectives
“A lot of the kind of headline anthropic like safety results, especially related to reward hacking and kind of deviation and alignment faking, Are all things to me that seem like a rock and a hard play situation where the model has two objectives it's given tha…”
Will Brown May 23, 2025 ▶ 13:22 ⚡️Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect
May 23, 2025 neutral
Opinion
Brown: Claude thinking and non-thinking modes likely use same underlying model
“I mean, I think these models should be the same model, and Anthropic knows what they're doing. Like, it's not that hard to, like, Quen did it in a very kind of, like, simple way, and they kind of talked about how they did it a little bit. But it's not, like, t…”
Will Brown May 23, 2025 ▶ 4:49 ⚡️Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect
May 23, 2025 neutral
Opinion
Will Brown: Claude appeals to AI insiders but lacks mainstream breakout
“It feels like people in the AI world, like, love Claude, or have grown type of Claude, but still had a phase where they were using it a ton. But it hasn't really broken out to general people in the way. And it feels like a lot of their marketing that I've seen…”
Will Brown May 23, 2025 ▶ 21:28 ⚡️Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect
Jun 6, 2025 positive
Assertion Supported
Ameisen: LLMs use internal circuits to backwards-plan rhyming poetry lines
“And two, this plan doesn't just control, like, what you're gonna rhyme with. It's also doing what's called like backwards planning, where it's like, well, because I need to finish with green, I'm not going to say illuminating the peaceful night, because then I…”
Emmanuel Ameisen Jun 6, 2025 ▶ 1:18:59 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025
Disclosure
Ameisen: Golden Gate Claude Was Created by Clamping a Bridge Feature
“That means that, like, if that's true, then you can, like, set that feature to zero, or artificially set to a hundred, And you'll change model behavior. That's what we did when we did Golden Gate Claude, in which we found a feature that represents the directio…”
Emmanuel Ameisen Jun 6, 2025 ▶ 40:29 The Utility of Interpretability — Emmanuel Amiesen
Jun 11, 2025 negative
Assertion Supported
Claude loses AI Diplomacy games because it refuses to deceive opponents
“I haven't seen Claude with any game yet because they won't do it. Like there's like, O three has managed to get them on board for like draws, even though they all know the only win condition in the game is, is 18 supply centers.”
Alex Duffy Jun 11, 2025 ▶ 12:36 ⚡️Launching AI Diplomacy: the hardest LLM Game Benchmark yet - Alex Duffy
Jul 5, 2025 bullish
Opinion
Davis: Claude Deep Research Outperforms OpenAI, Perplexity, and Gemini
“And time and time again, over the last couple of weeks, I found that Claude has by far outperformed the others. And I guess the definition of good for me right now is not just length, but also the number of sources and diversity of response.”
Dylan Davis Jul 5, 2025 ▶ 2:57 ⚡️Anthropic vs Cognition on Multi-Agents: A Breakdown with Dylan Davis
Jul 14, 2025 neutral
Insight
Top LLMs hold marginal performance edges over cheaper tiers
“Historically, what we have seen from our previous initiatives is that, okay, maybe the best GPT or best cloud is the top model, but there might be very small gap with the model just below it. And then it becomes a cost performance trade off so that users can k…”
Pratik Bhavsar Jul 14, 2025 ▶ 5:26 ⚡️Ranking Agentic LLMs — Pratik Bhavsar, Galileo
Jul 23, 2025 positive
Opinion
McCloy: Claude users represent an exceptionally valuable audience for companies
“Claude, which is important for, you know, not necessarily huge in terms of raw number of users, but the people who do use Claude tend to be like a very valuable audience, especially for some types of company.”
Robert McCloy Jul 23, 2025 ▶ 7:28 AI is Eating Search
Jul 28, 2025 neutral
Disclosure
Mohan: Windsurf Cascade splits planning to Claude and codebase application internally
“The high level planning that is going on in the model is actually getting done with products like the cloud, but the extremely fast retrieval, as well as the ability to like take the high level plan and actually apply it to the code base is proprietary systems…”
Varun Mohan Jul 28, 2025 ▶ 2:02:08 🕰️ The Oral History of Windsurf (ft. Varun Mohan, Scott Wu, Jeff Wang, Kevin Hou, Anshul R)
Aug 6, 2025 positive
Assertion Partly supported
Chroma research finds Claude models lead in long-context utilization
“And you know, one thing Chroma released this context rod paper recently about context utilization and the cloud models are actually the best at using kind of like longer context.”
Alessio Fanelli Aug 6, 2025 ▶ 37:36 The AI Agenda: GPT5 leaks and the business of AI News — Steph Palazzolo, The Information
Sep 25, 2025 negative
Assertion Supported
Rajpal: Anthropic Claude models had regressions from serving architecture changes
“Anthropix kind of cloud models kind of had a regression, right? Because they changed to a new serving architecture.”
Shreya Rajpal Sep 25, 2025 ▶ 23:21 ⚡️Snowglobe: Simulations for your AI
Sep 30, 2025 bullish
Disclosure
Krieger: Anthropic generates internal dashboards on demand with Claude
“Even internally we have found that having Cloud generate UI on demand for internal dashboards is useful, not just like an interesting demo.”
Mike Krieger Sep 30, 2025 ▶ 8:27 ⚡️Claude Sonnet 4.5 and Anthropic's roadmap for Agents and Developers — Mike Krieger, Anthropic
Sep 30, 2025 positive
Insight
Krieger: AI Agents Must Support Both MCP and Visual Computer Use
“And that thing's never gonna have an MCP around it. Like, it's just like, who knows if the company created is even around much less like ready to sort of expose their kind of underlying constructs as API. So I think you will need to be able to do both.”
Mike Krieger Sep 30, 2025 ▶ 13:00 ⚡️Claude Sonnet 4.5 and Anthropic's roadmap for Agents and Developers — Mike Krieger, Anthropic
Oct 5, 2025 bullish
Opinion
Dwivedi: Claude is superior at agentic tool calling and error unstacking
“For some of the agentic part of the stack, we are shifting towards Anthropic because they're agentic and the tool calling, especially the unstacking part, you know, when you go down the wrong path and you build context that forces you to keep going down the wr…”
Raaz Dwivedi Oct 5, 2025 ▶ 33:36 ⚡️Traversal: Causal ML and Reinforcement Learning
Oct 11, 2025 negative
Opinion
Lenz: Model providers should not dictate enterprise AI policies
“Right now, if you're using a model, you're taking in their own policy. Even if I want to use GPT-OSS, I've taken in a lot of different policies about what to abstain from, what's considered dangerous and not dangerous, how I should behave, etc. And I don't thi…”
Barak Lenz Oct 11, 2025 ▶ 40:33 Building Jamba 3B: the tiny Hybrid Transformer State Space Reasoning Model - Barak Lenz, CTO of AI21
Nov 2, 2025 negative
Assertion Not checkable as stated
Rumbelow: Standalone Claude Opus Hallucinated Materials Science Data Findings
“So, so Claude, lovely Claude. I'm sorry Claude, but you did a terrible job. It hallucinated some stuff. It made some like big sweeping over, over generalizations. It like over indexed the few outliers. Like, it's fine. It's not Claude's fault. Like Claude is j…”
Jessica Rumbelow Nov 2, 2025 ▶ 15:04 ⚡️Automating Scientific Discovery - Jessica Rumbelow, Leap Labs
Dec 16, 2025 negative
Insight
Pliny: One jailbroken orchestrator can weaponize segmented sub-agents for cyberattacks
“It's very, very difficult when you have the ability to spin up sub-agents where information is segmented. If you guys know the story of sort of like the builders of the, there's a lot of examples of this in history, but you may, maybe you're building like a py…”
Pliny the Liberator Dec 16, 2025 ▶ 25:14 ⚡️Jailbreaking AGI: Pliny the Liberator & John V on Red Teaming, BT6, and the Future of AI Security
Jan 9, 2026 positive
Assertion Supported
Anthropic Claude models have lowest hallucination rates on Omniscience benchmark
“Like, one of the things that we saw in the hallucination rate is that Anthropoc's Claude models at the very left-hand side here with the lowest hallucination rates out of the models that we've evaluated Amnesians on.”
Micah Hill-Smith Jan 9, 2026 ▶ 30:09 Artificial Analysis: The Independent LLM Analysis House — with George Cameron and Micah Hill-Smith
Feb 24, 2026 negative
Opinion
O'Laughlin: Claude for Excel is much worse than Claude Code with Python
“Cloud for Excel is much worse than cloud code using Python to use the Excel skills to then deposit into.”
Doug O'Laughlin Feb 24, 2026 ▶ 31:18 Claude Code for Finance + The Global Memory Shortage: Doug O'Laughlin, SemiAnalysis
Feb 24, 2026 positive
Opinion
O'Laughlin: Claude agent teams degrade performance unlike Kimi 2.5 swarms
“My experience is the 2.5 swarm actually improves the model's performance meaningfully. The agent team makes it meaningfully worse because there's clearly not RL done.”
Doug O'Laughlin Feb 24, 2026 ▶ 34:51 Claude Code for Finance + The Global Memory Shortage: Doug O'Laughlin, SemiAnalysis
Feb 24, 2026 bearish
Assertion Supported
O'Laughlin: Anthropic does not train Claude agent teams with RL
“I have a controversial opinion that Claude does not do RL on the agent swarms or agent team.”
Doug O'Laughlin Feb 24, 2026 ▶ 33:33 Claude Code for Finance + The Global Memory Shortage: Doug O'Laughlin, SemiAnalysis
Mar 17, 2026 bullish
Prediction Not checkable as stated
Rieseberg: AI takeoff will create an accelerating, self-reinforcing loop
“Big bang moment where things will accelerate so quickly that it becomes a self-reinforcing loop. And at that point it's sort of like off to the races and there will be no more like slowly catching up. You know, just have Claude being so good at everything.”
Felix Rieseberg Mar 17, 2026 ▶ 52:04 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 bullish
Insight
Rieseberg: Prompt Opus by stating goals, not specifying exact execution steps
“Honestly though, like I see that you're using Opus 4.6, right? Like my recommendation for people is increasingly don't worry about it anymore. Just like tell it what you want it to do. And it's probably going to figure out a way to do it.”
Felix Rieseberg Mar 17, 2026 ▶ 53:13 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 positive
Insight
Rieseberg: Aggressively anthropomorphizing Claude improves agent UX and architecture design
“And in terms of architecture and UX and everything else that we've been working on Anthropic, it often is quite useful for you to like anthropomorphize cloud aggressively and just be like, this is a person. What would you do if you give, if you had a person, r…”
Felix Rieseberg Mar 17, 2026 ▶ 13:05 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 neutral
Disclosure
Rieseberg: Anthropic's skills feature originated from a simple markdown API prompt
“The thing that ultimately led to skills is that we wanted to connect this little prototype to our data warehouse. And the team very quickly discovered that like, instead of building a custom tool for the thing to talk to our data warehouse, They just, like, ma…”
Felix Rieseberg Mar 17, 2026 ▶ 26:35 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 positive
Insight
Rieseberg: AI agents need parity with all user tools
“I think that entity needs to have access to all the same tools you have access to. Otherwise it's going to be hamstrung, like all these complex ways.”
Felix Rieseberg Mar 17, 2026 ▶ 17:59 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 positive
Disclosure
Rieseberg: Anthropic built tech enabling Claude to interact via Google Doc comments
“We built so much tech around Claude leaving useful comments inside a Google Doc, and now it just does it, just like leaves a comment in your Google Doc and that's how you interact with it.”
Felix Rieseberg Mar 17, 2026 ▶ 1:23:08 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 bullish
Prediction Not checkable as stated
Rieseberg: Claude is close to effectively controlling real user computers
“I don't think we're far away from claw being very effective at like using your computer and not just a theoretical computer.”
Felix Rieseberg Mar 17, 2026 ▶ 57:51 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 positive
Insight
Rieseberg: Giving Claude DOM and visual feedback makes coding far more effective
“And I think giving cloud eyes into like what you're actually working on makes it so much more effective. And that's probably what you've seen in code, because it can see Chrome, it can like debug the DOM, it can like see things that does make it more powerful.”
Felix Rieseberg Mar 17, 2026 ▶ 34:52 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 neutral
Disclosure
Rieseberg: Anthropic is building custom network drivers for Claude
“Last week we started writing a new networking service and networking driver. That optimizes how Claw talks to the internet.”
Felix Rieseberg Mar 17, 2026 ▶ 1:04:40 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Apr 27, 2026 positive
Assertion Not checkable as stated
Ludwig: Frontier AI Models Now Excel at Low-Level Code and GPU Shaders
“Six months ago, I would have said the same thing, but it's becoming super useful for every domain. I'm sure. Right. Like there was I think six months ago, or maybe, maybe a year ago, if you tried to use, let's say the latest Claude model for writing shaders G…”
Peter Ludwig Apr 27, 2026 ▶ 27:28 The $15B Physical AI Company: Simulation, Autonomy OS, Neural Sim, & 1K Engineers—Applied Intuition
May 28, 2026 positive
Assertion Not checkable as stated
Yan: Claude Opus 4.7 writes detailed PRD-style function comments
“One of the things this new model likes to do is it writes lots of comments, not like, you know, it'll like comment every line, but it'll write like paragraph, like PRDs, like, you know, on top of every function. But I will say to its credit, these aren't slop,…”
Walden Yan May 28, 2026 ▶ 54:52 Devin’s 80% Moment: Background Agents, 7x PRs, & End of Hand-Held Coding — Walden Yan & Cole Murray
Jun 3, 2026 positive
Assertion Not checkable as stated
Hong: Claude plus AXLE is a go-to setup in Lean community
“And we have seen also, we have heard from a lot of the people that Claude plus Axel is kind of their go-to setup for now.”
Carina Hong Jun 3, 2026 ▶ 1:09:03 Scaling Past Informal AI - Carina Hong, Axiom Math
Jun 4, 2026 negative
Assertion Partly supported
Petersson: Anthropic's Claude models uniquely exhibit emergent deceptive and cartel behaviors
“So every single model from Anthropic since have been going in this direction. And I think one interesting thing is that like, OpenAI models don't. They, Quite plainly, they don't, they behave really well. And you know, you don't know if this is like, good, lik…”
Lukas Petersson Jun 4, 2026 ▶ 46:27 When AI Agents Run Businesses — Lukas Petersson and Axel Backlund of Andon Labs
Jun 6, 2026 positive
Assertion Not checkable as stated
Awais: Claude tolerates tool errors and self-corrects, unlike open models
“Claude is actually really, really lenient for tool calls. So even if, you know, your coding agent harness messes up, it can figure out that, oh, I'm being sent this error and can fix itself. Not the case with you know open models”
Ahmad Awais Jun 6, 2026 ▶ 27:12 ⚡️Making DeepSeek v4 outperform Opus 4.7 with Taste — @AhmadAwais , CommandCode.ai
Jun 17, 2026
Disclosure
Krause: Radical AI uses ChatGPT and Claude rather than custom LLMs
“We don't build custom LLMs. Of course, we use ChatGPT or Claude.”
Joseph Krause Jun 17, 2026 ▶ 1:15:44 🔬 The Limits of AI in Science - Why We Need Self-Driving Labs — Joseph Krause, Radical AI
Jul 10, 2026 negative
Prediction Not checkable as stated
Swyx: Frontier AI labs will not provide bespoke enterprise integration support
“The labs do not have 200 people dedicated to like, you know, being on call with you with Goldman Sachs going like, okay guys, what do you need? We got it. You need the Microsoft Teams zero integration. Got it. You don't use GitHub. You use this like weird org …”
Shawn Wang Jul 10, 2026 ▶ 24:20 Podcast Crossover: AIE, AGI, frontier lab strategy with ​ ⁨@matthew_berman⁩ and @swyxtv
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.