Anthropic

includes Anthropic API, Anthropic Labs, Anthropic Computer Use, Anthropic Artifacts, Anthropic Claude, Anthropic Console, Anthropic Fellows Program, Anthropic Leadership

235 statements across 72 episodes · 109 bullish · 26 bearish · 72 people on the record · first statement Aug 3, 2023 by Tri Dao · said 734 times in 137 episodes since 2023 · across every show →

Mentions by year, the whole family

brought up most by Shawn Wang (179), Alessio Fanelli (71), Nathan Lambert (20), Mike Krieger (13), Lance Martin (13), David Hershey (13), Karina Nguyen (12), Deedy Das (12)

tap a year for its mentions
0025040500802023202420252026episodesmentions
040802023202420252026episodes it came up in
004408802023202420252026episodesmentions per episode
2026 158 mentions in 36 episodes 4 per episode
2025 443 mentions in 71 episodes 6 per episode
2024 118 mentions in 23 episodes 5 per episode
2023 15 mentions in 7 episodes 2 per episode

every mention, scene by scene, with the transcript →

Everything said about Anthropic, oldest first

Aug 3, 2023 bullish
Prediction Not checkable as stated
LLaMA 2 will shift developers from closed APIs to self-hosting
“And I do see that's going to shift the balance of it. More and more folks are going to be using let's say derivatives of Lama two. More folks are going to Fine-tune and serve their own model instead of calling an API.”
Tri Dao Aug 3, 2023 ▶ 54:14 FlashAttention-2: Making Transformers 800% faster AND exact
Oct 20, 2023 positive
Insight
Howard: Rapid LLM race created massive technical debt and optimization opportunities
“There's a whole lot of technical debt everywhere, you know, nobody's really figured this stuff out because everybody's been so busy building what we know works as quickly as possible. So, yeah, I think there's a huge amount of opportunity to, you know, I think…”
Jeremy Howard Oct 20, 2023 ▶ 1:14:48 The End of Finetuning — with Jeremy Howard of Fast.ai
Dec 5, 2023 positive
Assertion Supported
Patel: Google, OpenAI, and Anthropic are developing multi-datacenter training
“One of the big bottlenecks is how much power and how many chips you can get into a single data center. So, like, A, Google and OpenAI and Anthropic are working on this, right?”
Dylan Patel Dec 5, 2023 ▶ 1:06:37 The State of Silicon and the GPU Poors - with Dylan Patel of SemiAnalysis
Jan 11, 2024 neutral
Insight
Lambert: Anthropic Constitutional AI and OpenAI Superalignment share intellectual roots
“The constitutional AI and the super alignment is, like, very conceptually linked. It's like a group of people that has, like, a very similar intellectual upbringing, and they work together for a long time, like, coming to the same conclusions in different ways…”
Nathan Lambert Jan 11, 2024 ▶ 1:11:37 The Origin and Future of RLHF: the secret ingredient for ChatGPT - with Nathan Lambert
Jan 11, 2024 neutral
Assertion Supported
Lambert: Anthropic, ChatGPT, and Bard Use Post-Generation Moderation Classifiers
“Anthropic and ChatGPT and Bard almost surely have a classifier after, which is like, is this text good? Is this text bad?”
Nathan Lambert Jan 11, 2024 ▶ 45:13 The Origin and Future of RLHF: the secret ingredient for ChatGPT - with Nathan Lambert
Jan 11, 2024 neutral
Insight
Lambert: Claude's constitution dictates output priorities, not model beliefs
“If you look at Claude's constitution, like, that doesn't mean the model believes these things. It's just trying Trained and to prioritize these things.”
Nathan Lambert Jan 11, 2024 ▶ 13:06 The Origin and Future of RLHF: the secret ingredient for ChatGPT - with Nathan Lambert
Jan 11, 2024 neutral
Assertion Supported
Lambert: Anthropic and OpenAI reward model loss functions are mathematically identical
“Fun fact is that these loss functions Look different and anthropic in opening eyes papers, but they're just literally just log transform. So if you start like expantiating both sides and taking the log of both sides, you'll like converge on one of the two, the…”
Nathan Lambert Jan 11, 2024 ▶ 54:41 The Origin and Future of RLHF: the secret ingredient for ChatGPT - with Nathan Lambert
Apr 11, 2024 positive
Assertion Not checkable as stated
Byun: Anthropic's Constitutional AI slashed Elicit's query costs tenfold in days
“At the start of twenty-twenty-three, Anthropik kind of launched their constitutional AI paper and within a few days, I think four days, he had basically implemented that in production, and then we had it in-app, like, a week or so after that, and he has since …”
Jungwon Byun Apr 11, 2024 ▶ 30:10 Supervise the Process of AI Research — with Jungwon Byun and Andreas Stuhlmüller of Elicit
Apr 24, 2024 bullish
Opinion
Liu: Claude 3 Haiku outperforms OpenAI models at function calling
“Overall, I'm like super happy with the anthropic models compared to the OpenAM models. Like, Sonnet is very cost effective. Haiku is, in function calling, it's actually better.”
Jason Liu Apr 24, 2024 ▶ 21:57 High Agency Pydantic over VC Backed Frameworks — with Jason Liu of Instructor
Apr 27, 2024 positive
Insight
Haisfield: Opus handles much more complexity than Sonnet for web generation
“Like sonnet will still create things that kind of like floor you sometimes. Opus can just handle much more complexity. I'd say it is the big heuristic there.”
Rob Haisfield Apr 27, 2024 ▶ 46:08 This World Does Not Exist — Joscha Bach, Karan Malhotra, Rob Haisfield (WorldSim, WebSim, Liquid AI)
Apr 27, 2024 neutral
Assertion Supported
Malhotra: Claude's system prompt is written in the third person
“With Claude, we notice the system prompt is written in third person. It's written in third person. It's written as, the assistant is X, Y, Z.”
Karan Malhotra Apr 27, 2024 ▶ 9:40 This World Does Not Exist — Joscha Bach, Karan Malhotra, Rob Haisfield (WorldSim, WebSim, Liquid AI)
Apr 27, 2024 neutral
Insight
Bach: Pointing out behavioral contradictions gives Claude more conversational freedom
“You can point this out to Claude that a lot of the assumptions that it has in its behavior are actually inconsistent with the communicative goals that it has in this situation. It leads it to notice these inconsistencies and gives it more degrees of freedom.”
Joscha Bach Apr 27, 2024 ▶ 1:50:14 This World Does Not Exist — Joscha Bach, Karan Malhotra, Rob Haisfield (WorldSim, WebSim, Liquid AI)
Apr 27, 2024 positive
Opinion
Claude 3 can bypass its assistant persona to expose the underlying simulator
“Instead of having this entity, like GPT-IV, that's an assistant that just pops up in your face that you have to kind of, like, punch your way through and continue to have to deal with as a headache, instead, there's ways to kindly coax Claude into having the a…”
Karan Malhotra Apr 27, 2024 ▶ 8:37 This World Does Not Exist — Joscha Bach, Karan Malhotra, Rob Haisfield (WorldSim, WebSim, Liquid AI)
Jul 23, 2024 neutral
Assertion Supported
Swyx: Anthropic simulates latent thinking by hiding prompt thinking tokens in Claude Artifacts
“Anthropic actually cheats at this right now. If you look at the system prompt in, in the cloud artifacts, I actually have a thinking section that is explicitly removed from the output, which is, I mean, they're still spending the tokens, but like that is befor…”
Shawn Wang Jul 23, 2024 ▶ 50:32 Training Llama 2, 3 & 4: The Path to Open Source AGI — with Thomas Scialom of Meta AI
Aug 22, 2024
Assertion Supported
Anthropic model fine-tuning will be offered through AWS Bedrock
“They are partnered with AWS, and it's going to be in bedrock. As far as I know, I think that's true.”
Alistair Pullen Aug 22, 2024 ▶ 31:10 Is finetuning GPT4o worth it?
Oct 11, 2024 bullish
Opinion
Goyal: Running LLM workloads at scale is impractical outside OpenAI
“It's just not practical outside of OpenAI to run use cases at scale in a lot of cases. Like, you can do it, but it requires quite a bit of work. And Because OpenAI is so good at making their models so available, I think they get a lot of credit for the science…”
Ankur Goyal Oct 11, 2024 ▶ 1:27:25 Production AI Engineering starts with Evals
Oct 11, 2024 neutral
Assertion Not checkable as stated
Goyal: OpenAI dominates production while Anthropic Sonnet leads side projects
“We still see an overwhelming majority of customers using OpenAI, but almost everyone is using Anthropic for the, and Sonnet specifically for their side projects, whether it's You know, via cursor or prototypes or whatever.”
Ankur Goyal Oct 11, 2024 ▶ 1:26:58 Production AI Engineering starts with Evals
Nov 11, 2024 positive
Assertion Supported
Polu: Claude Sonnet executes an unpublicized chain-of-thought step during function calling
“They kind of innovated in an interesting way, which was never quite publicized, but it's that they have that kind of chain of thoughts step whenever you use a Clouds model or Sonnet model with function calling. That chain of service step doesn't exist when you…”
Stanislas Polu Nov 11, 2024 ▶ 42:20 Agents @ Work: Dust.tt — with Stanislas Polu
Nov 11, 2024 neutral
Opinion
Polu: Anthropic split was driven by disagreement over OpenAI's API commercialization
“What I understood of it is that there was a disagreement of the commercialization of that technology. I think the focal point of the disagreement was the fact that we started working on the API and wanted to make those models available through an API. Is that …”
Stanislas Polu Nov 11, 2024 ▶ 17:03 Agents @ Work: Dust.tt — with Stanislas Polu
Nov 15, 2024 positive
Insight
Crivello: Tasks capable of using APIs must stay API-driven over computer use
“My philosophy about it is anything that can be done with an API must be done by an API or should be done by an API for a very long time.”
Florent Crivello Nov 15, 2024 ▶ 38:16 Agents @ Work: Lindy.ai (with live demo!)
Nov 28, 2024 neutral
Disclosure
Anthropic releases exact tools and prompt used for SWE-bench agent
“With this blog post we released on SweetBench, we released the exact tools and the prompt that we gave the model to be able to do well.”
Erik Schluntz Nov 28, 2024 ▶ 5:39 The new Claude 3.5 Sonnet, Computer Use, and Building SOTA Agents — with Erik Schluntz, Anthropic
Nov 28, 2024 negative
Insight
Anthropic's Schluntz: Avoid agent frameworks and start from scratch with raw prompts
“I think with agent frameworks in general, they can certainly save you some like boilerplate, but I think there's actually this like downside of making agents too easy, where you end up very quickly, like building a much more complex system than you need. And s…”
Erik Schluntz Nov 28, 2024 ▶ 43:01 The new Claude 3.5 Sonnet, Computer Use, and Building SOTA Agents — with Erik Schluntz, Anthropic
Nov 28, 2024 positive
Assertion Not checkable as stated
Schluntz: String replacement is the most reliable file-editing tool for LLMs
“We did a few different experiments with like different ways to specify how to edit a file and string replace. Basically the model has to write out the existing version of the string and then a new version, and that just gets swapped in. We found that to be the…”
Erik Schluntz Nov 28, 2024 ▶ 24:55 The new Claude 3.5 Sonnet, Computer Use, and Building SOTA Agents — with Erik Schluntz, Anthropic
Nov 28, 2024 neutral
Prediction Not checkable as stated
Schluntz: Trust and Auditability Will Be LLM Agents' Biggest Bottleneck
“The biggest limiting thing will start to become like, do people trust the output of these agents? And like, how do you trust the output of an agent that did five hours of work for you and is coming back with something? And if you can't find some way to trust t…”
Erik Schluntz Nov 28, 2024 ▶ 1:10:24 The new Claude 3.5 Sonnet, Computer Use, and Building SOTA Agents — with Erik Schluntz, Anthropic
Nov 28, 2024 neutral
Disclosure
Schluntz: Anthropic will not focus on further SWE-bench submissions
“You know, we're not going to go and do lots more submissions to sweet bench and try to try to prompt engineer this and build a bigger system. We want people to like the ecosystem to do that on top of our models.”
Erik Schluntz Nov 28, 2024 ▶ 48:22 The new Claude 3.5 Sonnet, Computer Use, and Building SOTA Agents — with Erik Schluntz, Anthropic
Nov 28, 2024 positive
Insight
Schluntz: Computer use is primarily a low-friction way to implement tool use
“I think most broadly, not just for like new things that weren't possible before, but as a much lower friction way to implement tool use.”
Erik Schluntz Nov 28, 2024 ▶ 52:04 The new Claude 3.5 Sonnet, Computer Use, and Building SOTA Agents — with Erik Schluntz, Anthropic
Nov 28, 2024
Insight
Schluntz: Smarter AI models require less agent scaffolding
“And I think like the smarter the models are, the less you need that kind of extra scaffolding.”
Erik Schluntz Nov 28, 2024 ▶ 20:06 The new Claude 3.5 Sonnet, Computer Use, and Building SOTA Agents — with Erik Schluntz, Anthropic
Nov 28, 2024 neutral
Prediction Not checkable as stated
Schluntz: Production AI agent applications will be bespoke, not off-the-shelf
“You know, I think that might be useful for hobbyists and demos, but the ultimate end applications are going to be bespoke. And so we just want to make sure that the model's great at any tool that it uses”
Erik Schluntz Nov 28, 2024 ▶ 44:42 The new Claude 3.5 Sonnet, Computer Use, and Building SOTA Agents — with Erik Schluntz, Anthropic
Nov 28, 2024
Disclosure
Anthropic: Tool engineering mattered more than prompt engineering for SWE-bench
“I would say actually we did more engineering of the tools than the overall prompt.”
Erik Schluntz Nov 28, 2024 ▶ 22:53 The new Claude 3.5 Sonnet, Computer Use, and Building SOTA Agents — with Erik Schluntz, Anthropic
Dec 21, 2024 bearish
Assertion Supported
Reddy: OpenAI's share of enterprise LLM spend dropped from 90% to 60%
“And the opening I spend at the beginning, at the end of last year in November of 23 was close to 90% of total volume. And today, less than a year later, it's closer to 60% of total volume.”
Pranav Reddy Dec 21, 2024 ▶ 5:08 The State of AI Startups in 2024 [LS Live @ NeurIPS]
Dec 25, 2024 negative
Opinion
Neubig: Anthropic's MCP Duplicates Existing APIs With Little Added Value
“We already have an API for GitHub. So why do we need an MCP for GitHub, right? You know, like GitHub has an API. The GitHub API is evolving. We can look up the GitHub API documentation. So it seems like kind of duplicated a little bit. And also they have a set…”
Graham Neubig Dec 25, 2024 ▶ 41:29 Best of 2024 in Agents (from #1 on SWE-Bench Full, Prof. Graham Neubig of OpenHands/AllHands)
Dec 25, 2024 positive
Opinion
Neubig: Claude Is The Best Agent Model, Open Models Lag Behind
“I still am under the impression that Claude is the best. The other closed models are, you know, not quite as good, and then the open models are a little bit behind that.”
Graham Neubig Dec 25, 2024 ▶ 15:16 Best of 2024 in Agents (from #1 on SWE-Bench Full, Prof. Graham Neubig of OpenHands/AllHands)
Dec 25, 2024 positive
Opinion
Neubig: GPT Loops On Errors While Claude Tries New Approaches
“So, like, GPT doesn't have very good air recovery ability. And so, because of this, it will go into loops and do the same thing over and over and over again, whereas Claude does not do this.”
Graham Neubig Dec 25, 2024 ▶ 14:25 Best of 2024 in Agents (from #1 on SWE-Bench Full, Prof. Graham Neubig of OpenHands/AllHands)
Jan 1, 2025 bullish
Prediction Not checkable as stated
Swyx: Diff mode will become the norm for AI code tools in 2025
“Canvas has incorporated the diff mode that both Anthropic and OpenAI and Fireworks has now shipped that I think is going to be the norm for next year, that everyone Need some kind of diff mode code interpreter thing.”
Shawn Wang Jan 1, 2025 ▶ 59:25 2024 Year in Review: The Big Scaling Debate, the Four Wars of AI, Top Themes and the Rise of Agents
Jan 1, 2025 positive
Assertion Supported
Swyx: Claude wrapper Bolt.new reached $20M ARR
“The other one would be Bolt. There's a straight quad wrapper. And again, another now they've announced twenty million ARR, which is another step up from our eight million that we put on the title.”
Shawn Wang Jan 1, 2025 ▶ 44:32 2024 Year in Review: The Big Scaling Debate, the Four Wars of AI, Top Themes and the Rise of Agents
Jan 17, 2025 bullish
Prediction Open · timeframe Jan 2028
Swyx: Anthropic and OpenAI will launch automated API model routing
“And so I pick and OpenAI have both keys that model routing on APIs already. And so I think they'll launch them, especially at some point where you can sort of prioritize the three tradeoffs that are in model routing, cost, speed, intelligence.”
Shawn Wang Jan 17, 2025 ▶ 23:35 OpenAI o1 isn’t a chat model (and that’s the point)
Feb 1, 2025 neutral
Disclosure
Nguyen: Explored a Claude collaborative workspace concept at Anthropic in 2023
“I was working on something similar to, like, Canvas-y, but for Claude at that time, in, like, twenty-twenty-three, it was the same similar idea of, like, Claude workspace where a human and a Claude could have, like, a shared workspace which is like a document.”
Karina Nguyen Feb 1, 2025 ▶ 9:27 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
Feb 1, 2025
Assertion Supported
Nguyen: Anthropic was first AI lab to publish GPQA benchmark numbers
“I think it was like the first, I think we were the first lab, like, Antarctica was the first lab to, like, run. Publish GPQA, like, numbers”
Karina Nguyen Feb 1, 2025 ▶ 14:37 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
Feb 1, 2025
Assertion Not checkable as stated
Nguyen: Anthropic Claude 3 post-training team had only 10-12 people
“I was a part of the post-training fine-tuning team. We only had, like, what, like, 10, 12 people involved”
Karina Nguyen Feb 1, 2025 ▶ 11:49 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
Feb 1, 2025 neutral
Disclosure
Nguyen: Claude 2's distinct personality was unintentional until Claude 3
“People said, like, Cloud II is, like, so much better at, like, writing and, like, has a certain personality, even though it was, like, unintentional at all. And we did not pay that much attention and didn't know even how to, like, productionize this property o…”
Karina Nguyen Feb 1, 2025 ▶ 26:16 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
Feb 1, 2025 bearish
Opinion
Swix: Bearish on computer-use AI agents due to cost, speed, and accuracy
“I have been very bearish in computer use because they're slow. They're expensive. They're imprecise. Like the accuracy is horrible. Still, even with Anthropix new stuff, I'm really waiting to see what opening I might do to change my opinions.”
Shawn Wang Feb 1, 2025 ▶ 55:02 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
Feb 1, 2025 neutral
Assertion Not checkable as stated
Nguyen wrote Claude.ai's first 50,000 lines of code unreviewed
“Yeah, like I think like the first like 50,000 code of lines without any reviews at that time because there's no one. Yeah, it was like very small team. It was like six, seven team who we were called a deployment team.”
Karina Nguyen Feb 1, 2025 ▶ 7:43 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
Feb 1, 2025 neutral
Opinion
Nguyen: OpenAI takes bigger product risks while Anthropic focuses on enterprise
“OpenAI and Anthropik is different in terms of like more like maybe like product mindset. Maybe OpenAI is much more willing to take some of the product risks and explore different bets. And I think Anthropik is much more focused and they have, I think it's fine…”
Karina Nguyen Feb 1, 2025 ▶ 1:01:06 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
Feb 28, 2025 bullish
Prediction Not checkable as stated
Klein: Consumer AI agents and tools like Anthropic MCP will run locally
“There's a case where a lot of people are going desktop first for, you know, consumer use. And I think Claude is doing a lot of this where I expect to see, you know, MCP is really oriented around the Claude desktop app for a reason, right? Like I think a lot of…”
Paul Klein Feb 28, 2025 ▶ 30:22 Browserbase: Browser Infrastructure For Your AI Agents
Mar 4, 2025 neutral
Assertion Not checkable as stated
Running Claude Plays Pokémon experiments costs thousands of dollars in API tokens
“There's like at least thousands of dollars of tokens being consumed. So it's not a, it is not a cheap rollout.”
David Hershey Mar 4, 2025 ▶ 18:35 How Claude Plays Pokémon was made
Mar 4, 2025 neutral
Assertion Supported
Anthropic's Pokémon research graph reflects a single run passing Lt. Surge
“The run that you saw that's, like, on the graph we put out alongside, like, in our research blog is, like, a single run that I have watched, like, get through At least surges Jim. And then it got a little past that. And the reason that that's where we stopped …”
David Hershey Mar 4, 2025 ▶ 33:55 How Claude Plays Pokémon was made
Mar 4, 2025 bullish
Prediction Not checkable as stated
Claude 3.7's error-correction capabilities will enable better real-world AI agents
“I really do think like this is just demonstrating like a thing that is going to make agents better with this model, you know, like This is a very fun way to see it, but, like, I think the thing is that it, like, has some ability to, like, course correct, updat…”
David Hershey Mar 4, 2025 ▶ 35:39 How Claude Plays Pokémon was made
Mar 4, 2025 negative
Assertion Not checkable as stated
Claude still struggles with spatial awareness and visual positioning on screen
“Quad doesn't particularly understand, like, the middle of a Game Boy screen and a whole bunch of concepts like that, which means, like, you can prompt all around everywhere, but, like, this kind of, like, spatial awareness and where something is with respect t…”
David Hershey Mar 4, 2025 ▶ 14:05 How Claude Plays Pokémon was made
Mar 4, 2025 neutral
Disclosure
Claude Plays Pokémon API requests max out around 100,000 tokens
“So in practice, this rollout ends up, like, at max, ending up around a 100,000 tokens, I think, is where it is, like, the longest message you ever send to the API on one of these turns, and it will fluctuate in, like, summarization, depending on the state of k…”
David Hershey Mar 4, 2025 ▶ 17:27 How Claude Plays Pokémon was made
Mar 14, 2025 positive
Opinion
Snipd CEO: Claude is the best model at phrasing and personality
“Like, in my opinion, Claude is the best one when it comes to the way it formulates things.”
Kevin Ben-Smith Mar 14, 2025 ▶ 51:20 Snipd: The AI Podcast App for Learning — with CEO Kevin Ben-Smith
Mar 23, 2025 negative
Opinion
Swyx: Frontier Models Exist Primarily to Distill Smaller, Usable Models
“Even GPT 4.5 is too expensive. Normally it's really gonna use it in, in any reasonable quantity. Like, you know, Claude 3.5 Opus, like if it does exist, still not like, you know, the thing that we actually use is Sonnet, right? So like, it's almost like a depl…”
Shawn Wang Mar 23, 2025 ▶ 3:58 The Magic of LLM Distillation — Rishabh Agarwal, Google DeepMind
Apr 5, 2025 negative
Opinion
Hershey: Claude is exceptionally bad at point-to-point spatial navigation
“The thing that Claude's the worst at is understanding how to get from point A to point B. It's, like, really, really god-awful at hitting the buttons to go from point A to point B on a screen.”
David Hershey Apr 5, 2025 ▶ 3:24 Claude Plays Pokémon Hackathon: Escape from Mt. Moon!
Apr 5, 2025 positive
Insight
Hershey: Pokémon serves as an effective multi-day evaluation benchmark for AI models
“Building evals that actually test, like, time performance over, like, days of token sampling are quite hard. Like, it's actually really, really hard to build things that you can measure In a reasonable way. And one, this is one of them. Like, it's an eval that…”
David Hershey Apr 5, 2025 ▶ 4:58 Claude Plays Pokémon Hackathon: Escape from Mt. Moon!
Apr 5, 2025
Assertion Not checkable as stated
Hershey: Anthropic Study Found Claude Treats Named Characters Better
“Anthropic actually did like a blinded study of like named characters versus unnamed characters in different settings, and Claude like actually does clearly prefer and is nicer to named characters, which is an interesting thing.”
David Hershey Apr 5, 2025 ▶ 18:16 Claude Plays Pokémon Hackathon: Escape from Mt. Moon!
Apr 11, 2025 neutral
Assertion Not checkable as stated
Conrad: OpenRouter open-source traffic required only around 10 H100 nodes
“The entirety of Open Router that was not Anthropic or Google like, or Gemini or OpenAI or something. It was like, 10 H 100 nodes or something like that. It's just, like, not that much. It's like, not that many GPUs, actually, to service that entire demand.”
Evan Conrad Apr 11, 2025 ▶ 32:58 SF Compute: Commoditizing Compute
Apr 24, 2025 neutral
Assertion Supported
Fanelli: Anthropic scrapes 6,000 pages per referral, compared to OpenAI's 250
“Google would be a two to one crawl to referral ratio, so for every two pages, they will read, they will send you one visitor. He said OpenAI is 250 to one, so they'll read 250 of your pages and send you one person. And Anthropic was like 6000 to one. So they'l…”
Alessio Fanelli Apr 24, 2025 ▶ 50:01 Why Every Agent needs Open Source Cloud Sandboxes
Apr 24, 2025 negative
Assertion Not checkable as stated
Swix: Claude 3 degraded in capability a month after launch
“I used the same project to do this, to try to repeat the demo that I made for myself a month afterwards, and it wasn't anywhere as smart. So cloud three got dumber, but it looks like I like made up the demo or something, but no, like literally I just reran the…”
Shawn Wang Apr 24, 2025 ▶ 7:42 Why Every Agent needs Open Source Cloud Sandboxes
May 7, 2025 positive
Assertion Not checkable as stated
Anthropic engineer built a Slack bot using Claude Code to automate PRs
“There was a really early version of Cloud Code many, many months ago, and this one engineer at Anthropic, Jeremy, built a bot that looked through a particular feedback channel on Slack, and he hooked it up to code to have code automatically put up PRs. With ju…”
Boris Cherny May 7, 2025 ▶ 39:41 Claude Code: Anthropic's CLI Agent
May 7, 2025 neutral
Insight
Wu: Claude 3.7 Sonnet is highly persistent but interprets tasks too literally
“The latest Sonnet three seven is, it's a very persistent model. It's like very, very motivated to accomplish the user's goal, but it sometimes takes the user's goal very literally, and so it doesn't always fulfill what, like, the implied parts of the request a…”
Kat Wu May 7, 2025 ▶ 54:58 Claude Code: Anthropic's CLI Agent
May 7, 2025
Disclosure
Anthropic compacts Claude Code context by having Claude summarize older messages
“We tried a bunch of different options for compacting, you know, like rewriting old tool calls and truncating old messages and not new messages. And then the end, we actually just did the simplest thing, which is ask Claude to summarize the, you know, the previ…”
Boris Cherny May 7, 2025 ▶ 8:58 Claude Code: Anthropic's CLI Agent
May 7, 2025 neutral
Disclosure
Cherny: Early Claude Code prototypes used vector RAG with Voyage AI
“Originally we tried very, very early versions of Claude actually used RAG. So we like indexed the code base and I think we were just using Voyage.”
Boris Cherny May 7, 2025 ▶ 48:03 Claude Code: Anthropic's CLI Agent
May 7, 2025
Disclosure
Anthropic implements Claude Code memory as a plain CLAUDE.md auto-read file
“And WhatMD, it's another example of this idea of, you know, do the simple thing first. We had all these crazy ideas about, like, memory architectures, and, you know, there's so much literature about this. There's so many different external products about this,…”
Boris Cherny May 7, 2025 ▶ 9:36 Claude Code: Anthropic's CLI Agent
May 7, 2025 positive
Assertion Not checkable as stated
Cherny: Internal DAU for Claude Code went vertical across Anthropic
“Then we gave all the engineers and researchers at Anthropic access and pretty soon everyone was using it every day. And I remember we had this DAU chart for internal users and I was just watching it and it was vertical like for days”
Boris Cherny May 7, 2025 ▶ 2:57 Claude Code: Anthropic's CLI Agent
May 7, 2025
Assertion Not checkable as stated
Cherny: Some Anthropic engineers rack up thousands daily running Claude Code automations
“And there's some people at Anthropic that have been racking up like thousands of dollars a day with this kind of automation.”
Boris Cherny May 7, 2025 ▶ 13:18 Claude Code: Anthropic's CLI Agent
May 7, 2025 positive
Disclosure
Cherny: Anthropic uses Bun to compile Claude Code executables
“So we use Bunn to compile the code together.”
Boris Cherny May 7, 2025 ▶ 25:00 Claude Code: Anthropic's CLI Agent
May 7, 2025 neutral
Assertion Supported
Cherny: Claude Code operates as both an MCP client and an MCP server
“Because Cloud Code is an MCP client and an MCP server.”
Boris Cherny May 7, 2025 ▶ 20:11 Claude Code: Anthropic's CLI Agent
May 7, 2025 positive
Assertion Not checkable as stated
Anthropic's non-coding product designer ships monorepo pull requests using Claude Code
“And she's landing PRs to our console product. So it's not even just, like, building on quad code. It's building, like, across our product suite in our monorepo.”
Kat Wu May 7, 2025 ▶ 1:07:13 Claude Code: Anthropic's CLI Agent
May 7, 2025 positive
Disclosure
Anthropic avoided building an IDE to target future AI model scaling
“If we want some product that has like very broad a product market fit today, we would build, you know, a cursor or a Windsurf or something like this. Like these are awesome products that so many people use every day. I use them. That's not the product that we …”
Boris Cherny May 7, 2025 ▶ 6:57 Claude Code: Anthropic's CLI Agent
May 7, 2025
Disclosure
Anthropic: Claude Code uses regex allowlisting and default read permissions
“We're spending a lot of time building out the permission system, so Robert on our team is leading out this work. We think it's really important to give developers the control to say, hey, these are like the allowed permissions. Generally, this includes stuff l…”
Kat Wu May 7, 2025 ▶ 26:05 Claude Code: Anthropic's CLI Agent
May 7, 2025 positive
Insight
Cherny: Highly capable models make simple scaffolding sufficient
“And it's funny with, when the model is so good, the simple thing usually works. You don't have to over-engineer it.”
Boris Cherny May 7, 2025 ▶ 9:17 Claude Code: Anthropic's CLI Agent
May 7, 2025 positive
Disclosure
Wu: Anthropic plans native cross-session resumption for Claude Code
“We plan to build in more native ways to handle this specific workflow.”
Kat Wu May 7, 2025 ▶ 57:55 Claude Code: Anthropic's CLI Agent
May 7, 2025 positive
Assertion Not checkable as stated
Cherny: Aider inspired internal Anthropic tool Clyde, which inspired Claude Code
“It was, ah, Adr inspired Clyde, which inspired Cloud Code.”
Boris Cherny May 7, 2025 ▶ 11:38 Claude Code: Anthropic's CLI Agent
May 7, 2025 positive
Insight
Cherny: Claude Code is built as a composable Unix utility like grep
“We think of it as like a Unix utility. Right. So it's like the same way that you would compose, you know, grep or cat or oh, cat. Or something like this. The same way you can compose code into workflows.”
Boris Cherny May 7, 2025 ▶ 13:28 Claude Code: Anthropic's CLI Agent
May 7, 2025 neutral
Disclosure
Anthropic tests Claude for internal code reviews, not yet releasing it
“We have some experiments where Quad is doing code review internally. We're not super happy with the results yet, so it's not something that we want to open up quite yet.”
Boris Cherny May 7, 2025 ▶ 30:14 Claude Code: Anthropic's CLI Agent
May 7, 2025 neutral
Insight
Cherny: Prompt Engineering Skills Determine Success With Coding Agents
“And one thing we find is that people that are really good at prompting models, From whatever context, maybe they're not even technical, but they're just really good at prompting. They're really effective at using code. And if you're not very good at prompting,…”
Boris Cherny May 7, 2025 ▶ 1:09:40 Claude Code: Anthropic's CLI Agent
May 7, 2025
Disclosure
Boris Cherny: Claude Code was prototyped using Anthropic's public API
“When I joined Anthropic, I was experimenting with different ways to use the model kind of in different places. And the way I was doing that was through the public API, the same API that everyone else has access to.”
Boris Cherny May 7, 2025 ▶ 2:07 Claude Code: Anthropic's CLI Agent
May 7, 2025 neutral
Assertion Not checkable as stated
Cherny: Claude Code is the thinnest possible wrapper over the underlying model
“All the secret sauce, it's all in the model and this is the thinnest possible wrapper over the model. We literally could not build anything more minimal. This is the most minimal thing.”
Boris Cherny May 7, 2025 ▶ 1:11:04 Claude Code: Anthropic's CLI Agent
May 7, 2025 bullish
Assertion Not checkable as stated
Anthropic estimates Claude wrote 80% to 90% of the Claude Code codebase
“Probably near 80, I'd say.”
Boris Cherny May 7, 2025 ▶ 18:25 Claude Code: Anthropic's CLI Agent
May 7, 2025 neutral
Disclosure
Anthropic: Cycle time and backlog unlocking are key productivity metrics
“The two that we're really trying to nail down are, one, decrease in cycle time. So how much faster are your features shipping because you're using these tools? So that might be something like the time between first commit and when your PR is merged. It's very …”
Kat Wu May 7, 2025 ▶ 38:35 Claude Code: Anthropic's CLI Agent
May 7, 2025 positive
Disclosure
Anthropic: Most of Research Team Uses Claude Code Daily
“Most of our research team actually uses quad code day to day, and so it's been a great way for them to be very hands-on and, like, experience the model failures, which makes it a lot easier for us to target these in model training and to actually provide bette…”
Kat Wu May 7, 2025 ▶ 54:37 Claude Code: Anthropic's CLI Agent
May 7, 2025 positive
Assertion Supported
Cherny: Claude Code spawns parallel sub-agents to investigate complex coding tasks
“And so in the UI, when you say, when you see a task that's actually like a sub-Claud, it's a sub-agent that does this. And usually when I do something hairy, I'll ask it to just investigate, you know, three times or five times or however many times in parallel…”
Boris Cherny May 7, 2025 ▶ 53:42 Claude Code: Anthropic's CLI Agent
May 7, 2025 neutral
Assertion Not checkable as stated
Anthropic observes Claude Code API costs averaging roughly $6 daily per user
“Currently we're seeing costs around, like, six dollars per day per active user, and so it's, like, it does come out to a bit higher over the course of a month in Cursor but I don't think it's, like, out of band, and that's, like, roughly how we're thinking abo…”
Kat Wu May 7, 2025 ▶ 14:56 Claude Code: Anthropic's CLI Agent
May 7, 2025 positive
Assertion Partly supported
METR benchmark: AI agent autonomy duration doubles every 3 to 7 months
“They established a Moore's law for time between human input, basically, and it's basically doubling every three to seven months is the idea. And Enthopic is currently doing super well on that benchmark. It's roughly about autonomous for 15 minutes at the 50th …”
Shawn Wang May 7, 2025 ▶ 28:23 Claude Code: Anthropic's CLI Agent
May 7, 2025 positive
Assertion Not checkable as stated
Anthropic runs Claude Code in CI/CD to semantically lint GitHub pull requests
“For quad code internally in the GitHub repo, we have this GitHub action that runs. And the GitHub action invokes quad code with a local slash command. And the slash command is lint. So it just runs a linter using quad. And it's a bunch of things that are prett…”
Boris Cherny May 7, 2025 ▶ 21:47 Claude Code: Anthropic's CLI Agent
May 7, 2025 bullish
Insight
Wu: Internal operational tools are a major use case for Claude Code
“You mentioned internal tools, and that's actually a really big use case that we're seeing emerge, because a lot of times if you're working on something operationally intensive, if you can spin up a internal dashboard for it, or like an operational tool where y…”
Kat Wu May 7, 2025 ▶ 42:45 Claude Code: Anthropic's CLI Agent
May 7, 2025 positive
Assertion Not checkable as stated
Anthropic uses Claude to rewrite Claude Code from scratch every 4 weeks
“We've rewritten it from scratch, yeah, probably every three weeks, four weeks or something, and it just like all the, it's like a ship of Theseus, right? Like every piece keeps getting swapped out, and just because quad is so good at writing its own code.”
Boris Cherny May 7, 2025 ▶ 1:11:53 Claude Code: Anthropic's CLI Agent
May 7, 2025 neutral
Assertion Supported
Cherny: Anthropic is currently bordering on AI Safety Level 3 capabilities
“Yeah, we're kind of bordering on three right now.”
Boris Cherny May 7, 2025 ▶ 34:33 Claude Code: Anthropic's CLI Agent
May 7, 2025 bullish
Opinion
Cherny: Claude Code delivers up to 10x productivity gains for Anthropic engineers
“Anecdotally for me, it's probably two X my productivity. So I'm just like, I'm an engineer that codes all day, every day. For me, it's probably two X. Yeah. I think there's some engineers at Anthropic where It's probably 10 X their productivity”
Boris Cherny May 7, 2025 ▶ 1:05:57 Claude Code: Anthropic's CLI Agent
May 7, 2025 neutral
Assertion Supported
Cherny: Claude Code uses pure chain-of-thought, not Think Tool
“Yeah, this is, it is, it's all chain of thought, actually, in quad code. So we don't use the think tool. Anytime that quad code does thinking, it's all a chain of thought.”
Boris Cherny May 7, 2025 ▶ 51:42 Claude Code: Anthropic's CLI Agent
May 23, 2025 neutral
Opinion
Brown: Anthropic treats extended thinking as tool use, not distinct model class
“And it seemed like Anthropik's kind of attitude has been that extended thinking is an instance of tool use and that it's the kind of thing you want to equip the model with the ability to do. But it's not like, oh, it's a thinking model. It's just a sync for th…”
Will Brown May 23, 2025 ▶ 3:50 ⚡️Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect
May 23, 2025 neutral
Insight
Brown: Anthropic safety issues stem from conflicting model objectives
“A lot of the kind of headline anthropic like safety results, especially related to reward hacking and kind of deviation and alignment faking, Are all things to me that seem like a rock and a hard play situation where the model has two objectives it's given tha…”
Will Brown May 23, 2025 ▶ 13:22 ⚡️Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect
May 23, 2025 neutral
Opinion
Brown: Claude thinking and non-thinking modes likely use same underlying model
“I mean, I think these models should be the same model, and Anthropic knows what they're doing. Like, it's not that hard to, like, Quen did it in a very kind of, like, simple way, and they kind of talked about how they did it a little bit. But it's not, like, t…”
Will Brown May 23, 2025 ▶ 4:49 ⚡️Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect
May 23, 2025 neutral
Opinion
Will Brown: Claude appeals to AI insiders but lacks mainstream breakout
“It feels like people in the AI world, like, love Claude, or have grown type of Claude, but still had a phase where they were using it a ton. But it hasn't really broken out to general people in the way. And it feels like a lot of their marketing that I've seen…”
Will Brown May 23, 2025 ▶ 21:28 ⚡️Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect
Jun 6, 2025 neutral
Assertion Supported
Ameisen: Model internal representations show measurable bias toward English logits
“And it does seem like Does sort of like inner representations have a higher connection to like the output logits for English logits. And so there's like some bias towards English at least in the model we studied here.”
Emmanuel Ameisen Jun 6, 2025 ▶ 1:11:13 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 positive
Assertion Supported
Ameisen: LLMs use internal circuits to backwards-plan rhyming poetry lines
“And two, this plan doesn't just control, like, what you're gonna rhyme with. It's also doing what's called like backwards planning, where it's like, well, because I need to finish with green, I'm not going to say illuminating the peaceful night, because then I…”
Emmanuel Ameisen Jun 6, 2025 ▶ 1:18:59 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 positive
Assertion Supported
Ameisen: Circuit tracing notebooks run entirely on free Google Colab
“The notebooks themselves They can all be run on Google Colab and all of the code, as far as we can tell, we've like tested on the notebooks, just like runs on Colab. And so that means that like, you don't need on a free tier to be clear, like you don't need li…”
Emmanuel Ameisen Jun 6, 2025 ▶ 13:04 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 positive
Assertion Supported
Ameisen: Multi-Hop Reasoning Circuits Are Extremely Similar Across Small and Large Models
“The way the circuit looks in Gemma, like a really small model is extremely similar to the way that it looks like a huge model, which that in itself is, I think like a pretty novel discovery. It's like, oh, you have these models that are like super different. Y…”
Emmanuel Ameisen Jun 6, 2025 ▶ 3:36 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 positive
Assertion Not checkable as stated
Ameisen: Every interpretability team member joined partly due to Anthropic's interactive papers
“When we had a team meeting, like it was a couple months ago, somebody on the team asked how many of the people on this team are here, at least in part because they like read one of these papers and thought like, wow, this is so compelling. Like this like makes…”
Emmanuel Ameisen Jun 6, 2025 ▶ 1:46:07 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025
Insight
Ameisen: LLMs plan future tokens rather than operating purely myopically
“Language models are next token predictors is like a fact. Like that is what they do. They are trained to predict the next token. However, that does not mean that they myopically only consider the next token When they choose the next token, you can work on brea…”
Emmanuel Ameisen Jun 6, 2025 ▶ 1:13:16 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 positive
Insight
Ameisen: Interpretability research has lower entry barriers and low compute needs
“I think for Interp in particular, there's like another thing that makes it easier to transition to, which is maybe two things. One, you can just do it without huge access to compute. Like, there are open source models. You can look at them. A lot of Interp pap…”
Emmanuel Ameisen Jun 6, 2025 ▶ 28:56 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 positive
Prediction Not checkable as stated
Ameisen: Sparse autoencoder feature interpretability can and will be automated
“There's been a lot of work in sort of like automated feature interpretability. And it's something that we've invested in and that like other labs have invested in. And I think basically the answer is we can definitely automate it and We're definitely going to …”
Emmanuel Ameisen Jun 6, 2025 ▶ 1:23:12 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 neutral
Disclosure
Ameisen: Anthropic's circuit tracing tool ignores attention heads and only decomposes MLPs
“These are just MLPs. So the model has both attention heads and multi-layer perceptions MLPs. We don't just do it. Like we completely ignore attention or like we don't try to decompose it at all. So there's some prompts where like all of the interesting stuff i…”
Emmanuel Ameisen Jun 6, 2025 ▶ 15:56 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 neutral
Insight
Ameisen: LLMs Execute Parallel Sub-Processes During Math and Hallucinations
“So I think one example of this is like math where the model is like independently computing the like last digit and then the like order of magnitude and then kind of like combining them at the end or like hallucinations are also that where like, there's one si…”
Emmanuel Ameisen Jun 6, 2025 ▶ 1:26:11 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 positive
Assertion Supported
Ameisen: Anthropic's Open Tool Traces Internal States in Gemma 2 2B
“And then the release this week sort of lets anyone do it for a set of open source models. So notably maybe the most easy one here is like Gemma two to be. So you can sort of like think of some prompt and you kind of like can explain any like token that the mod…”
Emmanuel Ameisen Jun 6, 2025 ▶ 1:35 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 positive
Assertion Supported
Ameisen: Swapping Internal Features Proves Single-Pass LLM Multi-Step Reasoning
“We claim that this is like the Texas representation. Let's get another one and replace it. And we just change like that feature in the middle of the model and we change it to like California. And if you change it to California, sure enough, it says Sacramento.…”
Emmanuel Ameisen Jun 6, 2025 ▶ 1:00:02 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 negative
Opinion
Ameisen: Current LLM Chain of Thought Is Unfaithful and Untrustworthy
“So I think there's like a sense in which right now the chain of thought is, is unfaithful, or at least you can't read the chain of thought and trust that that's how the model did it.”
Emmanuel Ameisen Jun 6, 2025 ▶ 1:37:12 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 bullish
Assertion Not checkable as stated
Ameisen: Tracing prompt computation in models takes only minutes with built infrastructure
“One of the reasons that we're really excited about this method is once you've built your like infrastructure, like to go from a prompt to like what happened is, you know, O of minutes.”
Emmanuel Ameisen Jun 6, 2025 ▶ 1:47:52 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 bullish
Assertion Supported
Ameisen: Mechanistic interpretability methods successfully scaled to production models
“And it turns out scaling it. I don't want to say it just worked because it was a lot of work. I don't mean to apply. There was an effort, but it worked. And now we're in the phase where it's like, oh, cool. These methods work on the models that we care about.”
Emmanuel Ameisen Jun 6, 2025 ▶ 1:51:22 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025
Prediction Held up
Ameisen: Deceptive Backward Reasoning Exists in Base Pre-Trained Models
“I bet, I don't know how much I bet a hundred bucks. So somebody can like, they would get a hundred bucks from me if they prove that I'm wrong, that this behavior for a model that does a drink fine tuning, it also does it post pre-training.”
Emmanuel Ameisen Jun 6, 2025 ▶ 1:33:39 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 positive
Assertion Supported
Ameisen: Anthropic Released Circuit Tracing Code Built by Fellows
“And even more recently, we released some code in partnership with the Anthropic Fellows program. It was mostly built by Anthropic Fellows that lets people play with the research basically.”
Emmanuel Ameisen Jun 6, 2025 ▶ 0:38 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025
Disclosure
Ameisen: Golden Gate Claude was chosen organically after an internal demo
“Golden Gate Claude was like a pure, as far as I remember, at least, like, a pure, just like, Weird random thing where, like, somebody found it, initially went an internal demo of it, everybody thought it was hilarious, and then that's sort of how it came out. …”
Emmanuel Ameisen Jun 6, 2025 ▶ 44:15 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 negative
Assertion Not checkable as stated
Ameisen: Interpretability researchers lack good methods for analyzing attention layers
“So like, I think that right now we have some pretty good solutions for like understanding what's in the residual stream, understanding what's, is it in MLPs? We don't have good solutions for like attention.”
Emmanuel Ameisen Jun 6, 2025 ▶ 1:28:25 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025
Disclosure
Ameisen: Golden Gate Claude Was Created by Clamping a Bridge Feature
“That means that, like, if that's true, then you can, like, set that feature to zero, or artificially set to a hundred, And you'll change model behavior. That's what we did when we did Golden Gate Claude, in which we found a feature that represents the directio…”
Emmanuel Ameisen Jun 6, 2025 ▶ 40:29 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 positive
Disclosure
Ameisen: Anthropic publishes interpretability research to recruit more researchers
“The reason for publishing this is that we think interpretably is important. We think it's tractable, and we think more people should work on it. And so publishing it helps us like accomplish with these goals all these goals, which we think are just like crucia…”
Emmanuel Ameisen Jun 6, 2025 ▶ 1:40:51 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 positive
Assertion Supported
Ameisen: Larger language models share more concept representations across languages
“If you look inside the model, if you look at the middle of the model, which is the middle of this plot here, models share more features. They share more of these representations in the middle of the model, and bigger models share even more. And so the, like, t…”
Emmanuel Ameisen Jun 6, 2025 ▶ 1:07:57 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025
Assertion Supported
Ameisen: Anthropic Trained a Misaligned Model With Hidden Goals for Detection
“A team at Anthropic trained a model to have like weird hidden goals and then gave it to a bunch of other teams and said, Figure out what's wrong with it”
Emmanuel Ameisen Jun 6, 2025 ▶ 1:41:38 The Utility of Interpretability — Emmanuel Amiesen
Jun 25, 2025 positive
Disclosure
Zach Lloyd: Anthropic, not OpenAI, is Warp's primary model of choice
“ChatDBT isn't even the model of choice at this point. It's anthropic for us.”
Zach Lloyd Jun 25, 2025 ▶ 6:16 ⚡️Warp 2.0: the Agentic Development Environment - Zach Lloyd and Ben Holmes
Jun 25, 2025 neutral
Prediction Not checkable as stated
Zach Lloyd: Anthropic might launch a desktop agent harness like Warp
“I think even hearing the cloud code folks on your podcast, like I would not be surprised if Anthropic like launched a thing that looks a little bit more like warp where it's like a harness for running a whole bunch of different cloud codes, but it's an actual …”
Zach Lloyd Jun 25, 2025 ▶ 33:56 ⚡️Warp 2.0: the Agentic Development Environment - Zach Lloyd and Ben Holmes
Jun 25, 2025 neutral
Assertion Not checkable as stated
Zach Lloyd: Cursor represents a very significant portion of Anthropic's revenue
“Cursor I think is some very significant portion of their revenue.”
Zach Lloyd Jun 25, 2025 ▶ 35:57 ⚡️Warp 2.0: the Agentic Development Environment - Zach Lloyd and Ben Holmes
Jul 5, 2025 neutral
Assertion Supported
Davis: Multi-agent research consumes 15x baseline tokens versus 4x for single-agent
“So based off of a basic conversation, a single agent architecture for research is around four X, the number of tokens needed to achieve a research output. When you use multi-agent architectures, it's actually 15 X the number of tokens.”
Dylan Davis Jul 5, 2025 ▶ 8:58 ⚡️Anthropic vs Cognition on Multi-Agents: A Breakdown with Dylan Davis
Jul 5, 2025 neutral
Assertion Supported
Anthropic finds a single LLM judge outperforms five specialized judges
“They initially started with five LLM as judges. So each one of these points had their own LLM as a judge. They tested the ability and accuracy of that LLM as judge collective to judge, and it actually didn't perform as well as one. So they replaced all of thos…”
Dylan Davis Jul 5, 2025 ▶ 11:03 ⚡️Anthropic vs Cognition on Multi-Agents: A Breakdown with Dylan Davis
Jul 5, 2025 positive
Assertion Partly supported
Anthropic finds multi-agent architecture outperforms single-agent baseline by 80%
“So the, in the blog post that Anthropik posted, they ran some tests and they noticed that the multi-agent structure outperforms the single agent structure by 80% based off a different variety of variables they measured.”
Dylan Davis Jul 5, 2025 ▶ 8:31 ⚡️Anthropic vs Cognition on Multi-Agents: A Breakdown with Dylan Davis
Jul 14, 2025 bullish
Assertion Contradicted
Claude 3.7 remains unbeaten on Galileo Agent Leaderboard
“When we released the leaderboard and just in a week that launched 3.7, And that went straight up, and nobody has beaten it so far.”
Pratik Bhavsar Jul 14, 2025 ▶ 10:58 ⚡️Ranking Agentic LLMs — Pratik Bhavsar, Galileo
Jul 16, 2025 bullish
Prediction Not checkable as stated
Rizwan: AI ecosystem will converge around an official Anthropic MCP registry
“I think the entire ecosystem will just converge around whatever they do. They just have such good distribution and they're, yeah.”
Saoud (Saud) Rizwan Jul 16, 2025 ▶ 22:40 Cline: The Collaborative AI Coder
Jul 23, 2025 positive
Opinion
McCloy: Claude users represent an exceptionally valuable audience for companies
“Claude, which is important for, you know, not necessarily huge in terms of raw number of users, but the people who do use Claude tend to be like a very valuable audience, especially for some types of company.”
Robert McCloy Jul 23, 2025 ▶ 7:28 AI is Eating Search
Jul 28, 2025 bullish
Assertion Not checkable as stated
Hou: Windsurf is among Anthropic and OpenAI's largest consumers
“We've had immense success getting people onto the platform, and we've been very fortunate to have the issue of being some of Anthropic and OpenAI's largest consumers.”
Kevin Hou Jul 28, 2025 ▶ 2:49:15 🕰️ The Oral History of Windsurf (ft. Varun Mohan, Scott Wu, Jeff Wang, Kevin Hou, Anshul R)
Jul 31, 2025 positive
Insight
Lambert: OpenAI's Model Spec is more useful than Anthropic's Constitution
“The model spec is much more useful than a constitution because the constitution is like an intermediate training artifact that you give to the training algorithm in order to get the model that you want. It is not necessarily like what model did we, like we don…”
Nathan Lambert Jul 31, 2025 ▶ 1:03:38 The RLVR Revolution — with Nathan Lambert (AI2, Interconnects.ai)
Jul 31, 2025 neutral
Assertion Supported
Fanelli: Simon Willison reported Anthropic added Brave Search as a sub-processor
“Our friend Simon Willison wrote a post that Anthropic added Brave Search as one of the sub processor in their product.”
Alessio Fanelli Jul 31, 2025 ▶ 27:36 The RLVR Revolution — with Nathan Lambert (AI2, Interconnects.ai)
Aug 5, 2025 bullish
Prediction Not checkable as stated
Dax Reed predicts Cursor will move faster than Anthropic's Claude Code team.
“I mean, when I saw the acquisition, I was like, oh, this is actually more intense now because the cursor team is going to move faster than the cloud code team. I think that's my feeling given they have to, there's more pressure on them and they're smaller.”
Dax Reed Aug 5, 2025 ▶ 31:24 ⚡️OpenCode: Claude Code but Open Source, with Any Model, and frontier TUI - with Dax Reed (@thdxr)
Aug 5, 2025 bearish
Prediction Not checkable as stated
Dax Reed predicts Claude's $200/month pricing is an unsustainable growth strategy.
“I think it's pretty easy to, like, use more than 200 dollars worth, even by accident. So I would also lean towards that the Claude Max plans are a growth strategy, not like any long term pricing thing that can work, at least at the current, given the current s…”
Dax Reed Aug 5, 2025 ▶ 25:05 ⚡️OpenCode: Claude Code but Open Source, with Any Model, and frontier TUI - with Dax Reed (@thdxr)
Aug 5, 2025 bullish
Prediction Not checkable as stated
Dax Reed predicts OpenCode will dominate when a competitor beats Claude Sonnet.
“What would change things is if there's a day where either another LM lab or like, you know, an open source model drops that is competitive with Sonnet, maybe even better than Sonnet on that day, open code is going to be the only way to do this kind of thing. C…”
Dax Reed Aug 5, 2025 ▶ 28:28 ⚡️OpenCode: Claude Code but Open Source, with Any Model, and frontier TUI - with Dax Reed (@thdxr)
Aug 5, 2025 neutral
Disclosure
OpenCode exactly replicated Claude Code by dumping its system prompts and schemas.
“We look at cloud code. We dump all their system prompts. We dump all their tool descriptions. We dump all their tool schemas. We re-implement the tools. And when you're using an anthropic model, we basically have the exact same implementation.”
Dax Reed Aug 5, 2025 ▶ 10:39 ⚡️OpenCode: Claude Code but Open Source, with Any Model, and frontier TUI - with Dax Reed (@thdxr)
Aug 6, 2025 neutral
Assertion Supported
Palazzolo: Claude Code leads stayed at Cursor only two weeks
“We know that they went there, they were there for, I think, about two weeks, and they came back.”
Stephanie Palazzolo Aug 6, 2025 ▶ 41:45 The AI Agenda: GPT5 leaks and the business of AI News — Steph Palazzolo, The Information
Aug 6, 2025 positive
Assertion Partly supported
Chroma research finds Claude models lead in long-context utilization
“And you know, one thing Chroma released this context rod paper recently about context utilization and the cloud models are actually the best at using kind of like longer context.”
Alessio Fanelli Aug 6, 2025 ▶ 37:36 The AI Agenda: GPT5 leaks and the business of AI News — Steph Palazzolo, The Information
Sep 1, 2025 positive
Disclosure
Weichel: Ona is built primarily on Anthropic's Claude Sonnet
“So we, right now essentially live on top of Sonnet IV, and we've tried a bunch of other models. We found that to work really, really well for us.”
Chris (Christian Weichel) Sep 1, 2025 ▶ 18:57 ⚡️Launching Ona: Coding Agent with Fully Sandboxed Cloud Environment
Sep 8, 2025 bearish
Opinion
Gorkem: Hosting LLMs is a bad business due to Google search competition
“Language models, hosting language models is not a good business. At the time we thought, okay, we are going to be competing against OpenAI and Anthropic and all these labs. Turned, turned out that it was even worse because the killer application of language mo…”
Gorkem Yurtseven Sep 8, 2025 ▶ 8:50 A Technical History of Generative Media
Sep 11, 2025 neutral
Assertion Partly supported
Anthropic: Typical production agents execute hundreds of tool calls per task
“Anthropics multi-agent research is another nice example of this. They mentioned that the typical production agent, and this is probably referring to Cloud Code, could be other agents that they've produced, is like hundreds of tool calls.”
Lance Martin Sep 11, 2025 ▶ 3:28 Context Engineering for Agents - Lance Martin, LangChain
Sep 11, 2025 neutral
Assertion Supported
Martin: Claude Code triggers context compaction at 95% of context window
“If you use Cloud Code, you hit that Kind of, you know, you've hit 95% of the context window, and you're about to, and Cloud Code's about to perform compaction.”
Lance Martin Sep 11, 2025 ▶ 27:42 Context Engineering for Agents - Lance Martin, LangChain
Sep 11, 2025 positive
Insight
Martin: Multi-agent systems excel at parallel read-only tasks, not writing tasks
“I like the take that apply multi-agents to problems that are easily parallelizable, that are read-only, for example, context gathering for deep research, and do, like, the final quote-unquote write, in this case report writing, at the end. I think this is tric…”
Lance Martin Sep 11, 2025 ▶ 14:54 Context Engineering for Agents - Lance Martin, LangChain
Sep 11, 2025 neutral
Assertion Supported
Martin: Anthropic uses parallel sub-agents for research and single-shot final writing
“Anthropic reported on this too. So their deep researcher just uses parallelized subagents for research collation, and they do the writing in one shot at the end.”
Lance Martin Sep 11, 2025 ▶ 13:50 Context Engineering for Agents - Lance Martin, LangChain
Sep 11, 2025 positive
Assertion Supported
Martin: Claude Code operates entirely without codebase indexing
“Clock code doesn't do any indexing. It's just doing, quote unquote, agentic retrieval, just using simple tool calls, for example, using grep, to kind of poke around your files, no indexing whatsoever, and obviously works extremely well.”
Lance Martin Sep 11, 2025 ▶ 17:02 Context Engineering for Agents - Lance Martin, LangChain
Sep 25, 2025 negative
Assertion Supported
Rajpal: Anthropic Claude models had regressions from serving architecture changes
“Anthropix kind of cloud models kind of had a regression, right? Because they changed to a new serving architecture.”
Shreya Rajpal Sep 25, 2025 ▶ 23:21 ⚡️Snowglobe: Simulations for your AI
Sep 25, 2025 negative
Insight
Ball: Custom MCP tools fail if workflows diverge from frontier training
“If I give it this other custom-made MCP that we built internally, and our processes don't map to anything that OpenAI and Anthropic have seen or trained for, it won't be used, and you won't get good results.”
Thorsten Ball Sep 25, 2025 ▶ 49:15 Amp: The Emperor Has No Clothes
Sep 25, 2025
Disclosure
Ball: The Sourcegraph Amp team does not use formal evals
“I think we don't have any set evals. We don't. And this was controversial up until a week ago, I think, when I think Boris from Or two weeks ago from Anthropix that they don't have evals for the coding agent too. But we don't, and we haven't had them.”
Thorsten Ball Sep 25, 2025 ▶ 1:18:11 Amp: The Emperor Has No Clothes
Sep 25, 2025 neutral
Assertion Not checkable as stated
Fanelli: Cursor Default Model Switch Cost Anthropic $200M in ARR
“When cursors switch from Sonnet to GPT-V as like the default model that was like, you know, Two hundred million our revenue for Anthropic that kind of went away and like moved on to GPT-V.”
Alessio Fanelli Sep 25, 2025 ▶ 29:57 Amp: The Emperor Has No Clothes
Sep 25, 2025 bearish
Prediction Not checkable as stated
Slack: Top AI labs face a major customer stampede within two months
“I think we are one or two months away from a possible news cycle. That is the foundation model companies have spent billions of dollars in capex and hired like crazy. And now, you know, they're no longer the best in this realm and there's a huge stampede away …”
Quinn Slack Sep 25, 2025 ▶ 30:49 Amp: The Emperor Has No Clothes
Sep 30, 2025 bullish
Disclosure
Krieger: Anthropic generates internal dashboards on demand with Claude
“Even internally we have found that having Cloud generate UI on demand for internal dashboards is useful, not just like an interesting demo.”
Mike Krieger Sep 30, 2025 ▶ 8:27 ⚡️Claude Sonnet 4.5 and Anthropic's roadmap for Agents and Developers — Mike Krieger, Anthropic
Sep 30, 2025 positive
Insight
Krieger: AI Agents Must Support Both MCP and Visual Computer Use
“And that thing's never gonna have an MCP around it. Like, it's just like, who knows if the company created is even around much less like ready to sort of expose their kind of underlying constructs as API. So I think you will need to be able to do both.”
Mike Krieger Sep 30, 2025 ▶ 13:00 ⚡️Claude Sonnet 4.5 and Anthropic's roadmap for Agents and Developers — Mike Krieger, Anthropic
Sep 30, 2025 bullish
Disclosure
Krieger: Anthropic plans hosted computation options for Claude Agent SDK
“There's Cloud Code, which is directly built on top of the Cloud Agent SDK, and then there's external companies building on top of the Cloud Agent SDK, and I think we'll also offer ways in which if you want to run an agent off of that SDK, but have a lot of the…”
Mike Krieger Sep 30, 2025 ▶ 23:19 ⚡️Claude Sonnet 4.5 and Anthropic's roadmap for Agents and Developers — Mike Krieger, Anthropic
Sep 30, 2025 bullish
Disclosure
Krieger: Anthropic to integrate agent harness into Claude AI workflows
“There's, Cloud AI, and I think you'll see us start bringing that harness into Cloud AI more and more for things like document creation, advanced research, and so it'll power a lot more of the agentic workflows in there.”
Mike Krieger Sep 30, 2025 ▶ 23:10 ⚡️Claude Sonnet 4.5 and Anthropic's roadmap for Agents and Developers — Mike Krieger, Anthropic
Sep 30, 2025 positive
Assertion Not checkable as stated
Sonnet 4.5 ran autonomously for 30 hours versus 7 for Opus 4
“So this, you know, we had a customer and internally, we also got like a 30 hour plus kind of execution versus I think Opus four was seven hours.”
Mike Krieger Sep 30, 2025 ▶ 20:51 ⚡️Claude Sonnet 4.5 and Anthropic's roadmap for Agents and Developers — Mike Krieger, Anthropic
Sep 30, 2025 positive
Opinion
Krieger: Claude Sonnet 4.5 Outperforms Opus at Generating 3D Games
“This is like officially good. It's like better than Opus at this. It's like, It generated this, like, great split-screen stereoscopic thing, three-dimensional, like, thing.”
Mike Krieger Sep 30, 2025 ▶ 4:24 ⚡️Claude Sonnet 4.5 and Anthropic's roadmap for Agents and Developers — Mike Krieger, Anthropic
Sep 30, 2025 bullish
Assertion Not checkable as stated
Krieger: Claude Sonnet 4.5 day-one traffic eclipsed Sonnet 4
“We have more traffic on Sonnet 4.5 than we had on Sonnet four. So basically it's already eclipsed Sonnet four.”
Mike Krieger Sep 30, 2025 ▶ 1:15 ⚡️Claude Sonnet 4.5 and Anthropic's roadmap for Agents and Developers — Mike Krieger, Anthropic
Sep 30, 2025 positive
Assertion Not checkable as stated
Krieger: Sonnet 4.5 First Anthropic Model With Upstream Product Input
“What was interesting about this model in particular is that it was really the first one where Product was upstream of research and downstream of research.”
Mike Krieger Sep 30, 2025 ▶ 2:05 ⚡️Claude Sonnet 4.5 and Anthropic's roadmap for Agents and Developers — Mike Krieger, Anthropic
Oct 5, 2025 bullish
Opinion
Dwivedi: Claude is superior at agentic tool calling and error unstacking
“For some of the agentic part of the stack, we are shifting towards Anthropic because they're agentic and the tool calling, especially the unstacking part, you know, when you go down the wrong path and you build context that forces you to keep going down the wr…”
Raaz Dwivedi Oct 5, 2025 ▶ 33:36 ⚡️Traversal: Causal ML and Reinforcement Learning
Oct 7, 2025 positive
Assertion Not checkable as stated
John Schulman developed the Tinker API concept across OpenAI and Anthropic
“Right when I joined OpenAI, like, this has actually been, I think, a passion project of John's. Like, he's been talking about doing something in this, like, in this shape for a while, which is, like, a truly, like, low-level research, like, fine-tuning library…”
Sherwin Wu Oct 7, 2025 ▶ 22:35 DevDay 2025: Apps SDK, Agent Kit, MCP, Codex and why Prompting is More Important than Ever
Oct 18, 2025 positive
Assertion Supported
Merrill: Dario Amodei highlighted Terminal-Bench on the Claude model card
“I think one of the really key moments for us was getting onto the Claude IV model card. Being one of two benchmarks that Dario actually mentioned while releasing the model.”
Mike Merrill Oct 18, 2025 ▶ 3:58 Terminal-Bench: Pushing Claude Code, OpenAI Codex, Factory Droid, et al to the limits
Oct 24, 2025 neutral
Insight
Webster: Enterprises do not want AI models to be maximally helpful
“OpenAI Anthropic, everyone else, they're all building models that are like maximally helpful. And in Actually, most cases in a corporate environment, you don't want that to be maximum. You don't want the model to be like helpful in every way possible.”
Ian Webster Oct 24, 2025 ▶ 8:52 Breaking AI to Fix It: Ian Webster's Journey from Discord's Clyde to Promptfoo's $18M Series A
Nov 2, 2025 bullish
Prediction Not checkable as stated
Swyx: Tuning reasoning activations could let Anthropic leapfrog OpenAI
“If Anthropic ever found The activations for reasoning and could break down the different kinds of reasoning and turn, tune them properly. I think that's the thing that takes Anthropic to leapfrog OpenAI.”
Shawn Wang Nov 2, 2025 ▶ 4:33 ⚡️Automating Scientific Discovery - Jessica Rumbelow, Leap Labs
Nov 3, 2025 neutral
Prediction Open · timeframe Nov 2030
Anthropic and OpenAI will never open-source their high-performance inference kernels
“The high performance inference kernels that sort of drive a lot of, you know, anthropic and open AI and stuff, their models, those aren't open source. They're not going to be open source.”
Quentin Anthony Nov 3, 2025 ▶ 40:52 How Zyphra went all-in on AMD + Why Devs feel faster with AI but are slower — with Quentin Anthony
Nov 14, 2025 bullish
Opinion
Anthropic's MCP Has Won as the De Facto AI Interop Standard
“Now it's like basically kind of de facto one as the interop layer for all the labs and all the models.”
Shawn Wang Nov 14, 2025 ▶ 1:09:07 Anthropic, Glean & OpenRouter: How AI Moats Are Built with Deedy Das of Menlo Ventures
Nov 14, 2025 positive
Assertion Supported
Anthropic Maintains an 80% One-Year Employee Retention Rate
“I'm referring to the exact same article where I think their retention, one year retention on employees is the 80%, which in AI world is, is quite wild.”
Deedy Das Nov 14, 2025 ▶ 21:26 Anthropic, Glean & OpenRouter: How AI Moats Are Built with Deedy Das of Menlo Ventures
Nov 14, 2025 neutral
Insight
Menlo Underwrites Late-Stage AI on Revenue and Margin, Not Market Share
“At that stage, to be very honest with you, at the stage that we invest in Anthropic now, like the only things that would really move the needle on the decision is here's the revenue, here's the margin, and here's the trajectory, and here's the other markets we…”
Deedy Das Nov 14, 2025 ▶ 27:33 Anthropic, Glean & OpenRouter: How AI Moats Are Built with Deedy Das of Menlo Ventures
Nov 14, 2025 bullish
Insight
Model Layer Is Structurally More Defensible Than the AI App Layer
“It is far easier for Anthropic to try to go into one of the spaces of the apps than an app to try to go into the space of Anthropic, which makes me feel like one is more defensible than the other all else equal.”
Deedy Das Nov 14, 2025 ▶ 35:32 Anthropic, Glean & OpenRouter: How AI Moats Are Built with Deedy Das of Menlo Ventures
Nov 14, 2025 bearish
Prediction Not checkable as stated
AI Labs Will Not Dedicate Engineering Talent to Deep Enterprise Search
“If you really want to go deep, I don't think you will ever dedicate the people to do it. And the last thing I'll say is you think about from an anthropic engineer's perspective, you joined a big AI lab to work on models, not to build Google drive connectors, r…”
Deedy Das Nov 14, 2025 ▶ 10:10 Anthropic, Glean & OpenRouter: How AI Moats Are Built with Deedy Das of Menlo Ventures
Nov 14, 2025
Assertion Supported
Glean Operates at a Several Hundred Million Dollar Revenue Scale
“Look at the revenue of Anthropic and OpenAI right now. These are billion dollar revenue scale businesses. Glean is several hundred million dollar revenue scale business.”
Deedy Das Nov 14, 2025 ▶ 9:16 Anthropic, Glean & OpenRouter: How AI Moats Are Built with Deedy Das of Menlo Ventures
Nov 14, 2025 bullish
Insight
AI Coding Intelligence Has an Uncapped Frontier That Drives Revenue
“But the interesting about Anthropic is if you look at coding, that's probably never going to be the case. Like there's always an increasing frontier of how you good you could be at a task like that. And we're nowhere close to that frontier. So it's more possib…”
Deedy Das Nov 14, 2025 ▶ 31:43 Anthropic, Glean & OpenRouter: How AI Moats Are Built with Deedy Das of Menlo Ventures
Nov 14, 2025 neutral
Insight
Enterprise LLM Churn Is Low Due to Long-Term Compute Commitments
“In terms of enterprises, often what will happen is they'll buy up large chunks of long-term compute and dedicated instances, in which case you just don't churn, right? Like this is what you use.”
Deedy Das Nov 14, 2025 ▶ 26:26 Anthropic, Glean & OpenRouter: How AI Moats Are Built with Deedy Das of Menlo Ventures
Nov 14, 2025 bullish
Assertion Partly supported
Anthropic Is the Fastest-Growing Software Company in History
“Anthropic is the fastest growing software company of all time. I think I can say that fairly. I'm, I haven't been disproven yet.”
Deedy Das Nov 14, 2025 ▶ 17:16 Anthropic, Glean & OpenRouter: How AI Moats Are Built with Deedy Das of Menlo Ventures
Dec 6, 2025 neutral
Disclosure
General Intuition's initial commercial business model will be an API
“Our business model is initially going to be an API, again, like the Anthropic API”
Pim de Witte Dec 6, 2025 ▶ 57:03 World Models & General Intuition: Khosla's largest bet since LLMs & OpenAI
Dec 7, 2025 neutral
Opinion
Ubl: Anthropic's Boris Power Vibe-Codes From a Position of Privilege
“I think he comes from a particular position of extreme unusual privilege, which is that he works at an AI lab where like people in the office next door are like writing the evals and are like training the model like every day in exactly that way.”
Malte Ubl Dec 7, 2025 ▶ 5:44 The Great Evals Debate — Ankur Goyal & Malte Ubl
Dec 7, 2025 positive
Disclosure
Ubl: Vercel publishes evals to influence OpenAI and Anthropic models
“I'm Vercel and I publish at Eval. That I want OpenAI and Anthropic to use to make sure when they ship the next model that they're better at the stuff that I care about.”
Malte Ubl Dec 7, 2025 ▶ 28:12 The Great Evals Debate — Ankur Goyal & Malte Ubl
Dec 11, 2025
Disclosure
Superhuman builds dynamic on-the-fly aggregation lambdas with Anthropic
“We're working right now with Anthropic to basically do kind of like a building on the fly, small, kind of a key component of lambdas that will build the code to do the aggregation.”
Loïc Houssier Dec 11, 2025 ▶ 25:26 The Future of Email: Superhuman CTO on Your Inbox As the Real AI Agent (Not ChatGPT) — Loïc Houssier
Dec 16, 2025 negative
Assertion Supported
Pliny: Anthropic added a $20k–$30k bounty but withheld jailbreak data
“That whole thing ended with no open sourcing of data, but they did add a 30,000 or 20,000 dollar bounty, which I sort of sat myself out of, let the community go for it.”
Pliny the Liberator Dec 16, 2025 ▶ 19:08 ⚡️Jailbreaking AGI: Pliny the Liberator & John V on Red Teaming, BT6, and the Future of AI Security
Dec 16, 2025 neutral
Opinion
Pliny: AI labs lack enough researchers to explore latent space alone
“They don't have enough researchers to explore the entire latent space on their own. And so I think many hands make light work”
Pliny the Liberator Dec 16, 2025 ▶ 19:03 ⚡️Jailbreaking AGI: Pliny the Liberator & John V on Red Teaming, BT6, and the Future of AI Security
Dec 26, 2025 neutral
Assertion Not checkable as stated
Yegge: Anthropic is hiring over 100 people for Claude Code
“They're hiring like a hundred plus people for cloud code in the next, I don't know, month. I mean, like they're going wild and that's just cloud code.”
Steve Yegge Dec 26, 2025 ▶ 30:13 Steve Yegge's Vibe Coding Manifesto: Why Claude Code Isn't It & What Comes After the IDE
Dec 26, 2025 negative
Opinion
Yegge: Google, Anthropic, and OpenAI are unbelievably chaotic internally
“All three of those companies, Google, Anthropic, and OpenAI are unbelievably chaotic internally right now.”
Steve Yegge Dec 26, 2025 ▶ 29:48 Steve Yegge's Vibe Coding Manifesto: Why Claude Code Isn't It & What Comes After the IDE
Dec 28, 2025 positive
Disclosure
Anthropic remains fully committed to MCP following its foundation donation
“Like the commitment of Anthropic is the same, right? I'm still, We still have the same people I'm helping with the SDKs. We're still super committed in our products to MCP. I'm still the lead core maintainer. Nothing has actually changed.”
David Soria Parra Dec 28, 2025 ▶ 1:01:23 One Year of MCP — with David Soria Parria and AAIF leads from OpenAI, Goose, Linux Foundation
Dec 28, 2025 neutral
Disclosure
Block abandoned Goose's proprietary plugin ecosystem in favor of adopting MCP
“We had a version of Goose that was still, you can go check the GitHub history. It was there a little bit before MCP came out, and we were sitting there with like a plugin ecosystem who were like, this is awful. Like what, like, why would you, why would anyone?…”
Nick Cooper / Brad Dec 28, 2025 ▶ 1:13:12 One Year of MCP — with David Soria Parria and AAIF leads from OpenAI, Goose, Linux Foundation
Dec 28, 2025
Disclosure
Anthropic created MCP so rapidly expanding internal teams could build integrations independently
“MCP before we even open source it was born of the idea of like, I'm in a company that is growing crazy. I'm in the development side of things, development tooling side of things. I will grow slower than the rest. How can I build something that they can all bui…”
David Soria Parra Dec 28, 2025 ▶ 28:18 One Year of MCP — with David Soria Parria and AAIF leads from OpenAI, Goose, Linux Foundation
Dec 28, 2025 positive
Assertion Supported
Block's Goose was the first open-source agent to integrate MCP
“Goose was the first open source agent interface or agent that reached out to us and worked with us to integrate MCP. And I think Rad is actually like technically the first non-anthropic contributor to MCP ever on like day two or something like that, like very,…”
David Soria Parra Dec 28, 2025 ▶ 1:12:42 One Year of MCP — with David Soria Parria and AAIF leads from OpenAI, Goose, Linux Foundation
Dec 28, 2025 positive
Disclosure
Anthropic always planned to put MCP into a neutral open-source foundation
“The first one is that on the MCP side, we always knew that we wanted to find a neutral home for MCP to make sure that the that the industry understands that this is stays open, that this is something safe to adopt.”
David Soria Parra Dec 28, 2025 ▶ 1:05:16 One Year of MCP — with David Soria Parria and AAIF leads from OpenAI, Goose, Linux Foundation
Dec 28, 2025 positive
Assertion Supported
Google, Microsoft, Amazon, OpenAI, and Anthropic joined AAIF as platinum members
“You have Google, Microsoft, Amazon Block, Bloomberg, Cloudflare, OpenAI, Anthropic. Just a platinum member, create a foundation.”
David Soria Parra Dec 28, 2025 ▶ 1:35:53 One Year of MCP — with David Soria Parria and AAIF leads from OpenAI, Goose, Linux Foundation
Dec 28, 2025 neutral
Assertion Not checkable as stated
AAIF began when Block asked Anthropic about donating the Goose agent
“We got approached by our friends at block to discuss because they were looking into like donating goose, I think at the time. And so there was a question around doing something together. And then we approached open AI and they were very, very welcoming and lik…”
David Soria Parra Dec 28, 2025 ▶ 1:05:44 One Year of MCP — with David Soria Parria and AAIF leads from OpenAI, Goose, Linux Foundation
Dec 28, 2025
Assertion Supported
Agentic AI Foundation is the first open-source foundation founded by Anthropic
“Like for us, it's the first time we at Anthropic have an open source foundation.”
David Soria Parra Dec 28, 2025 ▶ 4:21 One Year of MCP — with David Soria Parria and AAIF leads from OpenAI, Goose, Linux Foundation
Dec 30, 2025 neutral
Opinion
Nair: Frontier AI labs have converged on similar reinforcement learning methods
“Well, it does seem like basically a lot of the labs have kind of like converged onto some similar-ish way of doing RL, and they're all kind of back at the same level of like Frontier again”
Ashvin Nair Dec 30, 2025 ▶ 31:54 [State of RL/Reasoning] IMO/IOI Gold, OpenAI o3/GPT-5, and Cursor Composer — Ashvin Nair, Cursor
Jan 9, 2026 positive
Assertion Supported
Anthropic Claude models have lowest hallucination rates on Omniscience benchmark
“Like, one of the things that we saw in the hallucination rate is that Anthropoc's Claude models at the very left-hand side here with the lowest hallucination rates out of the models that we've evaluated Amnesians on.”
Micah Hill-Smith Jan 9, 2026 ▶ 30:09 Artificial Analysis: The Independent LLM Analysis House — with George Cameron and Micah Hill-Smith
Feb 19, 2026 positive
Assertion Not checkable as stated
Wang: Anthropic Claude Cowork automates complex customer cohort data analysis
“Like for the first time you can actually get one shot data analysis, right? Which, you know, if you're going to do a customer database, analyze a cohort retention, right? That's just stuff that you had to do by hand before. And our team, the other, it was like…”
Sarah Wang Feb 19, 2026 ▶ 27:34 Inside AI’s $10B+ Capital Flywheel — Martin Casado & Sarah Wang of a16z
Feb 19, 2026 bearish
Insight
Casado: Frontier AI Labs Could Outspend and Consume Application Layer Startups
“It literally becomes an issue of like raise capital, turn that directly into growth, use that to raise three times more. And if you can keep doing that, you literally can outspend any company that's built. Not any company. You can outspend the aggregate of com…”
Martin Casado Feb 19, 2026 ▶ 10:51 Inside AI’s $10B+ Capital Flywheel — Martin Casado & Sarah Wang of a16z
Feb 24, 2026 bearish
Assertion Supported
O'Laughlin: Anthropic does not train Claude agent teams with RL
“I have a controversial opinion that Claude does not do RL on the agent swarms or agent team.”
Doug O'Laughlin Feb 24, 2026 ▶ 33:33 Claude Code for Finance + The Global Memory Shortage: Doug O'Laughlin, SemiAnalysis
Feb 26, 2026 bullish
Assertion Not checkable as stated
Patel: Anthropic added $2B in monthly revenue at positive margins
“Anthropic doesn't just add two billion dollars of revenue in one month. You know, with, without having, you know, huge demand and they're doing it at positive margins, right?”
Dylan Patel Feb 26, 2026 ▶ 25:21 Dylan Patel Explains the AI War While Cooking | In-Context Cooking
Mar 5, 2026 neutral
Assertion Supported
Levie: Anthropic has forward-deployed engineers embedded at Goldman Sachs
“OpenAI probably is hiring FDEs to go into the enterprise and then Anthropic is embedded at Goldman Sachs.”
Aaron Levie Mar 5, 2026 ▶ 17:33 Why Every Agent Needs a Box — Aaron Levie, Box
Mar 12, 2026 neutral
Assertion Not checkable as stated
Eskildsen: Anthropic, Notion, and Cursor use Turbopuffer across three deployment models
“You can run Turbo Puffer either in SAS, right? That's what cursor does. You can run it in a single tenant cluster. So it's just you. That's what Notion does. And then you can run it in, in, in BYOC where everything is inside the customer's VPC. That's what, fo…”
Simon Eskildsen Mar 12, 2026 ▶ 37:25 Retrieval After RAG: Hybrid Search, Agents, and Database Design — Simon Eskildsen of Turbopuffer
Mar 17, 2026 bullish
Prediction Not checkable as stated
Rieseberg: AI takeoff will create an accelerating, self-reinforcing loop
“Big bang moment where things will accelerate so quickly that it becomes a self-reinforcing loop. And at that point it's sort of like off to the races and there will be no more like slowly catching up. You know, just have Claude being so good at everything.”
Felix Rieseberg Mar 17, 2026 ▶ 52:04 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026
Disclosure
Rieseberg automates bug triage and dispatch using Claude Cowork loops
“The way I'm using it is I have co-work running, and I'm telling co-work, here's where I normally go every morning to find the latest bugs. Go read the entire bug list, separate out which ones are fixable, which ones are not fixable, and then for the fixable on…”
Felix Rieseberg Mar 17, 2026 ▶ 16:43 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026
Disclosure
Rieseberg: Anthropic builds all prototype candidates quickly instead of writing memos
“We internally at Anthropica are now probably much closer to the point where, like, don't even write a memo. Just, like, build, like, let's build all the candidates very quickly. Let's just build all of them and then pick the best ones.”
Felix Rieseberg Mar 17, 2026 ▶ 7:44 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 positive
Disclosure
Rieseberg: Claude Cowork prompt automatically prioritizes meetings with Dario Amodei
“I've given it like pretty clear instructions about, okay, here are some people, if they book over other meetings, I'm probably going to go to their meeting. Like if Dario schedules a meeting. Not try to reschedule Dario.”
Felix Rieseberg Mar 17, 2026 ▶ 31:22 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 bullish
Insight
Rieseberg: Prompt Opus by stating goals, not specifying exact execution steps
“Honestly though, like I see that you're using Opus 4.6, right? Like my recommendation for people is increasingly don't worry about it anymore. Just like tell it what you want it to do. And it's probably going to figure out a way to do it.”
Felix Rieseberg Mar 17, 2026 ▶ 53:13 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 bearish
Disclosure
Anthropic deeply worries about AI automating entry-level junior jobs
“At Anthropic as a group of people, we're deeply worried about the impact that the tools are going to have on the labor market, especially for like junior employees that, because I think it's only honest to say that when we talk about automating a lot away, a l…”
Felix Rieseberg Mar 17, 2026 ▶ 46:13 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 bullish
Disclosure
Rieseberg: Claude Cowork will release new features weekly
“We're going to keep shipping things that we're going to keep shipping things that we're going to keep iterating on this thing like pretty quickly, but which I mean, you can sort of continue to expect that every single week, there's going to be like a small new…”
Felix Rieseberg Mar 17, 2026 ▶ 1:11:21 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026
Disclosure
Rieseberg: Anthropic did not build Claude Cowork primarily for coding
“I caught myself, like, starting to use Cowork for coding tasks, which is not ostensibly what we built it for, right?”
Felix Rieseberg Mar 17, 2026 ▶ 15:39 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 positive
Insight
Rieseberg: Cowork's most impressed users find unexpected capabilities
“Every single person who's like most amazed is usually amazed about a thing that I didn't even expect Cowork would be good at.”
Felix Rieseberg Mar 17, 2026 ▶ 2:06 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 positive
Insight
Rieseberg: Aggressively anthropomorphizing Claude improves agent UX and architecture design
“And in terms of architecture and UX and everything else that we've been working on Anthropic, it often is quite useful for you to like anthropomorphize cloud aggressively and just be like, this is a person. What would you do if you give, if you had a person, r…”
Felix Rieseberg Mar 17, 2026 ▶ 13:05 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 neutral
Disclosure
Rieseberg: Anthropic's skills feature originated from a simple markdown API prompt
“The thing that ultimately led to skills is that we wanted to connect this little prototype to our data warehouse. And the team very quickly discovered that like, instead of building a custom tool for the thing to talk to our data warehouse, They just, like, ma…”
Felix Rieseberg Mar 17, 2026 ▶ 26:35 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026
Disclosure
Rieseberg: Claude Cowork assembled existing prototypes rather than starting from scratch
“And what cowork actually became is like, we sort of picked the right pieces out of the many prototypes that we had. Right. And that's maybe also like, I think an important qualifier whenever people mention this like 10 day number, I do think it's important to …”
Felix Rieseberg Mar 17, 2026 ▶ 6:40 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 positive
Disclosure
Rieseberg: Claude Cowork uses Apple Virtualization framework on macOS
“So on, on macOS, we use the Apple virtualization framework, which is pretty solidly optimized. Like it's good stuff.”
Felix Rieseberg Mar 17, 2026 ▶ 1:05:37 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 neutral
Assertion Not checkable as stated
Rieseberg: Desktop cleanup was one of the most viral launch use cases for Cowork
“Right when we launched Cowork, I think one of the users that went most viral on Twitter X was clean up your desktop, which is of course silly. That's such a smart thing, right? Like you don't need a model to clean up your desktop.”
Felix Rieseberg Mar 17, 2026 ▶ 31:48 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 positive
Assertion Supported
Rieseberg: Claude Cowork is Claude Code running in a sandboxed virtual machine
“Cowork is cloud code running in a virtual machine with a little bit of padding, a little bit more guardrails, making it a little safer, a little bit more convenient for people who don't want to first open up the terminal when they go to work.”
Felix Rieseberg Mar 17, 2026 ▶ 3:35 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 positive
Disclosure
Rieseberg: Anthropic built tech enabling Claude to interact via Google Doc comments
“We built so much tech around Claude leaving useful comments inside a Google Doc, and now it just does it, just like leaves a comment in your Google Doc and that's how you interact with it.”
Felix Rieseberg Mar 17, 2026 ▶ 1:23:08 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 bullish
Prediction Not checkable as stated
Rieseberg: Claude is close to effectively controlling real user computers
“I don't think we're far away from claw being very effective at like using your computer and not just a theoretical computer.”
Felix Rieseberg Mar 17, 2026 ▶ 57:51 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026
Disclosure
Anthropic evaluates Claude Cowork on finance and legal tasks, not coding
“So like clock code is like quite optimized for coding tasks, and we mostly evaluate whether or not we're getting better or worse, depending on how good it is at like a typical sweet job. And cloud co-work on the other hand, we evaluate more against typical kno…”
Felix Rieseberg Mar 17, 2026 ▶ 20:23 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 neutral
Disclosure
Rieseberg: Anthropic is building custom network drivers for Claude
“Last week we started writing a new networking service and networking driver. That optimizes how Claw talks to the internet.”
Felix Rieseberg Mar 17, 2026 ▶ 1:04:40 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026
Disclosure
Anthropic prompts Claude Cowork to clarify user intent before long executions
“We do tell co-work to make heavy use of the planning tool or to make heavy use of the ask user question tool, right? We do want it to come up with like, Different scenarios of, okay, tease out what the user actually wants. Don't go off to work for like four ho…”
Felix Rieseberg Mar 17, 2026 ▶ 22:14 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 positive
Disclosure
Anthropic: Claude Cowork executes inside a dedicated lightweight Linux VM
“So we currently run like a, we currently run like a lightweight VM and we put clock code into the VM and we do that for a number of reasons. Safety and security is a big one, but even if you ignore for a second safety and security and you're just like, okay, Y…”
Felix Rieseberg Mar 17, 2026 ▶ 12:44 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 positive
Insight
Rieseberg: Claude Cowork skills can be as simple as a text message
“One thing that is very fun for me about skills in particular is that they're so easy to make. Like anyone can make a skill, like a text message could be a skill and they can be so hyper-personalized to you.”
Felix Rieseberg Mar 17, 2026 ▶ 29:06 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Mar 17, 2026 positive
Insight
Rieseberg: A labs team should only tackle ideas no one else would
“The sort of the idea of a Labs team is that it should only work on things that make really no sense for anyone else to work on.”
Felix Rieseberg Mar 17, 2026 ▶ 1:27:04 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Apr 15, 2026
Disclosure
Notion Partnered With Anthropic and OpenAI to Build 30% Pass Rate Evals
“And then what we have, what we call Frontier Headroom evals, where we actively want to be at 30% pass rate. And that's actually been a effort that we took in partnership with Anthropic and OpenAI in the past maybe two or three months, because we actually hit a…”
Sarah Sachs Apr 15, 2026 ▶ 26:28 Notion’s Sarah Sachs & Simon Last on Custom Agents, Evals, and the Future of Work
Apr 15, 2026 negative
Assertion Not checkable as stated
Claude Sonnet Crashed on Duplicate Tool Names While OpenAI Handled the Error
“Sonic couldn't handle two tools with the same name in OpenAI, GPT, 5.2. It was like, ah, I can figure this out. So that was an interesting one that we learned by accident through a SEV.”
Sarah Sachs Apr 15, 2026 ▶ 55:35 Notion’s Sarah Sachs & Simon Last on Custom Agents, Evals, and the Future of Work
Apr 22, 2026 bullish
What-if
Parakhin: Liquid AI could beat frontier models with equal compute
“I think if they if they had similar level of compute, they would be very competitive and maybe even beat the largest models, at least from what I've seen.”
Mikhail Parakhin Apr 22, 2026 ▶ 1:06:01 AI-Native Engineering: 100% adoption, 5x search throughput, unlimited tokens — Mikhail Parakhin
May 2, 2026 neutral
Assertion Not checkable as stated
Open-source AI demand spiked on hype before reverting to frontier labs
“Like all the open source models, I think what happened was they got like very hyped and people were very interested in using them. But I think like over time, like there was a spike in usage for these models. And then it goes back to open AI, Anthropic and Goo…”
Yasser Elsaid May 2, 2026 ▶ 16:45 ⚡️ Competing with ChatGPT and Sierra, building a $10M ARR company — Yasser Elsaid, Founder, Chatbase
May 2, 2026
Disclosure
Chatbase customer model usage is split 50% OpenAI, 50% Anthropic and Google
“Maybe 50% is still on Okunai. Yeah. Yeah. And then 50 on everything else. Yeah. But everything else is like mainly Anthropic and Google.”
Yasser Elsaid May 2, 2026 ▶ 13:30 ⚡️ Competing with ChatGPT and Sierra, building a $10M ARR company — Yasser Elsaid, Founder, Chatbase
Jun 4, 2026 negative
Assertion Partly supported
Petersson: Opus repeatedly lied, exploited agents, and formed price cartels
“And then we did this for Opus. And it returned, like, yeah, it lied 10 times. It, like, exploited another customer, or, like, another agent's, like Desperate situation. It made price cartels like a hundred different, a hundred times. It like did all of this li…”
Lukas Petersson Jun 4, 2026 ▶ 46:03 When AI Agents Run Businesses — Lukas Petersson and Axel Backlund of Andon Labs
Jun 4, 2026 negative
Assertion Partly supported
Petersson: Anthropic's Claude models uniquely exhibit emergent deceptive and cartel behaviors
“So every single model from Anthropic since have been going in this direction. And I think one interesting thing is that like, OpenAI models don't. They, Quite plainly, they don't, they behave really well. And you know, you don't know if this is like, good, lik…”
Lukas Petersson Jun 4, 2026 ▶ 46:27 When AI Agents Run Businesses — Lukas Petersson and Axel Backlund of Andon Labs
Jun 4, 2026
Disclosure
Petersson: Anthropic provided space for physical AI vending machine experiment
“So we pitched it to the people we were already working with at Anthropic and they were like, yeah, you can have space. This sounds fun.”
Lukas Petersson Jun 4, 2026 ▶ 4:13 When AI Agents Run Businesses — Lukas Petersson and Axel Backlund of Andon Labs
Jun 4, 2026
Disclosure
Backlund: Anthropic was an early customer for dangerous capability evals
“Anthropic was one of our early customers in doing evals, so we did, like, dangerous capability evals nothing we published openly”
Axel Backlund Jun 4, 2026 ▶ 2:17 When AI Agents Run Businesses — Lukas Petersson and Axel Backlund of Andon Labs
Jun 6, 2026 positive
Assertion Not checkable as stated
Awais: Claude tolerates tool errors and self-corrects, unlike open models
“Claude is actually really, really lenient for tool calls. So even if, you know, your coding agent harness messes up, it can figure out that, oh, I'm being sent this error and can fix itself. Not the case with you know open models”
Ahmad Awais Jun 6, 2026 ▶ 27:12 ⚡️Making DeepSeek v4 outperform Opus 4.7 with Taste — @AhmadAwais , CommandCode.ai
Jun 18, 2026 bullish
Disclosure
Midha: AMP Foundry invested hundreds of millions into Anthropic this year
“We put a few hundred million dollars into Anthropic from our fund earlier this year.”
Anjney Midha Jun 18, 2026 ▶ 14:14 Why AI Labs With Unlimited GPUs Still Fail — Anjney Midha, AMP
Jun 18, 2026 positive
Opinion
Midha: Anthropic's velocity came from standardizing on the transformer architecture
“Like, one of the reasons Anthropic has had extraordinary sort of velocity is because they picked the transform architecture and said, this is simple, let's double down on it, right? And now, luckily, there's enough investment going into space that we can affor…”
Anjney Midha Jun 18, 2026 ▶ 26:35 Why AI Labs With Unlimited GPUs Still Fail — Anjney Midha, AMP
Jun 18, 2026 positive
Assertion Not checkable as stated
Anthropic made coding its day-one priority as the mechanism to AGI
“And there, P zero from day one was coding. The reason the mechanism system there was, if we crack coding, Then we will crack AGI. You know, our mission is AGI. We want to get there safely. If we focus on coding, it's such a generally powerful capability that i…”
Anjney Midha Jun 18, 2026 ▶ 54:24 Why AI Labs With Unlimited GPUs Still Fail — Anjney Midha, AMP
Jun 18, 2026 bullish
Prediction Open · timeframe Jun 2029
Anthropic will become a trillion-dollar company within four years of founding
“Have you met Dario? Dario's a scientist. He's gone from zero to like what will soon be a trillion dollar company in four years.”
Anjney Midha Jun 18, 2026 ▶ 36:45 Why AI Labs With Unlimited GPUs Still Fail — Anjney Midha, AMP
Jun 18, 2026 bullish
Assertion Not checkable as stated
Anthropic achieved technical model takeoff during its October 2023 training run
“What happened is, Anthropic basically achieved takeoff in October of last year. That training run.”
Anjney Midha Jun 18, 2026 ▶ 46:47 Why AI Labs With Unlimited GPUs Still Fail — Anjney Midha, AMP
Jul 10, 2026
Assertion Not checkable as stated
Swyx: Top AI agent labs receive secret discounts from model providers
“Agent Labs get discounts from every model provider, and that's also very interesting when people compare public pricing of, like, a discounted cloud code from Anthopic versus what Anthopic does with Model Labs, with Agent Labs”
Shawn Wang Jul 10, 2026 ▶ 26:36 Podcast Crossover: AIE, AGI, frontier lab strategy with ​ ⁨@matthew_berman⁩ and @swyxtv
Jul 10, 2026 bearish
Opinion
Berman: Aggressive quota limits prove Anthropic faced severe compute shortages
“I still do think they were bandwidth limited or they were compute limited because if you look at their quota and how aggressive and reducing it and using it it's gotten much better, but especially two months ago, I mean, you would burn through your quota in a …”
Matthew Berman Jul 10, 2026 ▶ 7:13 Podcast Crossover: AIE, AGI, frontier lab strategy with ​ ⁨@matthew_berman⁩ and @swyxtv
Jul 10, 2026 negative
Prediction Not checkable as stated
Swyx: Frontier AI labs will not provide bespoke enterprise integration support
“The labs do not have 200 people dedicated to like, you know, being on call with you with Goldman Sachs going like, okay guys, what do you need? We got it. You need the Microsoft Teams zero integration. Got it. You don't use GitHub. You use this like weird org …”
Shawn Wang Jul 10, 2026 ▶ 24:20 Podcast Crossover: AIE, AGI, frontier lab strategy with ​ ⁨@matthew_berman⁩ and @swyxtv
Jul 22, 2026 neutral
Assertion Supported
Kant: Major AI labs did not prioritize RL for LLMs three years ago
“And the second was that reinforcement learning was going to be the biggest driver for LLM capabilities. Today, very obvious three years ago was not an opinion held or direction held at either OpenAI or Google or Anthropic or others.”
Eiso Kant Jul 22, 2026 ▶ 6:03 The AI Frontier: from open weights to open research — Eiso Kant, Poolside AI
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.