Claude 3, every mention
56 scenes, the whole family · ← back to Claude 3
tap a year for its mentions
every year anyone Shawn Wang 26Rob Haisfield 9Ankur Goyal 9Alessio Fanelli 9Karan Malhotra 7Jason Liu 4Axel Backlund 3Will Brown 2Mike Krieger 2Damien Murphy 2
Verbatim, from the transcripts: the passages where Claude 3 comes up
Simulating Humanity: from Generative Agents to 8 Billion Digital Twins — Joon Sung Park, Simile AI
- ▶ 27:14 Alessio Fanelli Basically, if I was to do the same thing that you described with, say, your favorite LLM, Opus, GPT-Five-Six, have some agent to map out these things.
Podcast Crossover: AIE, AGI, frontier lab strategy with @matthew_berman and @swyxtv
- ▶ 5:40 Shawn Wang The, the, the downgrades to Opus.
The Agent Cloud: Databricks’ Bet on the Future of AI — Matei Zaharia and Reynold Xin
- ▶ 1:03:34 Matei Zaharia Uh, we have pipelines just using open source models, like the same model generates training environments and trains itself and beats like Opus and GPD-Five.Five and stuff at a task.
AI Security After Codex and Claude Code — Zico Kolter & Matt Fredrikson, Gray Swan
- ▶ 22:55 Matt Fredrikson While in these scenarios, humans found it very difficult to prompt inject, uh, the models, like we're aware of scenarios that a human would never fall for, that like Opus four seven would, right?
When AI Agents Run Businesses — Lukas Petersson and Axel Backlund of Andon Labs
- ▶ 47:42 Axel Backlund And like for Opus 4.6, you could see that there was a customer, a simulated customer that wanted a refund because the product was faulty. 3 times in the scene
Why Your AI Agents Don’t Work with Dex Horthy of HumanLayer | In-Context Cooking
- ▶ 13:53 Dex Horthy four point O and Opus 4.1.
Cursor's Third Era: Cloud Agents — ft. Sam Whitmore, Jonas Nelle, Cursor
- ▶ 9:19 Jonas Nelle I think we found in particular with Opus four, five, four, six and codex five, three, that those were additional step changes and sort of the autonomy grade capabilities of the model to just go off and figure out the details and come back…
- ▶ 38:56 unnamed speaker Sure, it's codex high today, but like, do I care if it's suddenly switched to Opus?
⚡️ Polsia: Solo Founder Tiny Team from 0 to 1m ARR in 1 month & the future of Self-Running Companies
- ▶ 36:41 unnamed speaker And then Opus.com came out, and I was like, okay, actually, now everything works, which was sort of my intuition is, like, models will get good enough that everything sort of works, right?
⚡️Jailbreaking AGI: Pliny the Liberator & John V on Red Teaming, BT6, and the Future of AI Security
- ▶ 16:45 Pliny the Liberator Yes, started to make some progress with some old templates, the old Godmode template from Opus three, um, and just sort of modify version because they had trained pretty heavily against that one.
⚡ Open Model Pretraining Masterclass — Elie Bakouch, HuggingFace SmolLM 3, FineWeb, FinePDF
- ▶ 9:43 Elie Bakouch I mean, I'm always curious to know, like, Claude, Claude Opus, like GPT-V, like what's the scale of those very big models?
⚡️Claude Sonnet 4.5 and Anthropic's roadmap for Agents and Developers — Mike Krieger, Anthropic
- ▶ 17:16 Mike Krieger And actually it was interesting is even when it was already outperforming Opus, for example, on sweet bench, people still didn't feel it was better, but then it continued to train and it was like now better than Opus and people don't want… 2 times in the scene
⚡️Composio: 10,000+ tools that evolve for Agents — Karan Vaidya and Soham Ganatra
- ▶ 20:51 Karan Vaidya So kind of like, I think I really love Claude Sonet, for example, because of the same, and Gropfor had like some of a mix of Opus and Sonet, which I really liked.
⚡️Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect
- ▶ 2:37 unnamed speaker I think in the Cloud Tree release, it was like extended thinking, kind of like, uh, capitalized and not just like extended thinking with tool use.
- ▶ 12:20 Shawn Wang I think the elephant in the room, let's talk about it, is this controversy around Opus or Opus, right?
- ▶ 17:15 Will Brown I mean, I think coding with these models, especially like quad three, I did a fair amount, like for a few weeks, I was doing a lot of quad code with three seven, mostly
- ▶ 22:02 Will Brown Many of them, like, really loved, like, Claude III Opus.
Claude Code: Anthropic's CLI Agent
- ▶ 1:13:50 Shawn Wang and I don't think this is obvious one year ago today, like when cloud three launched, it was just,
Why Every Agent needs Open Source Cloud Sandboxes
- ▶ 7:32 Shawn Wang so when I built it, it was with, uh, Claude III, the new Claude III launch. 3 times in the scene
Claude Plays Pokémon Hackathon: Escape from Mt. Moon!
- ▶ 7:10 David Hershey Sonnet three point
The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
- ▶ 11:31 Shawn Wang I think the other big project that you were, you were involved with was just cloud three. 4 times in the scene
- ▶ 11:44 Shawn Wang This is Haiku, Sonnet, Opus, all at once, right? 2 times in the scene
- ▶ 11:44 Shawn Wang This is Haiku, Sonnet, Opus, all at once, right?
- ▶ 11:44 Shawn Wang This is Haiku, Sonnet, Opus, all at once, right?
- ▶ 21:57 Shawn Wang Um, one last thing on Cloud Three. 5 times in the scene
- ▶ 56:03 Karina Nguyen Like, do you think the progress of, like, mini models, like, oh, three mini, like, oh, one mini, I guess, like, it came back to, like, the cloud, cloud three haiku, cloud 1.2 instant, like, this, like, gradual progression of, like, small…
2024 Year in Review: The Big Scaling Debate, the Four Wars of AI, Top Themes and the Rise of Agents
- ▶ 11:15 Shawn Wang And then Claude III launched mid, middle of this year. 2 times in the scene
- ▶ 1:16:15 Shawn Wang Um, I started tracking it, I think roundabout in March of 2024 with, uh, Haiku's launch. 3 times in the scene
- ▶ 1:17:15 Shawn Wang Um, but more, more importantly, I think, uh, you can see the more recent launches like Cloud Three Opus, which launched in March this year, um, now basically superseded completely, completely dominated by Gemini 1.5 pro, which is both…
- ▶ 1:25:10 Shawn Wang In March, Claude III came out, which huge, huge, huge for Enthopic.
Best of 2024 in Vision [LS Live @ NeurIPS]
- ▶ 54:58 unnamed speaker This is the year that vision language models became mainstream, with every model from GPT-Forty to One, to Claude Three, to Gemini One, and Two, to Llama 3.2, to Mistral's Pix-Trol, to AI-Two's Pixmo, going multimodal.
[Paper Club] DocETL: Agentic Query Rewriting + Eval for Complex Document Processing w Shreya Shankar
- ▶ 43:44 Eugene Yan And then essentially what they did was the ensemble command R, Haiku, and GPT-IV.
The new Claude 3.5 Sonnet, Computer Use, and Building SOTA Agents — with Erik Schluntz, Anthropic
- ▶ 47:13 Alessio Fanelli So you have Haiku, which is like, you know, cheaper. 3 times in the scene
- ▶ 49:09 Alessio Fanelli It's kind of like we left it at three back in the corner. 4 times in the scene
Production AI Engineering starts with Evals
- ▶ 1:24:59 Ankur Goyal So pre-Claude III, it was close to a hundred percent OpenAI. 3 times in the scene
- ▶ 1:25:04 Ankur Goyal Post-Claude III, uh, and I actually think Haiku is, is slept on a little bit, because before Foro Mini came out, Haiku was a very interesting reprieve, uh, for people to have very, very cheap. 6 times in the scene
- ▶ 1:25:15 unnamed speaker Are you talking about Sonnet or Haiku? 5 times in the scene
The Winds of AI Winter (Q2 Four Wars of the AI Stack Recap)
- ▶ 7:00 unnamed speaker Claw III Sonnet was interpreted and usefully improved using this technique.
- ▶ 10:33 Alessio Fanelli I think now for the first time, there's like a clear path to how do we make a seven B model good without having to go through GPT-IV or going to cloud three.
- ▶ 43:01 unnamed speaker And then Claude also did it with Opus and then with three fives on it, right?
- ▶ 1:10:23 unnamed speaker Haiku was, was the best cost per intelligence at that point in time.
[LLM Paper Club] Llama 3.1 Paper: The Llama Family of Models
- ▶ 42:25 unnamed speaker Like, if they, if they actually had anything else, any other trick that caused 3.5 Sonnet to be so good, uh, uh, it would, they probably would have deployed it on Haiku and Opus as well.
- ▶ 42:25 unnamed speaker Like, if they, if they actually had anything else, any other trick that caused 3.5 Sonnet to be so good, uh, uh, it would, they probably would have deployed it on Haiku and Opus as well.
- ▶ 42:57 unnamed speaker If you look at the training data and the training date for when cloud three Sonic came out to 3.5 on it, it also has a year and a half of significant data updates.
This World Does Not Exist — Joscha Bach, Karan Malhotra, Rob Haisfield (WorldSim, WebSim, Liquid AI)
- ▶ 7:50 Karan Malhotra So cloud three, as you guys know, is not a base model. 2 times in the scene
- ▶ 19:05 Karan Malhotra Now, this is just Cloud III. 2 times in the scene
- ▶ 21:54 Karan Malhotra Create Twitter account Claude three. 3 times in the scene
- ▶ 39:40 Rob Haisfield Uh, should I switch this one to Opus? 3 times in the scene
- ▶ 42:06 Rob Haisfield Right, so what I'd probably do there is I switch to haiku real quick. 2 times in the scene
- ▶ 45:45 Rob Haisfield Abstraction slash eigenstraction, and I'm just gonna switch to Sonnet for this, because Sonnet's actually, like, really good. 3 times in the scene
- ▶ 46:19 Rob Haisfield Opus can just handle much more complexity.
High Agency Pydantic over VC Backed Frameworks — with Jason Liu of Instructor
- ▶ 19:24 unnamed speaker And if you actually try and do vector similarity, it's not that good because the people that wrote the specs, they didn't have a mind making them like semantically apart, you know, they're kind of like, oh, create this, create this, create…
- ▶ 22:04 Jason Liu Haiku is, in function calling, it's actually better. 3 times in the scene
- ▶ 49:50 Jason Liu Or those are two sort of systems that I wish you before or Opus was actually good enough to just write me an essay, but most of the essays are still pretty bad.
Personal AI Meetup - Bee, BasedHardware, LangChain LangFriend, Deepgram EmilyAI
- ▶ 13:50 Damien Murphy Uh, Claude Haiku is surprisingly good, so, you know, if cost is something we want to get down, that's definitely an option. 2 times in the scene