Claude 3 Opus, every mention
24 scenes · ← back to Claude 3 Opus
tap a year for its mentions
every year anyone Alessio Fanelli 5Shawn Wang 4Rob Haisfield 4Axel Backlund 3Mike Krieger 2Will Brown 1Pliny the Liberator 1Matt Fredrikson 1Matei Zaharia 1Karan Vaidya 1
Verbatim, from the transcripts: the passages where Claude 3 Opus comes up
Simulating Humanity: from Generative Agents to 8 Billion Digital Twins — Joon Sung Park, Simile AI
- ▶ 27:14 Alessio Fanelli Basically, if I was to do the same thing that you described with, say, your favorite LLM, Opus, GPT-Five-Six, have some agent to map out these things.
Podcast Crossover: AIE, AGI, frontier lab strategy with @matthew_berman and @swyxtv
- ▶ 5:40 Shawn Wang The, the, the downgrades to Opus.
The Agent Cloud: Databricks’ Bet on the Future of AI — Matei Zaharia and Reynold Xin
- ▶ 1:03:34 Matei Zaharia Uh, we have pipelines just using open source models, like the same model generates training environments and trains itself and beats like Opus and GPD-Five.Five and stuff at a task.
AI Security After Codex and Claude Code — Zico Kolter & Matt Fredrikson, Gray Swan
- ▶ 22:55 Matt Fredrikson While in these scenarios, humans found it very difficult to prompt inject, uh, the models, like we're aware of scenarios that a human would never fall for, that like Opus four seven would, right?
When AI Agents Run Businesses — Lukas Petersson and Axel Backlund of Andon Labs
- ▶ 47:42 Axel Backlund And like for Opus 4.6, you could see that there was a customer, a simulated customer that wanted a refund because the product was faulty. 3 times in the scene
Why Your AI Agents Don’t Work with Dex Horthy of HumanLayer | In-Context Cooking
- ▶ 13:53 Dex Horthy four point O and Opus 4.1.
Cursor's Third Era: Cloud Agents — ft. Sam Whitmore, Jonas Nelle, Cursor
- ▶ 9:19 Jonas Nelle I think we found in particular with Opus four, five, four, six and codex five, three, that those were additional step changes and sort of the autonomy grade capabilities of the model to just go off and figure out the details and come back…
- ▶ 38:56 unnamed speaker Sure, it's codex high today, but like, do I care if it's suddenly switched to Opus?
⚡️ Polsia: Solo Founder Tiny Team from 0 to 1m ARR in 1 month & the future of Self-Running Companies
- ▶ 36:41 unnamed speaker And then Opus.com came out, and I was like, okay, actually, now everything works, which was sort of my intuition is, like, models will get good enough that everything sort of works, right?
⚡️Jailbreaking AGI: Pliny the Liberator & John V on Red Teaming, BT6, and the Future of AI Security
- ▶ 16:45 Pliny the Liberator Yes, started to make some progress with some old templates, the old Godmode template from Opus three, um, and just sort of modify version because they had trained pretty heavily against that one.
⚡ Open Model Pretraining Masterclass — Elie Bakouch, HuggingFace SmolLM 3, FineWeb, FinePDF
- ▶ 9:43 Elie Bakouch I mean, I'm always curious to know, like, Claude, Claude Opus, like GPT-V, like what's the scale of those very big models?
⚡️Claude Sonnet 4.5 and Anthropic's roadmap for Agents and Developers — Mike Krieger, Anthropic
- ▶ 17:16 Mike Krieger And actually it was interesting is even when it was already outperforming Opus, for example, on sweet bench, people still didn't feel it was better, but then it continued to train and it was like now better than Opus and people don't want… 2 times in the scene
⚡️Composio: 10,000+ tools that evolve for Agents — Karan Vaidya and Soham Ganatra
- ▶ 20:51 Karan Vaidya So kind of like, I think I really love Claude Sonet, for example, because of the same, and Gropfor had like some of a mix of Opus and Sonet, which I really liked.
⚡️Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect
- ▶ 12:20 Shawn Wang I think the elephant in the room, let's talk about it, is this controversy around Opus or Opus, right?
- ▶ 22:02 Will Brown Many of them, like, really loved, like, Claude III Opus.
The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
- ▶ 11:44 Shawn Wang This is Haiku, Sonnet, Opus, all at once, right?
2024 Year in Review: The Big Scaling Debate, the Four Wars of AI, Top Themes and the Rise of Agents
- ▶ 1:17:15 Shawn Wang Um, but more, more importantly, I think, uh, you can see the more recent launches like Cloud Three Opus, which launched in March this year, um, now basically superseded completely, completely dominated by Gemini 1.5 pro, which is both…
The new Claude 3.5 Sonnet, Computer Use, and Building SOTA Agents — with Erik Schluntz, Anthropic
- ▶ 49:09 Alessio Fanelli It's kind of like we left it at three back in the corner. 4 times in the scene
The Winds of AI Winter (Q2 Four Wars of the AI Stack Recap)
- ▶ 43:01 unnamed speaker And then Claude also did it with Opus and then with three fives on it, right?
[LLM Paper Club] Llama 3.1 Paper: The Llama Family of Models
- ▶ 42:25 unnamed speaker Like, if they, if they actually had anything else, any other trick that caused 3.5 Sonnet to be so good, uh, uh, it would, they probably would have deployed it on Haiku and Opus as well.
This World Does Not Exist — Joscha Bach, Karan Malhotra, Rob Haisfield (WorldSim, WebSim, Liquid AI)
- ▶ 39:40 Rob Haisfield Uh, should I switch this one to Opus? 3 times in the scene
- ▶ 46:19 Rob Haisfield Opus can just handle much more complexity.
High Agency Pydantic over VC Backed Frameworks — with Jason Liu of Instructor
- ▶ 19:24 unnamed speaker And if you actually try and do vector similarity, it's not that good because the people that wrote the specs, they didn't have a mind making them like semantically apart, you know, they're kind of like, oh, create this, create this, create…
- ▶ 49:50 Jason Liu Or those are two sort of systems that I wish you before or Opus was actually good enough to just write me an essay, but most of the essays are still pretty bad.