Opus 5
product on 2 shows · 7 statements across 2 episodes
the Y Combinator Startup Podcast
TBPN
7 statements about Opus 5, every show
Coogan: Claude Fable 5.1 achieved highest-ever AI index score
“Fable 5.1 get the highest result ever on the index. The scores also have Opus five, which got 63, and Fable, which got a 62.”
Hu: Anthropic Opus 5 achieved 30% on ARC-AGI
“You guys got, took Arc AGI three to 30%, which is incredible.”
Cherny: Claude Opus 5 Can Run Autonomously for Months Without Scaffolding
“For five, one example of something it does that I think no other model has done is it runs for a very long period of time. And especially when you combine Opus Five with auto mode, it's just like incredible. Like it can go for days, weeks, months at a time. It…”
Cherny: Anthropic Can No Longer Demonstrate Prompt Injection on Opus 5
“So essentially, if you combine a well-aligned model, so this is, like, essentially three years of research into alignment, With a prompt injection classifier, which we run for all traffic, and what this is doing is it's based on Chrysola's mechanistic interpre…”
Anthropic Deleted 80% of Claude Code System Prompt for Opus 5
“So yeah, we deleted 80% of the system prompt.”
Cherny: Claude Code Users Should Delete CLAUDE.md Scaffolding Every Six Months
“Every six months, delete your Cloud MD. Delete your skills. Delete your hooks. See what the model does, and it might surprise you. And actually for Opus Five, this is something we really do recommend, is just try deleting all of these things, because the model…”
Cherny: Claude still struggles with systems code, distributed systems, and UI verification
“So coding is solved for the kind of coding that I do. It's not solved for everyone. You know, there's still code bases that are like super deep systems code bases where quad still struggles. There's distributed systems where quad still struggles. There's reall…”