ARC-AGI-2, every mention
6 scenes · ← back to ARC-AGI-2
tap a year for its mentions
every year anyone Alessio Fanelli 7Greg Kamradt 4Ashvin Nair 1
Verbatim, from the transcripts: the passages where ARC-AGI-2 comes up
[State of RL/Reasoning] IMO/IOI Gold, OpenAI o3/GPT-5, and Cursor Composer — Ashvin Nair, Cursor
- ▶ 32:08 Ashvin Nair The anthropic, uh, models, like the, uh, Opus two, 4.5, it has this kind of like, uh, there's this like RKGI two plot that looks exactly like the API ones, right?
Terminal-Bench 2.0: the most impt coding agent benchmark of 2025 gets a v2! Launch + Q&A w/ founders
- ▶ 29:34 unnamed speaker I just, I can't be, um, you know, I can't have Greg here and not make comparisons with terminal bench two and RKGI two and like that developments, you know, something that, uh, Greg is sort of pushing is.
⚡️ARC-AGI-3: The Interactive Reasoning Benchmark
- ▶ 17:03 Alessio Fanelli I know in the B two, you also had the dollar per task thing. 2 times in the scene
- ▶ 25:39 Greg Kamradt So the way I think about it is if you had like a linear spectrum, RKGI one and two, it's a static list of like three or four JSON grids static. 2 times in the scene
- ▶ 27:10 Alessio Fanelli Uh, timeline on when you think B two will get close to like 50%, then a hundred percent, because I know graph for yesterday really 16%, which everybody was going crazy over, but it's still 16%. 5 times in the scene
- ▶ 29:19 Greg Kamradt Because our hypothesis is the thing that actually does beat Arc AGI too. 2 times in the scene