Assertion certainty 4/5 debate potential 3/5

Chollet: GPT-4 lacks fluid intelligence, but OpenAI's o3 model has it

Francois Chollet · Chasing Real AGI: Inside ARC Prize 2025 with Chollet & Knoop · Apr 3, 2025 · at 5:00

François Chollet, creator of the ARC-AGI benchmark and co-founder of NDEA, discusses how different AI model architectures handle fluid intelligence versus memorized skills.

0:00 / 0:04exact quote · 4.2s
▶ Watch the full episode on YouTube → 720p mp4 · rendered on demand · StarZero watermark
“GPT-IV does not have fluid intelligence, for instance, but O-III does.”

quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →

More from Francois Chollet

Assertion Supported
Chollet: 50,000x LLM scaling yielded flat progress on ARC benchmark
“Because between, like, GPT-II and GPT-IV. There's been this 50,000 X scale up of base models that has resulted in, in, in basically a flat curve. On Arc.”
Francois Chollet Apr 3, 2025 ▶ 11:44 Chasing Real AGI: Inside ARC Prize 2025 with Chollet & Knoop
Prediction Not checkable as stated
Chollet: Commercial AI models will increasingly adopt test-time search architectures
“Increasingly, you're gonna see commercial models that use test-time search, where instead of just trying to generate one single COT to adapt to the task, they're actually gonna run through this, you know, search.”
Francois Chollet Apr 3, 2025 ▶ 9:47 Chasing Real AGI: Inside ARC Prize 2025 with Chollet & Knoop
Assertion Supported
Chollet: Latest base LLMs score zero percent on ARC-AGI-2
“Today the latest base alarms, they're doing something like 10% on ARK-I. But on Arc two, they are doing zero percent.”
Francois Chollet Apr 3, 2025 ▶ 10:50 Chasing Real AGI: Inside ARC Prize 2025 with Chollet & Knoop
Insight
Chollet: Intelligence should be defined as skill acquisition efficiency
“And yeah, so I, to summarize that, you know, I see intelligence as skill acquisition efficiency. So it's not the fact that you can acquire skills, it's how efficiently You can do it. That's a measure of your intelligence.”
Francois Chollet Apr 3, 2025 ▶ 16:06 Chasing Real AGI: Inside ARC Prize 2025 with Chollet & Knoop
Opinion
Chollet: OpenAI o3 is the most advanced test-time adaptation model
“And OSTRI best I can tell is the most advanced the most successful test and adaptation model out there at this time.”
Francois Chollet Apr 3, 2025 ▶ 17:53 Chasing Real AGI: Inside ARC Prize 2025 with Chollet & Knoop
Insight
Chollet: AGI means humans can no longer easily create tasks AI fails
“You have AGI when it's no longer possible to easily come up with tasks that, you know, you and I can do naturally, but no AI system can do.”
Francois Chollet Apr 3, 2025 ▶ 49:37 Chasing Real AGI: Inside ARC Prize 2025 with Chollet & Knoop
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.