Kaggle

5 statements across 3 episodes · 3 bullish · 1 bearish · 3 people on the record · first statement Aug 17, 2024 by Jeremy Howard · said 37 times in 7 episodes since 2023 · across every show →

Mentions by year

brought up most by Jesse Hu (16), Jeremy Howard (7), Alessio Fanelli (6), Ari Morcos (1)

tap a year for its mentions
001322542023202420252026episodesmentions
0242023202420252026episodes it came up in
0042842023202420252026episodesmentions per episode
2026 6 mentions in 1 episode
2025 1 mention in 1 episode
2024 23 mentions in 4 episodes 6 per episode
2023 7 mentions in 1 episode

every mention, scene by scene, with the transcript →

Everything said about Kaggle, oldest first

Aug 17, 2024 bearish
Assertion Not checkable as stated
Howard: Decoder models must be far larger to match DeBERTa
“Now, the interesting thing is, you see, unlike Kaggle competitions, that decoder models still Are at least competitive with things like DiBerta VIII. But they have to be way bigger to be competitive with things like DiBerta VIII. And the only reason they are c…”
Jeremy Howard Aug 17, 2024 ▶ 35:33 Answer.ai & AI Magic with Jeremy Howard
Oct 19, 2024 positive
Assertion Supported
Hu: OpenAI o1-preview achieves bronze medals in 17% of MLE-bench competitions
“Their final results with a one preview and this a scaffolding from a different company was that they got a bronze medal. I don't think I've ever achieved once but I haven't competed that much in. 17% of competitions.”
Jesse Hu Oct 19, 2024 ▶ 45:06 [Paper Club] SWE-Bench [OpenAI Verified/Multimodal] + MLE-Bench with Jesse Hu
Oct 19, 2024 neutral
Assertion Supported
Hu: MLE-bench authors found obfuscating competition details did not show overfitting
“They do a lot of checks against overfitting on the Kaggle tasks themselves, and so they do something where they obfuscate some of the details of the Of the competitions, and then they rerun it. And I guess if they were overfitting on the competitions themselve…”
Jesse Hu Oct 19, 2024 ▶ 52:19 [Paper Club] SWE-Bench [OpenAI Verified/Multimodal] + MLE-Bench with Jesse Hu
Oct 19, 2024 positive
Assertion Supported
Hu: OpenAI o1-preview surpasses human Kaggle Grandmasters with seven gold medals
“Since a grandmaster requires five gold medals and oh, and preview gets an average of eight or sorry, seven gold medals. They're out competing even capital grandmasters.”
Jesse Hu Oct 19, 2024 ▶ 47:29 [Paper Club] SWE-Bench [OpenAI Verified/Multimodal] + MLE-Bench with Jesse Hu
May 24, 2026 positive
Assertion Supported
Sanseviero: Kaggle launched an exam-based benchmark leaderboard for AI agents
“Last week, they released a new system for agent evaluation. It's like a very, like, experimental initial benchmark, but pretty much allowing agents to take an exam and compete in a leaderboard, which is always fun.”
Omar Sanseviero May 24, 2026 ▶ 28:31 ⚡️ Google's Open AI Strategy — Omar Sanseviero, Google DeepMind
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.