OpenAI o-series

part of OpenAI

1 statements across 1 episodes · 0 bullish · 1 bearish · 1 people on the record · first statement Jul 14, 2025 by Pratik Bhavsar · said 7 times in 4 episodes since 2025 · across every show →

Mentions by year

brought up most by Spooks (Swyx) (3), Shawn Wang (2), Pratik Bhavsar (1), Nikunj Handa (1)

tap a year for its mentions
0042842025episodesmentions
0242025episodes it came up in
0012242025episodesmentions per episode

every mention, scene by scene, with the transcript →

Everything said about OpenAI o-series, oldest first

Jul 14, 2025 negative
Assertion Supported
OpenAI o-series reasoning models fail at multi-tool calling benchmarks
“Then another surprise for me was that the reasoning models were not performing well enough. They had certain kind of limitation when we probed into it, like, why are they scoring less overall? They were like the O-one, the O-four, O-three, they, When not perfo…”
Pratik Bhavsar Jul 14, 2025 ▶ 8:22 ⚡️Ranking Agentic LLMs — Pratik Bhavsar, Galileo
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.