HealthBench

1 statements across 1 episodes · 1 bullish · 0 bearish · 1 people on the record · first statement Aug 13, 2025 by Justine Moore · said 1 times in 1 episodes since 2025 · across every show →

Mentions by year

brought up most by Justine Moore (1)

tap a year for its mentions
0011112025episodesmentions
0112025episodes it came up in
000.50.5112025episodesmentions per episode

every mention, scene by scene, with the transcript →

Everything said about HealthBench, oldest first

Aug 13, 2025 positive
Assertion Supported
OpenAI's GPT-5 scored highest on the physician-trained HealthBench medical benchmark
“They talked about how GPT-V was kind of the highest scoring model on this thing called HealthBench, which is a benchmark they trained with like, 250 plus physicians. To measure how good an LLM is at answering medical questions.”
Justine Moore Aug 13, 2025 ▶ 11:44 This Week in AI: GPT-5 Ships, 4o Pulled Back, Grok Imagine Goes Social
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 1,000 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.