Generative Model Performance

topic on 1 show · 1 statements across 1 episodes

Latent Space

1 statements about Generative Model Performance, every show

Angelopoulos: Static benchmarks are intrinsically unable to evaluate generative models
“Static benchmarks are intrinsically, to some extent, unable to measure generative model performance. And the reason is because you cannot Pre-annotate all the outputs of a generative model. You change the model. It's like the distribution of your data is chang…”
Anastasios Angelopoulos Nov 1, 2024 ▶ 6:40 In the Arena: How LMSys changed LLM Benchmarking Forever

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.