Evaluations For Language Models
topic on 1 show · 1 statements across 1 episodes
1 statements about Evaluations For Language Models, every show
Sanyal: LLM evaluation does not require general-purpose reasoning models
“We feel like if you really focus on the fundamental problem that, hey, evaluations for language models is a very task-specific, constrained problem, and it does not require you to use you know, general purpose reasoning for that, and then there's fine-tuning y…”