reward model loss function

1 statements across 1 episodes · 0 bullish · 0 bearish · 1 people on the record · first statement Jan 11, 2024 by Nathan Lambert · across every show →

Everything said about reward model loss function, oldest first

Jan 11, 2024 neutral
Assertion Supported
Lambert: Anthropic and OpenAI reward model loss functions are mathematically identical
“Fun fact is that these loss functions Look different and anthropic in opening eyes papers, but they're just literally just log transform. So if you start like expantiating both sides and taking the log of both sides, you'll like converge on one of the two, the…”
Nathan Lambert Jan 11, 2024 ▶ 54:41 The Origin and Future of RLHF: the secret ingredient for ChatGPT - with Nathan Lambert
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.