RLHF
1 statements across 1 episodes · 1 bullish · 0 bearish · 1 people on the record · first statement Oct 10, 2025 by Jonathan Siddharth · across every show →
Everything said about RLHF, oldest first
Oct 10, 2025 positive
Siddharth: Verifiable domains allow self-play reinforcement learning to replace RLHF
“Now, for these verifiable domains like coding and math, instead of doing reinforcement learning with human feedback, you can do reinforcement learning. Because you can automatically check when you got the correct answer or not in these verifiable domains. And …”