Reward Problem
topic on 1 show · 1 statements across 1 episodes
1 statements about Reward Problem, every show
Laskin: Accurately verifying arbitrary outcomes is ASI-complete
“The reward problem in itself is at the time I called, I thought it was AGI complete. Now I'd say it's ASI complete, but by the time you have a neural network that can accurately verify any outcome, that is probably a super intelligence.”