tree enforcement learning with human feedback

1 statements across 1 episodes · 1 bullish · 0 bearish · 1 people on the record · first statement Sep 25, 2023 by Mira Murati · across every show →

Everything said about tree enforcement learning with human feedback, oldest first

Sep 25, 2023 positive
Prediction Not checkable as stated
Murati: Scaling RLHF alone might be sufficient to solve AI hallucinations
“And we also wanted to figure out the issue of hallucinations, which is always an extremely hard problem. But I do think that with this method of tree enforcement learning with human feedback, maybe that is all we need if we push this hard enough.”
Mira Murati Sep 25, 2023 ▶ 16:16 Where We Go From Here with OpenAI's Mira Murati
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 1,000 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.