reinforcement fine-tuning

also referred to as: reinforcement fine tuning

2 statements across 2 episodes · 1 bullish · 0 bearish · 2 people on the record · first statement Apr 10, 2025 by Brendan Foody · across every show →

Everything said about reinforcement fine-tuning, oldest first

Apr 10, 2025 bullish
Opinion
Foody: Reinforcement fine-tuning makes application-layer AI customization viable
“The reason I'm so optimistic about it taking off is that it's, like, profoundly data efficient, right? And it finally makes sense to customize models at the application layer.”
Brendan Foody Apr 10, 2025 ▶ 32:40 No Priors Ep. 110 | With Mercor CEO and Co-Founder Brendan Foody
Apr 24, 2025
Insight
Fulford: RFT is only worthwhile for out-of-distribution or make-or-break tasks
“I think if you have a very specific task that you think is so different to anything that the model was likely trained on and you try it a bunch of times yourself and you've tried a lot of different prompts and it's just really not good at it. So maybe it's gen…”
Isa Fulford Apr 24, 2025 ▶ 7:50 No Priors Ep. 112 | With OpenAI Deep Research, Isa Fulford
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 100 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.