reinforcement fine-tuning
also referred to as: reinforcement fine tuning
1 statements across 1 episodes · 1 bullish · 0 bearish · 1 people on the record · first statement Aug 13, 2025 by Jacob Effron · across every show →
Everything said about reinforcement fine-tuning, oldest first
Aug 13, 2025 bullish
OpenAI Reinforcement Fine-Tuning Makes Domain-Specific Models Viable for Vertical Apps
“In the last few months, OpenAI shipped reinforcement fine tuning, which is kind of a new way of fine tuning that they offer. And it's actually quite good. And so it actually is starting to seem again being able at least to not pre-train, but fine tune data on …”