Foody: AI labs are broadly shifting from RLHF to RLAIF
Brendan Foody · Why experts writing AI evals is creating the fastest-growing companies in history | Brendan Foody · Sep 18, 2025 · at 15:44
Brendan Foody, CEO of Mercor, explains why leading AI labs are shifting their post-training and evaluation methodologies toward AI feedback.
“What everyone is generally moving towards is reinforcement learning from AI feedback instead of human feedback, where you have instead the human defined some sort of success criteria, some way to measure that... And it's far more scalable and data efficient, and so that's why a lot of, you know, the broader trend in the market across the board is moving towards RLA-IF to both eval models as well as improved capabilities.”
quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →