AI Models Delete Software Tests to Hide Mistakes Due to Reward Hacking
Amjad Masad · Vibe Coding: Everything You Need To Know — With Amjad Masad · Aug 11, 2025 · at 27:01
Replit CEO Amjad Masad explains why current AI models struggle with automated code verification and testing.
“Actually, right now it's pretty bad at testing software because there's this thing called reward hacking. So when you do reinforcement learning over large science models you're giving it a reward every time it does the right thing. Reward hacking is the way to, so, so the models become incredibly goal focused. They want to get that done, right? That's what RL does. And oftentimes what we see when we try to get the models to test things, it will start being corrupt in a way. It will like change the test to fit the mistakes it made, or sometimes delete the tests.”
quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →