Human Preference
topic on 2 shows · 2 statements across 2 episodes
2 statements about Human Preference, every show
Lambert: Code maintainability is a human preference problem in RL
“The software stuff is not easy because it's almost like maintainability almost feels like a human preference type issue again, where somebody could look at it and be like, yeah, that's not as good, but adding the heuristic and trading seems very messy.”
Doshi: AI Image Models Homogenize Styles by Overfitting to Average Human Preference
“I think that the models are probably too curated and maybe overfit to be based on human preference. And human preference isn't your human preference. It's your preference. I mean, it's some kind of average of human preference.”