Model Interpretability
topic on 2 shows · 2 statements across 2 episodes
2 statements about Model Interpretability, every show
Bissell: AI safety research must scale with superintelligence or fight losing battle
“Ideally, you are setting up your research so that as super intelligence arrives, that is a tailwind. That's also bolstering our ability to like understand the models because otherwise you're fighting a losing battle. If it's like the systems are getting more a…”