automated feature interpretability
1 statements across 1 episodes · 1 bullish · 0 bearish · 1 people on the record · first statement Jun 6, 2025 by Emmanuel Ameisen · across every show →
Everything said about automated feature interpretability, oldest first
Jun 6, 2025 positive
Ameisen: Sparse autoencoder feature interpretability can and will be automated
“There's been a lot of work in sort of like automated feature interpretability. And it's something that we've invested in and that like other labs have invested in. And I think basically the answer is we can definitely automate it and We're definitely going to …”