AI alignment
3 statements across 3 episodes · 1 bullish · 0 bearish · 3 people on the record · first statement Sep 14, 2023 by Illia Polosukhin · across every show →
Everything said about AI alignment, oldest first
Sep 14, 2023 neutral
Polosukhin: Society needs human alignment rather than AI alignment
“So I have this view that we need human alignment instead of AI alignment. So right now, kind of when we talk about, you know, hey, we need to align AIs with like human values, but the reality is that, you know, all the problems that exist, they all exist becau…”
Aug 30, 2024 bullish
Steinberger: AI alignment is only solvable via recursive automated models
“The only way to sort of reasonably approach this is to iteratively ask your model to solve alignment and safety at that stage, not, not, you know, surely you can also ask it to solve your product level problems, but like that, that's nice, but that's not the f…”
Mar 5, 2025 neutral
Hendrycks: AI alignment is only a subset of AI safety
“So I view the distinction between alignment and safety as alignment as being a sort of subset of safety. Obviously you want the value systems of the AIs to be in keeping with or compatible with say the US public for USAIs or for you as an individual, but that …”