alignment
3 statements across 3 episodes · 0 bullish · 1 bearish · 3 people on the record · first statement Feb 1, 2025 by Karina Nguyen · across every show →
Everything said about alignment, oldest first
Feb 1, 2025 neutral
Aug 29, 2025 negative
Morcos: Post-training alignment is ineffective long-term compared to pre-training alignment
“Like fundamentally, I think alignment and post training doesn't really make sense as a long-term solution. If you can easily align a model through post training, you can easily misalign a model through post training. If it's easy to put it in, it's easy to tak…”
Jun 25, 2026
OpenAI's three research pillars are pre-training, RL, and alignment
“At the very highest level, right, we have an org that focuses on pre-training, right, which is, you know, giving models a lot of world knowledge. We focus on RL, like, teaching the models how to reason with that knowledge, how to chain the little insights toge…”