AI alignment

3 statements across 2 episodes · 1 bullish · 0 bearish · 2 people on the record · first statement Oct 16, 2025 by Jerry Tworek · across every show →

Everything said about AI alignment, oldest first

Oct 16, 2025
Insight
Tworek: AI alignment is a never-ending pursuit as human goals evolve
“And it's I think it's a never ending pursuit because like, even, even for humans, it's not super easy to define what's, what do we consider a light? And I think as our civilization will evolve, it will, the notion of alignment and the goals of humanity will, K…”
Jerry Tworek Oct 16, 2025 ▶ 1:00:22 How GPT-5 Thinks — OpenAI VP of Research Jerry Tworek
Oct 23, 2025
Insight
Schrittwieser: AI safety must span the entire stack, not just RL
“Yeah, I wouldn't view it alignment adjust like an RL problem. I think it sort of, it goes throughout the whole stack. You might, you know, for example, filter the pre-training data in some way. You might, after training, you might have classifiers that, you kn…”
Julian Schrittwieser Oct 23, 2025 ▶ 1:03:10 Are We Misreading the AI Exponential? Julian Schrittwieser on Move 37 & Scaling RL (Anthropic)
Oct 23, 2025 positive
Insight
Pre-training aids AI alignment by implicitly instilling human values
“I definitely think we would keep using pre-training data, not just from an efficiency point of view as well, but also I think there is interesting safety angles, because by pre-training and, you know, all this human knowledge, we're implicitly creating an agen…”
Julian Schrittwieser Oct 23, 2025 ▶ 22:07 Are We Misreading the AI Exponential? Julian Schrittwieser on Move 37 & Scaling RL (Anthropic)
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.