AI Alignment

topic on 11 shows · 20 statements across 17 episodes

the Y Combinator Startup Podcast American Optimist Latent Space Lenny's Podcast No Priors WTF is with Nikhil Kamath the MAD Podcast the a16z Podcast Big Technology TBPN 20VC

20 statements about AI Alignment, every show

Bostrom: AI misuse is a governance challenge, not a technical one
“You're focusing there on the misuse potential that this people might choose to do bad things with AI technology. And that certainly is one big category of risk, right? But that's not primarily a technical challenge. It's more ultimately a governance challenge …”
Nick Bostrom Aug 19, 2026 ▶ 15:03 Nick Bostrom: Worries About AI Existential Risk Just Became More Concrete
BIG TECHNOLOGY Prediction Not checkable as stated
Bostrom: Weak aligned superintelligence could help align stronger superintelligence
“If you get a kind of weak super intelligence that is For the most part aligned, we might then be able to use that to make a more powerful form of super intelligence that is more reliably aligned.”
Nick Bostrom Aug 19, 2026 ▶ 19:02 Nick Bostrom: Worries About AI Existential Risk Just Became More Concrete
Bostrom: Natural language AI interfaces provide a safer alignment buffer before superintelligence
“This gives us more sort of surface area to work with. Like you can more easily understand and interact with these systems because they have human double concepts and you can talk with them.”
Nick Bostrom Aug 19, 2026 ▶ 42:45 Nick Bostrom: Worries About AI Existential Risk Just Became More Concrete
Ries: Human alignment, not technical alignment, is AI's top unsolved problem
“This is the number one unsolved problem in AI. It's not the tech, we're making great progress on the technical alignment problem, but we haven't made jack progress on the human alignment problem, which is that we've known since the development of Conway's laws…”
Eric Ries May 10, 2026 ▶ 1:30:31 How Anthropic, Costco, and Patagonia all build incorruptible companies | Eric Ries
20VC Insight
Midha: Human alignment is a bigger challenge than technical AI alignment
“AI alignment, don't get me wrong, is hard, but not the hardest problem. Human alignment is really the problem right now.”
Anjney Midha Apr 14, 2026 ▶ 0:00 The Early Days of Anthropic & How 21 of 22 VCs Rejected It | The Four Bottlenecks in AI | Anj Midha · 20VC with Harry Stebbings
a16z Insight
Emmett Shear: Goal-based alignment covers only a tiny fraction of human experience
“Goals are one level of alignment. You can align something around goals. The kind of goals we're talking about here are one level of alignment. You can align something around goals by like if you can explicitly articulate in concept and in description, the stat…”
Emmett Shear Nov 17, 2025 ▶ 21:50 Emmett Shear on Building AI That Actually Cares: Beyond Control and Steering
TBPN Opinion
Johnson: Project Blueprint is fundamentally about AI alignment
“What a lot of people don't realize is this endeavor is entirely about AI alignment.”
Bryan Johnson Oct 30, 2025 ▶ 4:41 Don't Die CEO Bryan Johnson on Living Forever, AI & His $60M Blueprint For Immortality
MAD Insight
Pre-training aids AI alignment by implicitly instilling human values
“I definitely think we would keep using pre-training data, not just from an efficiency point of view as well, but also I think there is interesting safety angles, because by pre-training and, you know, all this human knowledge, we're implicitly creating an agen…”
Julian Schrittwieser Oct 23, 2025 ▶ 22:07 Are We Misreading the AI Exponential? Julian Schrittwieser on Move 37 & Scaling RL (Anthropic)
MAD Insight
Schrittwieser: AI safety must span the entire stack, not just RL
“Yeah, I wouldn't view it alignment adjust like an RL problem. I think it sort of, it goes throughout the whole stack. You might, you know, for example, filter the pre-training data in some way. You might, after training, you might have classifiers that, you kn…”
Julian Schrittwieser Oct 23, 2025 ▶ 1:03:10 Are We Misreading the AI Exponential? Julian Schrittwieser on Move 37 & Scaling RL (Anthropic)
MAD Insight
Tworek: AI alignment is a never-ending pursuit as human goals evolve
“And it's I think it's a never ending pursuit because like, even, even for humans, it's not super easy to define what's, what do we consider a light? And I think as our civilization will evolve, it will, the notion of alignment and the goals of humanity will, K…”
Jerry Tworek Oct 16, 2025 ▶ 1:00:22 How GPT-5 Thinks — OpenAI VP of Research Jerry Tworek
Fisher: Economic pressure for long-horizon agents will drive AI alignment progress
“I'm actually really positive and bullish that there is this economic pressure in a good way To make progress on alignment because long horizon agents require it.”
Jordan Fisher Oct 7, 2025 ▶ 18:19 Ask These Questions Before Starting An AI Startup · Y Combinator
Y COMBINATOR Prediction Not checkable as stated
Anthropic's Joseph: Certain AI Alignment Pieces Will Move to Pre-Training
“I do think at some point there will be, like, some pieces of alignment that, like, you do want to export back into pre-training because that might be a way to, like, Put them in with more strength, like, more robustness, kind of, or more core to the intelligen…”
Nick Joseph Sep 30, 2025 ▶ 46:03 Anthropic Head of Pretraining on Scaling Laws, Compute, and the Future of AI · Y Combinator
Mann: Language Models Understand Human Values in a Core Way
“And since then, my estimation of how hard the problem would be has gone down significantly actually because things like language models actually do really understand human values in a core way. The problem is definitely not solved, but I'm more hopeful than I …”
Ben Mann Jul 20, 2025 ▶ 35:42 Anthropic co-founder: AGI predictions, leaving OpenAI, what keeps him up at night | Ben Mann
NO PRIORS Insight
Hendrycks: AI alignment is only a subset of AI safety
“So I view the distinction between alignment and safety as alignment as being a sort of subset of safety. Obviously you want the value systems of the AIs to be in keeping with or compatible with say the US public for USAIs or for you as an individual, but that …”
Dan Hendrycks Mar 5, 2025 ▶ 3:46 No Priors Ep. 105 | With Director of the Center of AI Safety Dan Hendrycks
NO PRIORS Insight
Steinberger: AI alignment is only solvable via recursive automated models
“The only way to sort of reasonably approach this is to iteratively ask your model to solve alignment and safety at that stage, not, not, you know, surely you can also ask it to solve your product level problems, but like that, that's nice, but that's not the f…”
Eric Steinberger Aug 30, 2024 ▶ 11:56 No Priors Ep. 79 | With Magic.dev CEO and Co-Founder Eric Steinberger
Bach: Coercing LLMs into good behavior is unsustainable
“At the moment, the idea that we build LLMs that are being coerced with good behavior is not really sustainable. Because if they cannot prove that the behavior is actually good I think we are doomed.”
Joscha Bach Apr 27, 2024 ▶ 1:49:18 This World Does Not Exist — Joscha Bach, Karan Malhotra, Rob Haisfield (WorldSim, WebSim, Liquid AI)
NO PRIORS Insight
Polosukhin: Society needs human alignment rather than AI alignment
“So I have this view that we need human alignment instead of AI alignment. So right now, kind of when we talk about, you know, hey, we need to align AIs with like human values, but the reality is that, you know, all the problems that exist, they all exist becau…”
Illia Polosukhin Sep 14, 2023 ▶ 10:09 No Priors Ep. 32 | With NEAR’s Illia Polosukhin
Andreessen: AI alignment has fundamentally become social engineering and politics
“And for that second risk expressed as AI alignment, what you're dealing with fundamentally is social engineering. And when you're dealing with social engineering, you're necessarily dealing with politics.”
Marc Andreessen Jul 18, 2023 ▶ 56:58 Ep 65: Marc Andreessen and the Case for AI Optimism · Joe Lonsdale
WTF Insight
Mayya: Rule-based AI alignment fails because edge cases cannot be enumerated
“Alignment is about figuring out all these edge cases and saying, don't do this, basically putting it in a sheet. And the problem is we just don't know what it'll end up doing even this world. So nobody can write all the edge cases.”
Varun Mayya May 14, 2023 ▶ 1:41:34 Ep #4 | WTF is ChatGPT: Heaven or Hell? | w/ Nikhil, Varun Mayya, Tanmay, Umang & Aprameya · Nikhil Kamath
a16z Opinion
Masad: AI alignment reflects Silicon Valley sensibilities, not average humans
“I think a lot of what's called AI alignment today is not really aligning with what the average human being wants. It's aligning with like what the sort of Silicon Valley average sensibility is, which I don't think it's good.”
Amjad Masad Feb 16, 2023 ▶ 57:52 The 1000x Developer with Amjad Masad

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.