Language Modeling

topic on 5 shows · 7 statements across 7 episodes

the Y Combinator Startup Podcast the Knowledge Project No Priors the MAD Podcast 20VC

7 statements about Language Modeling, every show

KNOWLEDGE PROJECT Assertion Not checkable as stated
Brockman: 2017 sentiment neuron proved semantics emerge from next-character prediction
“Because it's 2017, and it's really the first time that we saw semantics arise from training on A language modeling objective. So you train on learn the next character, predict the next character, and then suddenly you get a neural net that understands sentimen…”
Greg Brockman Apr 22, 2026 ▶ 7:07 Ai Goes Parabolic | OpenAI Co-Founder Greg Brockman
Finn: Pre-training and curated fine-tuning unlocked robotic laundry folding
“And this was actually to take some inspiration from the world of language modeling to actually instead of just training a policy on all of our data, can we pre-train on all the data? And then fine tune on a highly, on a curated, consistent, high quality set of…”
Chelsea Finn Jul 22, 2025 ▶ 9:21 Chelsea Finn: Building Robots That Can Do Anything · Y Combinator
MAD Assertion Not checkable as stated
Gomez: Google failed to lean into language modeling early, unlike OpenAI
“To say they didn't lean hard enough into language modeling, like just pure Sequence modeling of text on the internet. That's, I think the accurate statement. That's what OpenAI did early and uniquely well.”
Aidan Gomez Jun 5, 2025 ▶ 11:48 Inside the Paper That Changed AI Forever - Cohere CEO Aidan Gomez on 2025 Agents
MAD Insight
Chip Huyen: Lack of labeled data requirements makes language modeling uniquely scalable
“You don't need to curate, like, labels, like, reference data, so that, that you can use a train models that make language modeling, like, so, so much easier to scale than other types of tasks.”
Chip Huyen Jan 16, 2025 ▶ 16:32 What You MUST Know About AI Engineering | Chip Huyen, Author of “AI Engineering”
20VC Disclosure
Gomez: Industry took two to three years to realize language model scaling worked
“With language modeling and the whole scaling project, I thought the world would catch on way faster to that piece. It started to become really obvious, but then it was two, three years before everyone woke up, and it sort of hit the world.”
Aidan Gomez Aug 19, 2024 ▶ 20:56 Aidan Gomez: What No One Understands About Foundation Models | E1191 · 20VC with Harry Stebbings
NO PRIORS Assertion Supported
Gu: Mamba successfully applied state-space models to language modeling
“Recently proposed a model called Mamba which was kind of brought these to language modeling and showed really good results there.”
Albert Gu Jun 27, 2024 ▶ 2:24 No Priors Ep. 70 | With Cartesia Co-Founders Karan Goel & Albert Gu
NO PRIORS Insight
Shazeer Calls Next-Token Prediction an AI-Complete Problem
“And like, the problem is super simple to define. It's just like predict the next word, the fat cat sat on the, you know, like, okay, well, you know, what comes next? Like it's extremely easy to define. And if you can do a great job of it, like, you know, then …”
Noam Shazeer Apr 25, 2023 ▶ 3:03 No Priors Ep. 12 | With Noam Shazeer

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.