PyTorch

product on 20 shows · 42 statements across 29 episodes · said 378 times in 109 episodes since 2017

Latent Space 197 the MAD Podcast 84 the Y Combinator Startup Podcast 22 20VC 21 Acquired 11 No Priors 11 BG2 Pod 6 the a16z Podcast 5 All-In 4 Big Technology 3 TBPN 3 the Neon Show 2 Invest Like the Best 2 In Depth 1 Cheeky Pint 1 the Startup Ideas Podcast 1 A Product Market Fit Show 1 Sourcery 1 Capital Allocators 1 Catalyst 1

Mentions by year, every show

tap a year for its mentions
007520150402017201820192020202120222023202420252026episodesmentions
020402017201820192020202120222023202420252026episodes it came up in
007.52015402017201820192020202120222023202420252026episodesmentions per episode

Latent Space 197the MAD Podcast 84the Y Combinator Startup Podcast 2220VC 21Acquired 11No Priors 11BG2 Pod 6the a16z Podcast 512 more shows

2026 45 mentions in 21 episodes 2 per episode
2025 97 mentions in 27 episodes 4 per episode
2024 131 mentions in 31 episodes 4 per episode
2023 87 mentions in 24 episodes 4 per episode
2021 6 mentions in 4 episodes 2 per episode
2020 11 mentions in 1 episode
2017 1 mention in 1 episode

every mention on every show, scene by scene, with the transcript →

42 statements about PyTorch, every show

LATENT SPACE Assertion Supported
TorchLean enables defining PyTorch-like neural networks directly in Lean
“So what it really enables is that you can now write neural networks essentially in Lean. So instead of writing in, like, PyTorch, it's like a PyTorch-like abstraction, but you can, like, kind of, you know, write it in Lean, and so it can be fully formalized in…”
Anima Anandkumar Aug 26, 2026 ▶ 6:45 🔬 Why Transformers Hit a Wall the Moment Physics Shows Up — Anima Anandkumar, Caltech
MAD Assertion Supported
CMU undergrad AI course has students build an LLM from scratch
“You build a LLM completely from scratch. You use PyTorch, but you build one from scratch that, you know, can be a chatbot. You train it on data. You RL it to solve math problems with tool calls. You do all of this. And this is a undergrad level course.”
Zico Kolter May 7, 2026 ▶ 1:11:36 OpenAI Board Member Zico Kolter: Modern AI Is Just 200 Lines of Code
CAPITAL ALLOCATORS Assertion Contradicted
Baker: Google JAX and Meta PyTorch teams fought publicly on X until leaders called truce
“The JAX team at Google got into a giant fight with the PyTorch team on Meta on X, and the heads of each company's respective AI division had to make a public truce and instruct their troops to stop fighting.”
Gavin Baker Mar 2, 2026 ▶ 1:03:00 Gavin Baker – Truth-Seeking and Crossover Investing at Atreides (EP.489)
Agentic AI warrants a dedicated foundation because agent developers are a distinct discipline
“Agents are a distinct enough set of technology that it merits its own community. [5135] Swyx (Shawn Wang): Separate from data and AI. [5136] Jim Zemlin: Yeah, because like a PyTorch dev isn't really doing a ton of stuff in agent land, right?”
Jim Zemlin Dec 28, 2025 ▶ 1:25:28 One Year of MCP — with David Soria Parria and AAIF leads from OpenAI, Goose, Linux Foundation
Training frontends matter little if attention and MLP kernels are highly optimized
“Most of that is an attention and MOPs, right? So if you have good kernels for attention, MOPs and norms and so on, then it doesn't much matter what the front end to, you know, send tensors to and from those kernels is”
Quentin Anthony Nov 3, 2025 ▶ 6:11 How Zyphra went all-in on AMD + Why Devs feel faster with AI but are slower — with Quentin Anthony
20VC Opinion
In AI Inference, Nobody Cares About Nvidia's CUDA or PyTorch
“In inference, the truth is, nobody cares about CUDA. Nobody even cares about PyTorch. All right. What they want is an API.”
Andrew Feldman Oct 6, 2025 ▶ 24:59 Cerebras CEO, Andrew Feldman on Why Raise $1BN and Delay the IPO & Why NVIDIA’s Worried About Growth · 20VC with Harry Stebbings
Y COMBINATOR Disclosure
Anthropic built custom distributed training to scale beyond Facebook's infrastructure
“We don't want to outsource this to some package because A, we're about to go to a bigger scale, like PyTorch, for instance, they had a package for doing this. But we were going to go to a bigger scale than Facebook had been to. And you don't want to have a dep…”
Nick Joseph Sep 30, 2025 ▶ 12:24 Anthropic Head of Pretraining on Scaling Laws, Compute, and the Future of AI · Y Combinator
Y COMBINATOR Disclosure
Joseph: Anthropic had to hack PyTorch profiler for large-scale GPU clusters
“The PyTorch profiler was, like, pretty good, actually, throughout for a single GPU. You want to, like, profile a GPU, the PyTorch profile would work. But if you wanted to profile a job on 100,000 of GPUs, that, like, hadn't really been done much, and then that…”
Nick Joseph Sep 30, 2025 ▶ 16:37 Anthropic Head of Pretraining on Scaling Laws, Compute, and the Future of AI · Y Combinator
LATENT SPACE Assertion Supported
Sohmers: AMD's PyTorch fork lagged official releases by 6-9 months
“And AMD had their own separate you know, non-mainline PyTorch distribution for years. That was always six to nine months behind any new PyTorch releases.”
Thomas Sohmers Aug 18, 2025 ▶ 20:06 ⚡️Accelerators @ 3x NVIDIA H200 perf, Made in the USA - Thomas Sohmers + Mitesh Agrawal, Positron AI
PyTorch democratized model training, but AI inference remains undemocratized
“And so things like PyTorch came on the scene and I think PyTorch gets all credit for democratizing model training, right? It's taught to pretty much every computer science student that graduates. That's a huge deal, but nobody democratized inference. Inference…”
Chris Lattner Jun 13, 2025 ▶ 35:44 The Shape of Compute (Chris Lattner of Modular)
MAD Assertion Not checkable as stated
Meta spent five years rebuilding PyTorch's backend for internal scale
“It took us five years. Took us five years to get the stage supporting almost all internal needs using deep learning and mass and massive scale.”
Lin Qiao Mar 27, 2025 ▶ 6:48 Why This Ex-Meta Leader is Rethinking AI Infrastructure | Lin Qiao, CEO, Fireworks AI
MAD Assertion Supported
Lin Qiao: OpenAI switched completely from TensorFlow to PyTorch
“OpenAI switched to use PyTorch fully.”
Lin Qiao Mar 27, 2025 ▶ 8:00 Why This Ex-Meta Leader is Rethinking AI Infrastructure | Lin Qiao, CEO, Fireworks AI
MAD Insight
Lin Qiao: PyTorch's primary success lesson is that simplicity scales
“I think one of the biggest success we saw from the PyTorch experience is simplicity scales.”
Lin Qiao Mar 27, 2025 ▶ 16:26 Why This Ex-Meta Leader is Rethinking AI Infrastructure | Lin Qiao, CEO, Fireworks AI
MAD Assertion Not checkable as stated
Lin Qiao: Meta had hundreds of engineers building PyTorch and its infrastructure
“We have hundreds of engineers building PyTorch and infrastructure around PyTorch, but at the same time, I believe PyTorch within Meta probably has thousands of users.”
Lin Qiao Mar 27, 2025 ▶ 19:23 Why This Ex-Meta Leader is Rethinking AI Infrastructure | Lin Qiao, CEO, Fireworks AI
LATENT SPACE Disclosure
Comfyanonymous had never written PyTorch before October 2022
“So basically October, 20, 22, just like I hadn't written a line of PyTorch before that. So it's completely new.”
comfyanonymous (Comfy) Jan 4, 2025 ▶ 2:45 AI Engineering for Art - with comfyanonymous
Comfyanonymous: PyTorch Lacks Fine-Grained Memory Control for Complex Pipelines
“The problem with PyTorch is it's high levels. Don't have that much fine-grained control over, like specific memory stuff, so kind of have to leave, like, the memory freeing to Python and PyTorch, which is, can be annoying sometimes.”
comfyanonymous (Comfy) Jan 4, 2025 ▶ 33:13 AI Engineering for Art - with comfyanonymous
Fu: Changing one PyTorch line requires a week of CUDA development
“If we decided to change one thing in PyTorch, like one line of PyTorch code is like a week of CUDA code at least.”
Dan Fu Dec 24, 2024 ▶ 29:38 2024 in Post-Transformer Architectures: State Space Models, RWKV [Latent Space LIVE! @ NeurIPS 2024]
LATENT SPACE Assertion Supported
PyTorch was originally built for researchers without considering production requirements
“PyTorch actually started as the framework for researchers. Don't care about production at all.”
Lin Qiao Nov 25, 2024 ▶ 4:22 Why Compound AI + Open Source will beat Closed AI — with Lin Qiao, CEO of Fireworks AI
BG2 Assertion Not checkable as stated
Huang: cuDNN revolutionized deep learning by powering frameworks like PyTorch
“So we revolutionized deep learning because of our domain specific library called QDNN. Without QDNN, nobody talks about QDNN because it's one layer underneath PyTorch and, you know, and TensorFlow and back in the old days, CAFE and Theano and now Triton, and t…”
Jensen Huang Oct 13, 2024 ▶ 14:51 Ep17. Welcome Jensen Huang | BG2 w/ Bill Gurley & Brad Gerstner · Bg2 Pod
a16z Insight
Schmidt: Synchronized AI training stems from PyTorch convenience abstractions, not optimal convergence
“But the one monolithic thing was actually just like a technical bot. Like it was from the fact that like we had PyTorch and then they like, you know, or like at Karis or any of the other ones. And they're like, Well, if you want, you can train on multiple GPUs…”
Jeff Schmidt Oct 1, 2024 ▶ 1:04:01 The Quest for Community-Trained Open Source AI Models
Karpathy: Without PyTorch, developers are 'naked in the world'
“So, PyTorch is really, really nice, and this is just some of the things that PyTorch offers. So, without PyTorch, we're kind of naked in the world, right?”
Andrej Karpathy Sep 21, 2024 ▶ 5:09 llm.c's Origin and the Future of LLM Compilers - Andrej Karpathy at CUDA MODE
LATENT SPACE Assertion Supported
Karpathy: llm.c was 20% faster and used 30% less memory than PyTorch
“At the time of that post, we were using, in LL and that's in 30% less memory, and we were 20% faster in training, just the truth.”
Andrej Karpathy Sep 21, 2024 ▶ 19:10 llm.c's Origin and the Future of LLM Compilers - Andrej Karpathy at CUDA MODE
Karpathy: Python and PyTorch are crutches for finite human intelligence
“The use of Python and PyTorch and everything else is just a crutch, because we humans are finite. We have finite knowledge, intelligence, and attention.”
Andrej Karpathy Sep 21, 2024 ▶ 22:17 llm.c's Origin and the Future of LLM Compilers - Andrej Karpathy at CUDA MODE
Tay: Complex architecture modifications fail due to an implementation lottery
“A lot of architecture changes, right, the moment they are, like, tedious to implement, like, nobody, like, SuiGuru is a simple thing, right? [4306] Yi Tay: Just split it and then get it. [4307] Yi Tay: It's a very simple thing to implement. [4309] Yi Tay: Mayb…”
Yi Tay Jul 5, 2024 ▶ 1:11:30 The 10,000x Yolo Researcher Metagame — with Yi Tay of Reka
LATENT SPACE Assertion Not checkable as stated
Huang: EasyContext was the first functional PyTorch Ring Attention implementation
“Easy context was the first PyTorch implementation that applied it with native libraries that worked pretty well. And then we adapted it ourselves in order to configure it for our cluster network topology.”
Mark Huang May 31, 2024 ▶ 27:08 How to train a Million Context LLM — with Mark Huang of Gradient.ai
LATENT SPACE Assertion Contradicted
Chintala: PyTorch is around 190,000 lines of code
“PyTorch is like a 190,000 lines of code or something at this point.”
Soumith Chintala Mar 6, 2024 ▶ 6:37 Open Source AI is AI we can Trust — with Soumith Chintala of Meta AI
LATENT SPACE Prediction Not checkable as stated
Chintala: George Hotz's TinyGrad requires major breakthroughs to match PyTorch
“There's no, like, I don't think, like, unless we have, like, great breakthroughs, like, George's vision is achievable, like, or, like, he should be thinking about a narrower problem, such as, I'm only gonna make this for, like, work for self-driving car con ne…”
Soumith Chintala Mar 6, 2024 ▶ 9:40 Open Source AI is AI we can Trust — with Soumith Chintala of Meta AI
Chintala: Deep PyTorch-Mojo integration lacks synergy because Mojo replaces PyTorch frontend
“Mojo as a fundamental frontend would be replacing PyTorch, not, like, augmenting PyTorch. So, in that sense, I don't see a synergy in more deeply, like, integrating Mojo.”
Soumith Chintala Mar 6, 2024 ▶ 15:57 Open Source AI is AI we can Trust — with Soumith Chintala of Meta AI
LATENT SPACE Prediction Not checkable as stated
Chintala: Apple's MLX will fail server-side due to lack of differentiation
“If they end up expanding onto the server side, and they'll probably build something like PyTorch as well, right? Like, eventually, that'll where it will land. And I think there, they will kind of fail on the, like, lack of differentiation. Like, it wouldn't be…”
Soumith Chintala Mar 6, 2024 ▶ 18:12 Open Source AI is AI we can Trust — with Soumith Chintala of Meta AI
LATENT SPACE Assertion Supported
Chintala: PyTorch is used in Mars rover simulations, drug discovery, and Tesla
“It's used in Mars rover simulations, to drug discovery, to Tesla cars, and there's a huge diversity of, like, applications in which it is used in.”
Soumith Chintala Mar 6, 2024 ▶ 30:47 Open Source AI is AI we can Trust — with Soumith Chintala of Meta AI
LATENT SPACE Assertion Supported
Chintala: CERN uses PyTorch and GANs for particle physics research
“I think the scariest was when I went to visit CERN at some point, and they said they were using it, PyTorch, and they were using GANs at the same time for, like, particle physics research, and I was scared more about the fact that they were using GANs than the…”
Soumith Chintala Mar 6, 2024 ▶ 31:21 Open Source AI is AI we can Trust — with Soumith Chintala of Meta AI
NO PRIORS Assertion Supported
Papermaster: AMD is one of two qualified hardware platforms on PyTorch
“We're, One of two qualified offerings on on PyTorch, and so all of that testing is being done you know, routinely with the regression testing that's run literally every night on any software release.”
Mark Papermaster Feb 29, 2024 ▶ 13:30 No Priors Ep. 53 | With AMD CTO Mark Papermaster
Firshman: AI engineers do not need low-level PyTorch expertise
“The metaphor here is that you don't need to be digging down into like this sort of PyTorch level if you don't want to in the same way as a software engineer in the nineties. You don't need to be like understanding how network stacks work to be able to build a …”
Ben Firshman Feb 28, 2024 ▶ 1:16:37 A Brief History of the Open Source AI Hacker - with Ben Firshman of Replicate
20VC Disclosure
Grimshaw: Benchmark invested in LangChain, Cerebras, and ex-PyTorch leads
“Obviously with investment in Langtrain and we have a few others at Benchmark Cerebrus, which is you know, an AI chip for training and inference. We backed a team that worked on and was the leads of PyTorch and has got a new company and some others.”
Miles Grimshaw Sep 18, 2023 ▶ 1:23:12 Miles Grimshaw: The 5 Pillars of Venture Capital & Why Co-Pilot is an Incumbent Strategy | E1061 · 20VC with Harry Stebbings
MAD Insight
Biewald: PyTorch beat TensorFlow through developer empathy, not eager execution
“I don't think they really, I think people tell this, the story of sort of the silver bullet. Of like you know, the eager execution model. But I think the reality is they just built a product with so much more empathy.”
Lukas Biewald Aug 9, 2023 ▶ 47:35 Startup to Industry Standard: Lukas Biewald Explains How W&B Scaled MLOps for OpenAI, NVIDIA & More
20VC What-if
Douwe Kiela: Current AI breakthroughs would not happen without Meta's PyTorch
“So PyTorch really without PyTorch, none of this stuff would be happening right now. And so it's really like fundamental for all of the AI breakthroughs.”
Douwe Kiela Jun 30, 2023 ▶ 4:59 Douwe Kiela: Why Data Size Matters More Than Model Size; Why Open Source Isn't Going to Win | E1032 · 20VC with Harry Stebbings
LATENT SPACE Assertion Supported
Hotz: Tinygrad is about 5x slower than PyTorch on Nvidia GPUs
“The correctness for both forwards and backwards passes is there, but on Nvidia, it's about five X slower than PyTorch right now.”
George Hotz Jun 20, 2023 ▶ 21:15 Ep 18: Petaflops to the People — with George Hotz of tinycorp
LATENT SPACE Prediction Not checkable as stated
Hotz: Tinygrad could replicate PyTorch's API in two engineer-months
“Replicating the PyTorch API. Is something I can do with a couple, you know, like an engineer month or two.”
George Hotz Jun 20, 2023 ▶ 31:32 Ep 18: Petaflops to the People — with George Hotz of tinycorp
20VC Assertion Supported
LeCun: All of OpenAI and the AI world runs on PyTorch
“ChatGPT was developed on PyTorch. Okay. All OpenAI runs on PyTorch. The entire world, in fact, runs on PyTorch, except Google, because they have their own thing, right?”
Yann LeCun May 15, 2023 ▶ 25:39 Yann LeCun: Meta’s New AI Model LLaMA; Why Elon is Wrong about AI; Open-source AI Models | E1014 · 20VC with Harry Stebbings
MAD Disclosure
DoorDash standardized its core machine learning platform on LightGBM and PyTorch
“We landed on using a framework that enables tree-based models. And we picked light GBM for that after trying a few different packages and also deep learning. And for that, we then used PyTorch. And so we started with those two core libraries.”
Alok Gupta Feb 1, 2021 ▶ 6:29 Fireside Chat: Alok Gupta (Head of Data Science & ML, DoorDash) with Matt Turck (Partner, FirstMark)
MAD Disclosure
Pesenti: Facebook is going end-to-end all-in on PyTorch
“We're definitely going all in as PyTorch, you know, end to end. So I think initially when we launched the Onyx strategy, it was more like a multi-framework world. And we had actually two framework internally between PyTorch and Cafe Two, but we're still suppor…”
Jerome Pesenti Jun 10, 2020 ▶ 55:46 Fireside Chat: Jerome Pesenti (Head of AI, Facebook) with Matt Turck (Partner, FirstMark)
MAD Insight
Karpinski: Alternating language layers in AI frameworks prevents compiler optimizations
“They have what I've, I would describe as a sandwich problem, which is that you end up sandwiching a lot of system code with user code, and then, like, adding more and more layers of that, and as you've sandwiched, like, you know, seven or eight layers of that,…”
Stefan Karpinski Sep 28, 2017 ▶ 6:39 A Fresh Approach to Technical Computing // Viral Shah & Stefan Karpinski, Julia Computing

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.