PyTorch

includes PyTorch FSDP, PyTorch / XLA

23 statements across 14 episodes · 6 bullish · 6 bearish · 14 people on the record · first statement Jun 20, 2023 by George Hotz · said 199 times in 43 episodes since 2023 · across every show →

Mentions by year, the whole family

brought up most by Andrej Karpathy (27), Shawn Wang (26), George Hotz (25), Chris Lattner (11), comfyanonymous (Comfy) (6), Alessio Fanelli (6), Evan Feinberg (5), Thomas Sohmers (4)

tap a year for its mentions
00758150152023202420252026episodesmentions
08152023202420252026episodes it came up in
0047.58152023202420252026episodesmentions per episode
2026 13 mentions in 7 episodes 2 per episode
2025 44 mentions in 15 episodes 3 per episode
2024 103 mentions in 15 episodes 7 per episode
2023 39 mentions in 6 episodes 7 per episode

every mention, scene by scene, with the transcript →

Everything said about PyTorch, oldest first

Jun 20, 2023 neutral
Assertion Supported
Hotz: Tinygrad is about 5x slower than PyTorch on Nvidia GPUs
“The correctness for both forwards and backwards passes is there, but on Nvidia, it's about five X slower than PyTorch right now.”
George Hotz Jun 20, 2023 ▶ 21:15 Ep 18: Petaflops to the People — with George Hotz of tinycorp
Jun 20, 2023 bullish
Prediction Not checkable as stated
Hotz: Tinygrad could replicate PyTorch's API in two engineer-months
“Replicating the PyTorch API. Is something I can do with a couple, you know, like an engineer month or two.”
George Hotz Jun 20, 2023 ▶ 31:32 Ep 18: Petaflops to the People — with George Hotz of tinycorp
Feb 28, 2024 neutral
Insight
Firshman: AI engineers do not need low-level PyTorch expertise
“The metaphor here is that you don't need to be digging down into like this sort of PyTorch level if you don't want to in the same way as a software engineer in the nineties. You don't need to be like understanding how network stacks work to be able to build a …”
Ben Firshman Feb 28, 2024 ▶ 1:16:37 A Brief History of the Open Source AI Hacker - with Ben Firshman of Replicate
Mar 6, 2024 positive
Assertion Supported
Chintala: PyTorch is used in Mars rover simulations, drug discovery, and Tesla
“It's used in Mars rover simulations, to drug discovery, to Tesla cars, and there's a huge diversity of, like, applications in which it is used in.”
Soumith Chintala Mar 6, 2024 ▶ 30:47 Open Source AI is AI we can Trust — with Soumith Chintala of Meta AI
Mar 6, 2024 bearish
Prediction Not checkable as stated
Chintala: Apple's MLX will fail server-side due to lack of differentiation
“If they end up expanding onto the server side, and they'll probably build something like PyTorch as well, right? Like, eventually, that'll where it will land. And I think there, they will kind of fail on the, like, lack of differentiation. Like, it wouldn't be…”
Soumith Chintala Mar 6, 2024 ▶ 18:12 Open Source AI is AI we can Trust — with Soumith Chintala of Meta AI
Mar 6, 2024 negative
Opinion
Chintala: Deep PyTorch-Mojo integration lacks synergy because Mojo replaces PyTorch frontend
“Mojo as a fundamental frontend would be replacing PyTorch, not, like, augmenting PyTorch. So, in that sense, I don't see a synergy in more deeply, like, integrating Mojo.”
Soumith Chintala Mar 6, 2024 ▶ 15:57 Open Source AI is AI we can Trust — with Soumith Chintala of Meta AI
Mar 6, 2024 neutral
Assertion Supported
Chintala: CERN uses PyTorch and GANs for particle physics research
“I think the scariest was when I went to visit CERN at some point, and they said they were using it, PyTorch, and they were using GANs at the same time for, like, particle physics research, and I was scared more about the fact that they were using GANs than the…”
Soumith Chintala Mar 6, 2024 ▶ 31:21 Open Source AI is AI we can Trust — with Soumith Chintala of Meta AI
Mar 6, 2024 bearish
Prediction Not checkable as stated
Chintala: George Hotz's TinyGrad requires major breakthroughs to match PyTorch
“There's no, like, I don't think, like, unless we have, like, great breakthroughs, like, George's vision is achievable, like, or, like, he should be thinking about a narrower problem, such as, I'm only gonna make this for, like, work for self-driving car con ne…”
Soumith Chintala Mar 6, 2024 ▶ 9:40 Open Source AI is AI we can Trust — with Soumith Chintala of Meta AI
Mar 6, 2024 neutral
Assertion Contradicted
Chintala: PyTorch is around 190,000 lines of code
“PyTorch is like a 190,000 lines of code or something at this point.”
Soumith Chintala Mar 6, 2024 ▶ 6:37 Open Source AI is AI we can Trust — with Soumith Chintala of Meta AI
May 31, 2024 positive
Assertion Not checkable as stated
Huang: EasyContext was the first functional PyTorch Ring Attention implementation
“Easy context was the first PyTorch implementation that applied it with native libraries that worked pretty well. And then we adapted it ourselves in order to configure it for our cluster network topology.”
Mark Huang May 31, 2024 ▶ 27:08 How to train a Million Context LLM — with Mark Huang of Gradient.ai
Jul 5, 2024 neutral
Insight
Tay: Complex architecture modifications fail due to an implementation lottery
“A lot of architecture changes, right, the moment they are, like, tedious to implement, like, nobody, like, SuiGuru is a simple thing, right? [4306] Yi Tay: Just split it and then get it. [4307] Yi Tay: It's a very simple thing to implement. [4309] Yi Tay: Mayb…”
Yi Tay Jul 5, 2024 ▶ 1:11:30 The 10,000x Yolo Researcher Metagame — with Yi Tay of Reka
Sep 21, 2024 positive
Opinion
Karpathy: Without PyTorch, developers are 'naked in the world'
“So, PyTorch is really, really nice, and this is just some of the things that PyTorch offers. So, without PyTorch, we're kind of naked in the world, right?”
Andrej Karpathy Sep 21, 2024 ▶ 5:09 llm.c's Origin and the Future of LLM Compilers - Andrej Karpathy at CUDA MODE
Sep 21, 2024 negative
Insight
Karpathy: Python and PyTorch are crutches for finite human intelligence
“The use of Python and PyTorch and everything else is just a crutch, because we humans are finite. We have finite knowledge, intelligence, and attention.”
Andrej Karpathy Sep 21, 2024 ▶ 22:17 llm.c's Origin and the Future of LLM Compilers - Andrej Karpathy at CUDA MODE
Sep 21, 2024 positive
Assertion Supported
Karpathy: llm.c was 20% faster and used 30% less memory than PyTorch
“At the time of that post, we were using, in LL and that's in 30% less memory, and we were 20% faster in training, just the truth.”
Andrej Karpathy Sep 21, 2024 ▶ 19:10 llm.c's Origin and the Future of LLM Compilers - Andrej Karpathy at CUDA MODE
Nov 25, 2024 neutral
Assertion Supported
PyTorch was originally built for researchers without considering production requirements
“PyTorch actually started as the framework for researchers. Don't care about production at all.”
Lin Qiao Nov 25, 2024 ▶ 4:22 Why Compound AI + Open Source will beat Closed AI — with Lin Qiao, CEO of Fireworks AI
Dec 24, 2024
Insight
Fu: Changing one PyTorch line requires a week of CUDA development
“If we decided to change one thing in PyTorch, like one line of PyTorch code is like a week of CUDA code at least.”
Dan Fu Dec 24, 2024 ▶ 29:38 2024 in Post-Transformer Architectures: State Space Models, RWKV [Latent Space LIVE! @ NeurIPS 2024]
Jan 4, 2025
Disclosure
Comfyanonymous had never written PyTorch before October 2022
“So basically October, 20, 22, just like I hadn't written a line of PyTorch before that. So it's completely new.”
comfyanonymous (Comfy) Jan 4, 2025 ▶ 2:45 AI Engineering for Art - with comfyanonymous
Jan 4, 2025 negative
Insight
Comfyanonymous: PyTorch Lacks Fine-Grained Memory Control for Complex Pipelines
“The problem with PyTorch is it's high levels. Don't have that much fine-grained control over, like specific memory stuff, so kind of have to leave, like, the memory freeing to Python and PyTorch, which is, can be annoying sometimes.”
comfyanonymous (Comfy) Jan 4, 2025 ▶ 33:13 AI Engineering for Art - with comfyanonymous
Jun 13, 2025 neutral
Opinion
PyTorch democratized model training, but AI inference remains undemocratized
“And so things like PyTorch came on the scene and I think PyTorch gets all credit for democratizing model training, right? It's taught to pretty much every computer science student that graduates. That's a huge deal, but nobody democratized inference. Inference…”
Chris Lattner Jun 13, 2025 ▶ 35:44 The Shape of Compute (Chris Lattner of Modular)
Aug 18, 2025 negative
Assertion Supported
Sohmers: AMD's PyTorch fork lagged official releases by 6-9 months
“And AMD had their own separate you know, non-mainline PyTorch distribution for years. That was always six to nine months behind any new PyTorch releases.”
Thomas Sohmers Aug 18, 2025 ▶ 20:06 ⚡️Accelerators @ 3x NVIDIA H200 perf, Made in the USA - Thomas Sohmers + Mitesh Agrawal, Positron AI
Nov 3, 2025
Insight
Training frontends matter little if attention and MLP kernels are highly optimized
“Most of that is an attention and MOPs, right? So if you have good kernels for attention, MOPs and norms and so on, then it doesn't much matter what the front end to, you know, send tensors to and from those kernels is”
Quentin Anthony Nov 3, 2025 ▶ 6:11 How Zyphra went all-in on AMD + Why Devs feel faster with AI but are slower — with Quentin Anthony
Dec 28, 2025 positive
Opinion
Agentic AI warrants a dedicated foundation because agent developers are a distinct discipline
“Agents are a distinct enough set of technology that it merits its own community. [5135] Swyx (Shawn Wang): Separate from data and AI. [5136] Jim Zemlin: Yeah, because like a PyTorch dev isn't really doing a ton of stuff in agent land, right?”
Jim Zemlin Dec 28, 2025 ▶ 1:25:28 One Year of MCP — with David Soria Parria and AAIF leads from OpenAI, Goose, Linux Foundation
Aug 26, 2026
Assertion Supported
TorchLean enables defining PyTorch-like neural networks directly in Lean
“So what it really enables is that you can now write neural networks essentially in Lean. So instead of writing in, like, PyTorch, it's like a PyTorch-like abstraction, but you can, like, kind of, you know, write it in Lean, and so it can be fully formalized in…”
Anima Anandkumar Aug 26, 2026 ▶ 6:45 🔬 Why Transformers Hit a Wall the Moment Physics Shows Up — Anima Anandkumar, Caltech
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.