Nvidia A100, every mention

28 scenes · ← back to Nvidia A100

tap a year for its mentions
00841582023202420252026episodesmentions
0482023202420252026episodes it came up in
001.342.582023202420252026episodesmentions per episode

every year anyone Dylan Patel 7Nader Khalil 4Chris Lattner 3Vipul Ved Prakash 2George Hotz 2Batuhan Taskaya 2Alessio Fanelli 2Yi Tay 1Tri Dao 1Thomas Sohmers 1

Verbatim, from the transcripts: the passages where Nvidia A100 comes up

loading…

Next 100x in AI: Inference, Networking, & Self-Optimizing Models — Philip Kiely & Ali Taha, Baseten Aug 3, 2026 · 1 mention

  • ▶ 1:09:51 Philip Kiely Like the whole full-oh, save full-oh movement, like, you don't gotta have a save llama three movement, you just gotta have an eight 100 somewhere.

The Future of AI Infra: from Kubernetes to Agent Sandboxes — Akshat Bubna, Modal CTO Jul 8, 2026 · 1 mention

  • ▶ 3:36 unnamed speaker Yeah, just like add a 100.

Agent Inference at the "Speed of Light" — How NVIDIA moves like a $4.3 Trillion Startup Mar 8, 2026 · 4 mentions

  • ▶ 4:00 Nader Khalil And, um, whenever we would talk to users, they wanted a GPU, they wanted an A-one hundred. 4 times in the scene

Dylan Patel Explains the AI War While Cooking | In-Context Cooking Feb 26, 2026 · 1 mention

  • ▶ 42:46 Dylan Patel It was a large GPU, um, and it was like having, it was like the best memory, the best networking, everything, sort of the best as possible, um, and sort of like one size fits all, uh, with the main line of like A-one hundred, H-one…

[State of AI Papers 2025] Fixing Research with Social Signals, OCR & Implementation — Team AlphaXiv Dec 31, 2025 · 1 mention

  • ▶ 5:18 unnamed speaker If you host it on your own A-One-Hundreds, and just like, you, you batch things properly, probably Deep Seek is best bang for your buck.

How Zyphra went all-in on AMD + Why Devs feel faster with AI but are slower — with Quentin Anthony Nov 3, 2025 · 1 mention

Building Jamba 3B: the tiny Hybrid Transformer State Space Reasoning Model - Barak Lenz, CTO of AI21 Oct 11, 2025 · 1 mention

  • ▶ 5:27 Barak Lenz So we designed J to have a version that fits on a single GPU, a single AY 100 or H one, 80 gigabytes.

A Technical History of Generative Media Sep 8, 2025 · 2 mentions

  • ▶ 22:50 Batuhan Taskaya And we, like, Kubernetes version at Google Cloud was fine in 2022 when we wanted to get eight A-one-hundreds. 2 times in the scene

⚡️Accelerators @ 3x NVIDIA H200 perf, Made in the USA - Thomas Sohmers + Mitesh Agrawal, Positron AI Aug 18, 2025 · 1 mention

  • ▶ 14:44 Thomas Sohmers And the A 100 comparison here is interesting because in most of these cases, they're actually, they,

The Shape of Compute (Chris Lattner of Modular) Jun 13, 2025 · 3 mentions

SF Compute: Commoditizing Compute Apr 11, 2025 · 1 mention

  • ▶ 22:06 Evan Conrad We just, like, assumed we could go to, like, Lambda, um, or something, and, like, buy thousands of, at the time, A-One-Hundreds.

[Paper Club] BERT: Bidirectional Encoder Representations from Transformers Nov 27, 2024 · 1 mention

  • ▶ 41:15 unnamed speaker Yeah, 20 dollars, they did, like, eight A-One hundreds for an hour, and they're able to match the glue store of basic BERT with their recipe.

Segment Anything 2: Memory + Vision = Object Permanence — with Nikhila Ravi and Joseph Nelson Aug 7, 2024 · 1 mention

  • ▶ 52:54 Joseph Nelson The smallest model is thirty-eight million parameters and can run at 45 FPS on an A-one hundred, right?

The 10,000x Yolo Researcher Metagame — with Yi Tay of Reka Jul 5, 2024 · 1 mention

  • ▶ 47:12 Yi Tay For, for a long period of time, we had, 500 A-One hundreds, because we, we, we, we, we made a commitment, like, uh, and they were constantly being delayed, I think, because of H-One hundred, supply demand, whatever, like, like, reasons…

How AI is Eating Finance - with Mike Conover of Brightwave Jun 11, 2024 · 1 mention

  • ▶ 9:57 Mike Conover The Matic contributors to an investment decision, um, or, or developing your thesis that, um, in response to export controls on a 100 cards, uh, China has put in place licensors on the transfer of germanium and gallium, which are not rare…

Why Google failed to make GPT-3 -- with David Luan of Adept Mar 27, 2024 · 1 mention

  • ▶ 12:11 David Luan One interesting set of stuff is just like, you know, like knowing that a 100 generation that like quad sparsity was going to be a thing.

Truly Serverless Infra for AI Engineers - with Erik Bernhardsson of Modal Feb 19, 2024 · 1 mention

  • ▶ 16:30 Erik Bernhardsson Yeah, it's like, you just, like, say, you know, on the function decorator, you're like, GPU equals, you know, a 100, and then, or like, GPU equals, you know, uh, a 10 or a T four, something like that, and then you get that GPU, and like,…

Building an open AI company - with Ce and Vipul of Together AI Feb 8, 2024 · 4 mentions

  • ▶ 27:40 Vipul Ved Prakash They're mostly A-hundreds and H-hundreds.
  • ▶ 32:20 Alessio Fanelli And this post he said, our model indicates that together it's better off using two a 180 gig system rather than a each 100 based system. 2 times in the scene
  • ▶ 42:07 Vipul Ved Prakash It's, uh, um, you know, the value we get from doing specific optimization, even, even for, you know, what works well for a particular model on A hundreds with a particular bus.

The AI-First Graphics Editor - with Suhail Doshi of Playground AI Jan 2, 2024 · 2 mentions

  • ▶ 21:33 Suhail Doshi And we did that for academic research because there's a whole bunch of, you know, we come across people all the time in academia and they have like, they have access to like one a 100 or eight at best.
  • ▶ 1:02:19 unnamed speaker Uh, you, I, I, you had a tweet about like how many A 100 you have, but I feel like it's out of date probably.

The State of Silicon and the GPU Poors - with Dylan Patel of SemiAnalysis Dec 5, 2023 · 6 mentions

RWKV: Reinventing RNNs for the Transformer Era Aug 31, 2023 · 1 mention

FlashAttention-2: Making Transformers 800% faster AND exact Aug 3, 2023 · 1 mention

  • ▶ 12:28 Tri Dao If you're using, um, a-one-hundred, and you, you list the GPU memory, it's like, 40 gigs or 80 gigs, so that's, that's the, that's the HBM.

Ep 18: Petaflops to the People — with George Hotz of tinycorp Jun 20, 2023 · 2 mentions

  • ▶ 41:53 George Hotz Um, so, the bandwidth is the, is roughly 10 X less than what you can get with NV-linked A-Hundreds. 2 times in the scene
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.