Vision Transformer, every mention

5 scenes across 2 shows · ← back to Vision Transformer

tap a year for its mentions
002142202420252026episodesmentions
012202420252026episodes it came up in
001122202420252026episodesmentions per episode

Latent Space 5the MAD Podcast 2

every year every show Latent Space 5 the MAD Podcast 2

Verbatim, from the transcripts: passages where Vision Transformer comes up on Latent Space, the MAD Podcast

loading…

Inside xAI: Building Grok Imagine in 3 Months, Videogen vs World Models, and Video Agents— Ethan He Jun 1, 2026 · 2 mentions

AI is Already Building AI — Google DeepMind’s Mostafa Dehghani Apr 2, 2026 · 2 mentions

  • ▶ 0:31 Matt Turck Today, my guest is Mustafa Degani, a top AI researcher at Google DeepMind, and a core contributor to some of the most influential architectural breakthroughs of the last decade, including Universal Transformers, the Vision Transformer, and…
  • ▶ 39:56 Matt Turck Another fundamentally important contribution to the, to the field that, uh, you did was the visual transformer paper in 2022.

Best of 2024 in Vision [LS Live @ NeurIPS] Dec 22, 2024 · 2 mentions

  • ▶ 10:18 Isaac Robinson SAM can be used on a single image, in which case the only difference between SAM and SAM is that image encoder, which SAM used a standard VIT. 2 times in the scene

[Paper Club] Molmo + Pixmo + Whisper 3 Turbo - with Vibhu Sapra, Nathan Lambert, Amgadoz Oct 13, 2024 · 1 mention

  • ▶ 34:10 Vibhu Sapra So for the vision encoder, we release, uh, all of our released models use OpenAI's VIT large Clip model, which provides consistently good results.
Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.