Optical Character Recognition

topic on 3 shows · 4 statements across 4 episodes

Latent Space How I Built This the a16z Podcast

4 statements about Optical Character Recognition, every show

LATENT SPACE Disclosure
Zhang: Meta intentionally avoided OCR-heavy images during SAM 3 training data sampling
“In fact, during our data engine, we intentionally do not sample OCR-heavy images.”
Pengchuan Zhang Dec 18, 2025 ▶ 37:11 SAM 3: The Eyes for AI — Nikhila & Pengchuan (Meta Superintelligence), ft. Joseph Nelson (Roboflow)
HOW I BUILT THIS Assertion Supported
von Ahn: Early book scanning software failed on 30% of words
“Computers cannot, or at the time, could not recognize many of the words, about 30% of the words computers could not recognize.”
Luis von Ahn Dec 18, 2023 ▶ 22:14 reCAPTCHA and Duolingo: Luis von Ahn (2020)
HOW I BUILT THIS Assertion Supported
Von Ahn: Early OCR failed on 30% of scanned book words
“Computers cannot, or at the time, could not recognize many of the words. About 30% of the words computers could not recognize.”
Luis von Ahn May 25, 2020 ▶ 22:04 reCAPTCHA and Duolingo: Luis von Ahn
a16z Assertion Contradicted
David Rumsey: OCR cannot be performed on historical paper maps
“Optical character recognition in old maps, but it's impossible to do because the text is going up and down and they're different fonts and they're artistically.”
David Rumsey Jan 2, 2019 ▶ 36:48 a16z Podcast | Exploding the Map

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.