Dylan Fox

Founder & CEO, AssemblyAI · 1 appearance on the record.

computed by AI from the episodes · how this works → · full disclaimer →

founderexecutiveengineer@YouveGotFox ↗LinkedIn ↗assemblyai.com ↗

Dylan Fox founded AssemblyAI in 2017 to build speech AI models and APIs for speech-to-text transcription and audio intelligence. Before starting AssemblyAI, he worked as a machine learning and software engineer at Cisco Systems developing AI and NLP products.

13statements → 6claims → 3claims resolved → 67%fully supported → 3.77/5average certainty → 1.54/5average debate potential →

2 supported 0 partly supported 1 contradicted 3 not checkable as stated how the 6 claims stand · each chip opens the sources

6 assertions · 2 opinions · 1 insight · 4 disclosures · every statement was checked. The predictions and assertions are the 6 claims: statements the public record can support or contradict. 3 are resolved, and 3 name no date, number or outcome precise enough to check. Everything else (opinions, insights, what ifs, disclosures) can never be settled by the record, so it carries no assessment.

The record, in short

What the tape says about how Dylan argues and how the claims held up. Everything they said, and everything said about them, is in the tabs below.

Their most notable supported claim

Assertion Supported
State-of-the-art speech recognition models still carry a 15% error rate
“State-of-the-art automatic speech recognition still has, like, a 15% error rate on a lot of data sets”
Dylan Fox May 1, 2023 ▶ 20:05 Generative AI for Speech Recognition | AssemblyAI Founder & CEO, Dylan Fox

Their most notable contradicted claim

Assertion Contradicted
AssemblyAI's JAX contribution sped up Whisper model training by 10x
“We actually, I think, published like, made a contribution to Jax to make it, like, 10 times faster to train Whisper.”
Dylan Fox May 1, 2023 ▶ 26:18 Generative AI for Speech Recognition | AssemblyAI Founder & CEO, Dylan Fox

Expressed certainty vs assessment result

none yet certainty 1
none yet certainty 2
0% certainty 3
100% certainty 4
none yet certainty 5

weighted support: a fully supported claim counts one, a partly supported claim counts half. Each filled bar is clickable and opens exactly those claims; "none yet" means nothing said at that certainty level has resolved yet

How they sound: speaking style how? →

245 words/min while actually speaking · 36.6 um and uh per 1k words

No argument clarity score for Dylan Fox: no usable question→answer exchanges on raw tape (a fair score needs 8+). We do not score a sample that small. Roundtable and news formats yield far fewer direct exchanges than interviews.

Measured by listening to the audio itself: 4,536 words across 1 episode of raw-level tape, transcribed verbatim with every um and uh kept, each one attributed only where the alignment onto our timed stream is unambiguous. These are measurements of speaking style. We do not rank them: across this corpus, fluency and argument quality are nearly uncorrelated (ρ≈0.2), and smooth talking does not signal clear thinking. How it's measured →

Everything Dylan Fox said on the MAD Podcast that made the record, most notable first. Filter by type, assessment or year in the ledger →

Opinion
Major cloud providers are too big to ship good developer products
“I'm surprised the big cloud companies, and apologies if anyone here works there, I'm surprised they can't ship, you know, better developer products, but it's, I think they're maybe just too big at this point.”
Dylan Fox May 1, 2023 ▶ 23:20 Generative AI for Speech Recognition | AssemblyAI Founder & CEO, Dylan Fox
Insight
Most speech AI value comes from downstream workflows, not just transcription
“Where a lot of value is created is you're taking the transcription, and then you're using it as an input to do something else.”
Dylan Fox May 1, 2023 ▶ 3:15 Generative AI for Speech Recognition | AssemblyAI Founder & CEO, Dylan Fox
Disclosure
Fox: AssemblyAI delays exploratory model investment until developers prove value
“Some of these other things that are more exploratory, like, we're not gonna put a ton of effort into those until we see that our customers the developers that use our API are actually able to find value and create value with those.”
Dylan Fox May 1, 2023 ▶ 7:42 Generative AI for Speech Recognition | AssemblyAI Founder & CEO, Dylan Fox
Assertion Not checkable as stated
Most commercial speech models historically trained on roughly 50,000 hours of audio
“And our models prior, and most commercial speech recognition models trained on like, 50,000 hours”
Dylan Fox May 1, 2023 ▶ 14:09 Generative AI for Speech Recognition | AssemblyAI Founder & CEO, Dylan Fox
Disclosure
AssemblyAI is training its next speech model on ~4 million hours of audio
“We're actually training conformer two or what might call it 1.5, but whatever this accessory will be is training right now. And that's something around four million hours of labeled audio data.”
Dylan Fox May 1, 2023 ▶ 14:19 Generative AI for Speech Recognition | AssemblyAI Founder & CEO, Dylan Fox
Assertion Not checkable as stated
AssemblyAI processes over 100 million audio files monthly via API
“We've processed, ah, yeah, it's like over a hundred million audio files a month that are flowing through the API, and that's growing pretty quickly.”
Dylan Fox May 1, 2023 ▶ 16:30 Generative AI for Speech Recognition | AssemblyAI Founder & CEO, Dylan Fox
Assertion Supported
State-of-the-art speech recognition models still carry a 15% error rate
“State-of-the-art automatic speech recognition still has, like, a 15% error rate on a lot of data sets”
Dylan Fox May 1, 2023 ▶ 20:05 Generative AI for Speech Recognition | AssemblyAI Founder & CEO, Dylan Fox
Assertion Not checkable as stated
AssemblyAI has processed almost two billion audio files to date
“We've processed almost two billion audio files through our system.”
Dylan Fox May 1, 2023 ▶ 0:47 Generative AI for Speech Recognition | AssemblyAI Founder & CEO, Dylan Fox
Disclosure
Fox: AssemblyAI serves over 1,000 customers and tens of thousands of monthly developers
“We've got over a thousand customers tens of thousands a month of developers that are building with the API.”
Dylan Fox May 1, 2023 ▶ 2:10 Generative AI for Speech Recognition | AssemblyAI Founder & CEO, Dylan Fox
Assertion Supported
AssemblyAI trained Conformer-1 on 650,000 hours of labeled audio data
“So we trained it on, like, 60 terabytes of audio data, like, labeled audio data. So it was, I think, something like 650,000 hours of audio data.”
Dylan Fox May 1, 2023 ▶ 13:59 Generative AI for Speech Recognition | AssemblyAI Founder & CEO, Dylan Fox
Assertion Contradicted
AssemblyAI's JAX contribution sped up Whisper model training by 10x
“We actually, I think, published like, made a contribution to Jax to make it, like, 10 times faster to train Whisper.”
Dylan Fox May 1, 2023 ▶ 26:18 Generative AI for Speech Recognition | AssemblyAI Founder & CEO, Dylan Fox
Disclosure
AssemblyAI employs about 40 full-time staff dedicated to improving speech models
“We got like 40 people full time working on this, you know, and you're gonna get all the benefits of that.”
Dylan Fox May 1, 2023 ▶ 27:44 Generative AI for Speech Recognition | AssemblyAI Founder & CEO, Dylan Fox
Opinion
Large organizations remain confused about who should manage internal AI projects
“I think right now, larger organizations sometimes are, like, still confused, like, who's gonna manage this AI project, you know?”
Dylan Fox May 1, 2023 ▶ 33:15 Generative AI for Speech Recognition | AssemblyAI Founder & CEO, Dylan Fox

Appearances (1)

EpisodeDateSpeaking time
Generative AI for Speech Recognition | AssemblyAI Founder & CEO, Dylan Fox May 1, 2023 25m
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.