The Ledger, every show
Every statement that passed quotation and attribution checks, across all 44 shows. Pick shows below, then mix any filter with any other.
shows 




every show 44 of 44
Fox: Most AI note-takers use AssemblyAI as voice infrastructure
“And so the reason, you know, most AI note takers are using assembly as the voice infrastructure is because our models are the most scalable, right?”
Fox: Callers hang up immediately if voice agents disclose they are AI
“If our customers, if they're building a voice agent and you disclose up front that you're an AI, people just hang up versus if you don't, people continue.”
Fox: AssemblyAI processes four times YouTube's daily audio volume on peak weeks
“And so now, you know, on a given week, on a peak week, there'll be something like over a hundred and twenty million conversations, voice conversations going through our platform, over two million hours of voice, which as of December of this past year was four,…”
Fox: AssemblyAI used Claude to rebuild and migrate its entire website
“Like, you know, the most recent example is we had Cloud rebuild our whole website off of Webflow and just deployed on Vercel.”
Fox: Customers abandon fine-tuned open-source voice models due to maintenance burdens
“We've seen some customers maybe six months ago or a year ago, they'll take an open source model and fine tune it or something. And then, you know, they're now like, okay, this is like, shit, this is like really behind and it's like nightmare to maintain and I …”
Fox: Voice will augment, not replace, screens over the next couple years
“And so I don't think that means that computers are just going to be like a screen and you just talk to it, because actually we'd get tired of talking, but I think it's going to be a dimension that's added to everything, and so that over the next couple years w…”
Fox: Voice AI became reliable data capture only in the past year
“Like for the first time probably ever in the past year, it's like a reliable form of data capture.”
Fox: In five years, kids will expect all devices to understand voice
“Across humanoid robots, consumer electronics, like, you're gonna see voice as this dimension that you just, like, expect And what, the example I think about is, you know, when you see pictures of, like, kids, like, trying to, like, swipe on TVs or something, b…”
Fox: Consumer voice hardware and software applications will surge within a year
“I think over the next year, we'll see a lot of, a lot more consumer applications, hardware and software, where voice is a core dimension. So, toys, games consumer electronics”
Most commercial speech models historically trained on roughly 50,000 hours of audio
“And our models prior, and most commercial speech recognition models trained on like, 50,000 hours”
AssemblyAI processes over 100 million audio files monthly via API
“We've processed, ah, yeah, it's like over a hundred million audio files a month that are flowing through the API, and that's growing pretty quickly.”
State-of-the-art speech recognition models still carry a 15% error rate
“State-of-the-art automatic speech recognition still has, like, a 15% error rate on a lot of data sets”
Fox: AssemblyAI weekly conversation volume grew 800 percent over three years
“One stat I was just looking at before I came over here was like the amount of weekly conversations that Assembly handles through our APIs every week is up over 800% over the last three years.”
Fox: AssemblyAI handles nearly 100 million daily API calls from developers
“There's, you know, almost a hundred million API calls a day coming against our API, about a million developers, a little over a million developers on the platform now.”
Fox: Multi-speaker acoustic disambiguation remains a major hurdle for humanoid robots
“One of the main problems the humanoid robots face today, because a lot of them are using our APIs If you have three people standing next to the robot, it doesn't know who to listen to. And it has a hard time disambiguating who's saying what.”
Fox: AssemblyAI's upcoming speech model trained on 12.5M audio hours
“The first model I ever trained for speech to text
Was on 10,000 hours of audio data.
And then now the model that we're going to release soon, that's trained on 12 and a half million hours of audio data.
It's on, you know, hundreds of TPUs.”
Fox entered YC as a solo founder 30 days past the deadline
“And so I submitted an application to YC for their summer batch. This is like summer of 20 17. And it was 30 days past the deadline. So I was like single founder past the deadline. There's no way I'm going to get in. You know, I was planning to just work on it,…”
AssemblyAI's post-YC seed round was funded entirely by angel investors
“No institutional funds invested. It was only angels that invested”
AssemblyAI has processed almost two billion audio files to date
“We've processed almost two billion audio files through our system.”