Everything Dylan Fox said on any show that made the record, most notable first. Each card names its show and opens the statement there.
Most commercial speech models historically trained on roughly 50,000 hours of audio
“And our models prior, and most commercial speech recognition models trained on like, 50,000 hours”
AssemblyAI is training its next speech model on ~4 million hours of audio
“We're actually training conformer two or what might call it 1.5, but whatever this accessory will be is training right now. And that's something around four million hours of labeled audio data.”
AssemblyAI processes over 100 million audio files monthly via API
“We've processed, ah, yeah, it's like over a hundred million audio files a month that are flowing through the API, and that's growing pretty quickly.”
State-of-the-art speech recognition models still carry a 15% error rate
“State-of-the-art automatic speech recognition still has, like, a 15% error rate on a lot of data sets”
Fox: AssemblyAI weekly conversation volume grew 800 percent over three years
“One stat I was just looking at before I came over here was like the amount of weekly conversations that Assembly handles through our APIs every week is up over 800% over the last three years.”
Fox: AssemblyAI handles nearly 100 million daily API calls from developers
“There's, you know, almost a hundred million API calls a day coming against our API, about a million developers, a little over a million developers on the platform now.”
Fox: Multi-speaker acoustic disambiguation remains a major hurdle for humanoid robots
“One of the main problems the humanoid robots face today, because a lot of them are using our APIs If you have three people standing next to the robot, it doesn't know who to listen to. And it has a hard time disambiguating who's saying what.”
Fox: AssemblyAI is developing on-device models for low-power hardware
“We're, for example, working on on-device models, so models that can run, like, on a phone or on a, you know, really low-powered piece of hardware you know, like a remote control for a TV or something”
Fox: AssemblyAI's upcoming speech model trained on 12.5M audio hours
“The first model I ever trained for speech to text
Was on 10,000 hours of audio data.
And then now the model that we're going to release soon, that's trained on 12 and a half million hours of audio data.
It's on, you know, hundreds of TPUs.”
Fox entered YC as a solo founder 30 days past the deadline
“And so I submitted an application to YC for their summer batch. This is like summer of 20 17. And it was 30 days past the deadline. So I was like single founder past the deadline. There's no way I'm going to get in. You know, I was planning to just work on it,…”
AssemblyAI's post-YC seed round was funded entirely by angel investors
“No institutional funds invested. It was only angels that invested”
AssemblyAI passed $1M ARR with fewer than ten employees
“We had passed, like, a million dollars in ARR. Like, we had, like, decent traction at that time. We were growing pretty quickly, and we were still a super small team. I think we were, like, sub-ten people.”
Fox: Post-YC founders should slow down and plan two-year horizons
“I would have advised my, myself to take a longer term view in the first year, especially coming out of YC, like slow down. Now that you're done YC, what do you want to accomplish over the next two years? Like, where do you want to be in two years? And then jus…”
AssemblyAI has processed almost two billion audio files to date
“We've processed almost two billion audio files through our system.”
Fox: AssemblyAI serves over 1,000 customers and tens of thousands of monthly developers
“We've got over a thousand customers tens of thousands a month of developers that are building with the API.”
AssemblyAI trained Conformer-1 on 650,000 hours of labeled audio data
“So we trained it on, like, 60 terabytes of audio data, like, labeled audio data. So it was, I think, something like 650,000 hours of audio data.”
AssemblyAI's JAX contribution sped up Whisper model training by 10x
“We actually, I think, published like, made a contribution to Jax to make it, like, 10 times faster to train Whisper.”
AssemblyAI employs about 40 full-time staff dedicated to improving speech models
“We got like 40 people full time working on this, you know, and you're gonna get all the benefits of that.”
Large organizations remain confused about who should manage internal AI projects
“I think right now, larger organizations sometimes are, like, still confused, like, who's gonna manage this AI project, you know?”
An early AssemblyAI customer used the API for 24/7 broadcast brand tracking
“Yeah, so one of our first customers was building this product where they were analyzing TV and radio stations, 24 seven, and then they were looking for certain brand mentions that were spoken and alerting those brands when their names were mentioned. That was …”