The Ledger, every show
Every statement that passed quotation and attribution checks, across all 44 shows. Pick shows below, then mix any filter with any other.
shows 




every show 44 of 44
Douwe Kiela: EU AI regulation will completely destroy innovation
“What Europe is going to try to do is overregulate everything and just completely destroy
innovation.”
Douwe Kiela: AI will displace the majority human workforce within 5-10 years
“If you look at how OpenAI and Anthropic and these places define AGI, it's as systems achieving capabilities that allow them to effectively do the work of humans for the majority of economically valuable human tasks. Then we're not that far away from it. And so…”
Kiela: DeepSeek's total development cost was at least 100x its $6M training
“So I would guess that they spent at least a hundred X The amount of that, that single training run, right?”
Kiela: AI systems will probably never reach 100 percent accuracy
“When are we getting to a hundred percent accuracy? And I had to give them the bad news that probably never.”
Kiela: Companies Will Stop Trying to Build In-House AI Within Years
“Very often they think they can do it in house.
And so I think that belief is probably going to go away in the next couple of years where people realize that this stuff is a little bit more difficult than they,
Initially thought.”
Kiela: Top-Down Enterprise AI Budgets Will Dry Up After Failed Pilots
“That money is temporary and it's going to dry up. And so it's, we're going to have a couple of cool pilots and demos, and then at some point it doesn't work and we will move on to the next hype train.”
Douwe Kiela: AI Model Parameter Sizes Will Stop Growing Rapidly
“I think Sam Altman had this interesting quote where he was saying that he thought models would stop growing in size. GPT-IV kind of hit this ceiling. I think that's probably right, but not really because size doesn't matter. It's just that data size matters ev…”
Douwe Kiela: GPT-4 Will Disrupt Data Annotators Like Mechanical Turk
“One of the use cases I've been seeing now for GPT-IV is actually that people are using it to generate data and then they're training on that data with cheaper models. So GPT-IV might end up disrupting, not knowledge workers necessarily, but it might just disru…”
Douwe Kiela: Artificial Specialized Intelligence will be achieved much faster than AGI
“And I think you're totally right that that's a much easier problem to solve much quicker and then slowly grow with the capabilities of these models.”
Douwe Kiela: GPT-4 will disrupt Mechanical Turk before knowledge workers
“And so, so GPT-IV might end up disrupting, not like knowledge workers necessarily, but it might just disrupt like mechanical Turk and is just a, an annotator on steroids.”
Douwe Kiela: OpenAI built a massive, underutilized data moat from ChatGPT
“And they haven't even really trained as far as I know on the data that comes out of ChatGPT going viral, right? So they had ChatGPT, it went viral. This led to this giant, giant data mode that they haven't even really used, used yet.”
Douwe Kiela: GPT-4's coding skills may be inflated by dataset contamination
“Data contamination where a bunch of these language models are trained on the things that they are being evaluated on. So GPT-IV looks like it's an amazing coder, but it might also just be trained on the data that it's evaluated on, which means that it's not ac…”
Douwe Kiela: Foundation model builders will rely on external AI security audits
“And so that's a, an interesting part of the market, but I don't think that, that the actual foundation model builders like open AI and contextual are going to build that technology in house. It's probably gonna be an external sort of audit.”
Douwe Kiela: AI hype disillusionment will dry up funding and threaten startups
“At some point there's going to be a disillusionment with the technology and then funding might dry up and then these places are really in trouble.”
Kiela: AI is heading toward specialized language models over generalists
“Where we're headed is that we will have more specialized language models.”
Kiela: FAISS was the first vector database
“In the initial paper, we used a vector database or a face. So the words vector database didn't exist at the time. But so face was the first vector database.”
Douwe Kiela: I co-created Retrieval-Augmented Generation (RAG) at Facebook AI
“We're specifically basing it on this retrieval augmented generation, which is something that me and my colleagues at fair came up with in.”
Douwe Kiela: Humans will never fully understand large neural network outputs
“So we're not going to be able to really know why a neural net, what does what it does at the scale that neural networks operate at.”
Douwe Kiela: Model-agnostic AI startups will have a competitive advantage
“At this particular point in time, probably yes, but just because the field is moving so incredibly quickly. And so I think that in the next year, we're going to see lots of other models coming out. And if you can have a language model, agnostic AI company that…”
Douwe Kiela: AutoGPT does not actually work despite the hype
“And there's an auto GPT thing that is going to change the world, but doesn't actually work.”
Kiela: FAIR was the first team to build a generative RAG model
“Why RAG became the way you name these things is because it's generative, right? So we were the first ones to have a generative model there.”
Kiela: Aligning language models needs only 100 examples via Anchored Preference Optimization
“So you can train on this when you only have like a hundred examples, you can really make a meaningful, meaningful difference.”
Kiela: By 2030, Workers Will Manage Fleets of AI Co-Pilots
“What I think is going to happen is that there's a couple of CEOs on the stage here, and I'm sure there's a couple of CEOs in the audience, so we will all be our own CEO of our little company of AI co-pilots that are going to be doing a lot of very boring, mund…”
Douwe Kiela: Meta's Original LLaMA Was Trained Entirely on Open Data
“So the LAMA model was not trained on any proprietary data. It was just trained on open data on the web.”
Douwe Kiela: Inability to delete LLM data creates GDPR compliance issues
“So we can't really remove information from them, which is kind of tricky from a GDPR perspective.”
Douwe Kiela: Generative AI was not ready for enterprise adoption post-ChatGPT
“We saw this kind of great excitement in the world, but at the same time, a lot of disappointment about it not being quite ready yet for real world adaption in, in enterprises where you actually want to use this technology.”
Douwe Kiela: Meta's LLaMA model used zero proprietary training data
“So the Lama model was not trained on any proprietary data. It was just trained on open data on the web”
Douwe Kiela: AI model protection will drive a new cybersecurity wave
“Oh, absolutely. Yeah. Yeah. So that's completely going to change everything.”
Douwe Kiela: The AI market will be tiered across multiple specialized models
“It's going to be lots of models at different parts different layers of this pyramid being used for different kinds of applications.”
Douwe Kiela: An enterprise AI adoption tidal wave is coming
“I think the tidal wave is coming. There are just big problems that we have to overcome, and these are the things I just talked about, right? So hallucination, attribution compliance up-to-dateness, data privacy, latency and I think the whole field is, is movin…”