People, every show

Douwe Kiela

Research Scientist Director, Google DeepMind. On 3 shows, 4 appearances, plus 1 compilation re-air not counted. The Shows tab opens the full record on each.

scientistfounderexecutiveacademic@douwekiela ↗LinkedIn ↗cl.cam.ac.uk/~dk427 ↗Wikipedia ↗

Douwe Kiela is best known for pioneering Retrieval-Augmented Generation (RAG) and co-authoring the foundational 2020 paper while at Facebook AI Research. He later served as Head of Research at Hugging Face and co-founded Contextual AI to develop enterprise-grade contextual language models and grounded AI agents.

3shows
4appearances
66statements
9resolved
6supported
2contradicted
67%fully supported
3said about them ↓

Everything Douwe Kiela said on any show that made the record, most notable first. Each card names its show and opens the statement there.

20VC Opinion
Douwe Kiela: Incumbent AI firms promote existential risk to entrench regulatory advantage
“So the people who are pushing this narrative are really the people who are benefiting from this being the narrative. So these are the incumbent AI companies who are, who want to have either the market regulated. Right. Because then they benefit because they ca…”
Douwe Kiela Jun 30, 2023 ▶ 35:01 Douwe Kiela: Why Data Size Matters More Than Model Size; Why Open Source Isn't Going to Win | E1032 · 20VC with Harry Stebbings
MAD Opinion
Kiela: Core language model development is almost solved and plateauing
“It's not even really about language models anymore. That has almost been solved, right? That's kind of why you see things plateauing off a little bit as well.”
Douwe Kiela Mar 6, 2025 ▶ 11:20 Top AI Researcher on GPT 4.5, DeepSeek and Agentic RAG | Douwe Kiela, CEO, Contextual AI
MAD Opinion
Kiela: Long-context LLMs are inherently incredibly wasteful
“Long context models are inherently incredibly wasteful. You're paying for all this compute, and that's maybe why some of the companies that are trying to really sell long context model, long context window models, they will make more money from that, right?”
Douwe Kiela Mar 6, 2025 ▶ 29:23 Top AI Researcher on GPT 4.5, DeepSeek and Agentic RAG | Douwe Kiela, CEO, Contextual AI
20VC Opinion
Kiela: Believing open-source AI can match OpenAI is incredibly naive
“I would like it to be true that with open source, we could just keep up with all of that. But I think that's just incredibly naive. Open AI has this very deep understanding of how people want to use language models. Basically nobody else has, and they have thi…”
Douwe Kiela Nov 24, 2023 ▶ 7:57 The Ultimate AI Roundtable: What Happens Now in AI, Why Google are Vulnerable | E1085 · 20VC with Harry Stebbings
20VC Opinion
Douwe Kiela: The leaked 'no AI moat' Google memo is completely wrong
“I think in terms of data modes and maybe you, you've seen this come by actually, there was this Google memo from an internal Google employee who had written that open AI and Google have no mode. I think for me as a AI researcher, when I read that memo, I was l…”
Douwe Kiela Jun 30, 2023 ▶ 23:45 Douwe Kiela: Why Data Size Matters More Than Model Size; Why Open Source Isn't Going to Win | E1032 · 20VC with Harry Stebbings
20VC Opinion
Douwe Kiela: Getting struck by lightning is more likely than AI extinction
“Probably the chance of me getting hit by lightning, like right now is much higher than that happening.”
Douwe Kiela Jun 30, 2023 ▶ 34:32 Douwe Kiela: Why Data Size Matters More Than Model Size; Why Open Source Isn't Going to Win | E1032 · 20VC with Harry Stebbings
20VC Prediction Not checkable as stated
Douwe Kiela: EU AI regulation will completely destroy innovation
“What Europe is going to try to do is overregulate everything and just completely destroy innovation.”
Douwe Kiela Jun 30, 2023 ▶ 38:28 Douwe Kiela: Why Data Size Matters More Than Model Size; Why Open Source Isn't Going to Win | E1032 · 20VC with Harry Stebbings
20VC Prediction Open · timeframe Jun 2033
Douwe Kiela: AI will displace the majority human workforce within 5-10 years
“If you look at how OpenAI and Anthropic and these places define AGI, it's as systems achieving capabilities that allow them to effectively do the work of humans for the majority of economically valuable human tasks. Then we're not that far away from it. And so…”
Douwe Kiela Jun 30, 2023 ▶ 52:02 Douwe Kiela: Why Data Size Matters More Than Model Size; Why Open Source Isn't Going to Win | E1032 · 20VC with Harry Stebbings
20VC Opinion
Douwe Kiela: OpenAI and Anthropic are the Lycos and AltaVista of AI
“So if all the stars align, then OpenAI and Anthropic and all of these places, they had this great first generation technology. So they're kind of like the Lycos and Alta Vista of search engines. And the technology we have is more like PageRank. And that would …”
Douwe Kiela Jun 30, 2023 ▶ 53:20 Douwe Kiela: Why Data Size Matters More Than Model Size; Why Open Source Isn't Going to Win | E1032 · 20VC with Harry Stebbings
MAD Insight
Kiela: DeepSeek proved frontier AI models can rely on synthetic data
“We have kind of an existence proof now that it's actually not that hard to do this and so you don't need to invest all that much in, in data, and you can use synthetic data and get a pretty good model out of that”
Douwe Kiela Mar 6, 2025 ▶ 3:12 Top AI Researcher on GPT 4.5, DeepSeek and Agentic RAG | Douwe Kiela, CEO, Contextual AI
MAD Assertion Not checkable as stated
Kiela: DeepSeek's total development cost was at least 100x its $6M training
“So I would guess that they spent at least a hundred X The amount of that, that single training run, right?”
Douwe Kiela Mar 6, 2025 ▶ 9:58 Top AI Researcher on GPT 4.5, DeepSeek and Agentic RAG | Douwe Kiela, CEO, Contextual AI
MAD Insight
Kiela: Fine-tuning cannot inject new knowledge into AI models
“One common misconception about fine tuning is a lot of people think that you can inject new knowledge into a model using fine tuning. And that is not true.”
Douwe Kiela Mar 6, 2025 ▶ 28:19 Top AI Researcher on GPT 4.5, DeepSeek and Agentic RAG | Douwe Kiela, CEO, Contextual AI
MAD Insight
Kiela: Advanced RAG systems break down when scaling to a million PDFs
“You can build a very awesome demo on a couple of PDFs and things will probably work. But then you have to scale it up to a million PDFs, and then everything breaks down. And the reason for that is that a lot of these kind of advanced RAG systems still actually…”
Douwe Kiela Mar 6, 2025 ▶ 35:14 Top AI Researcher on GPT 4.5, DeepSeek and Agentic RAG | Douwe Kiela, CEO, Contextual AI
MAD Prediction Not checkable as stated
Kiela: AI systems will probably never reach 100 percent accuracy
“When are we getting to a hundred percent accuracy? And I had to give them the bad news that probably never.”
Douwe Kiela Mar 6, 2025 ▶ 48:21 Top AI Researcher on GPT 4.5, DeepSeek and Agentic RAG | Douwe Kiela, CEO, Contextual AI
SAASTR Prediction Not checkable as stated
Kiela: Companies Will Stop Trying to Build In-House AI Within Years
“Very often they think they can do it in house. And so I think that belief is probably going to go away in the next couple of years where people realize that this stuff is a little bit more difficult than they, Initially thought.”
Douwe Kiela Jan 17, 2024 ▶ 16:27 How Enterprise Companies are Buying AI (or Not) with ContextualAI, Anthropic, and Glean
SAASTR Prediction Not checkable as stated
Kiela: Top-Down Enterprise AI Budgets Will Dry Up After Failed Pilots
“That money is temporary and it's going to dry up. And so it's, we're going to have a couple of cool pilots and demos, and then at some point it doesn't work and we will move on to the next hype train.”
Douwe Kiela Jan 17, 2024 ▶ 17:02 How Enterprise Companies are Buying AI (or Not) with ContextualAI, Anthropic, and Glean
SAASTR Insight
Kiela: Enterprise AI Demands End-to-End Retrieval Models Over Parametric Monoliths
“And so we think that's not what we now call a language model. So it's not one big parametric monster. It's something a bit more elegant that has the retrieval kind of built in. It's a retrieval augmented language model and is strained end to end so that it can…”
Douwe Kiela Jan 17, 2024 ▶ 20:14 How Enterprise Companies are Buying AI (or Not) with ContextualAI, Anthropic, and Glean
SAASTR Insight
Kiela: Enterprises do not need model fine-tuning when RAG is available
“You don't have to fine tune your model. It feels very intuitive. We have this great data set. We own it. It's our data. So we need to do something useful with it. So we need to fine tune our own language model. And so the companies who are offering that servic…”
Douwe Kiela Jan 17, 2024 ▶ 23:58 How Enterprise Companies are Buying AI (or Not) with ContextualAI, Anthropic, and Glean
20VC Prediction Not checkable as stated
Douwe Kiela: AI Model Parameter Sizes Will Stop Growing Rapidly
“I think Sam Altman had this interesting quote where he was saying that he thought models would stop growing in size. GPT-IV kind of hit this ceiling. I think that's probably right, but not really because size doesn't matter. It's just that data size matters ev…”
Douwe Kiela Sep 15, 2023 ▶ 13:29 20VC: The Biggest AI Leaders on What Matters More; Model Size or Data Size & Where Does The Value in AI Accrue; to Startups or to Incumbents
20VC Prediction Not checkable as stated
Douwe Kiela: GPT-4 Will Disrupt Data Annotators Like Mechanical Turk
“One of the use cases I've been seeing now for GPT-IV is actually that people are using it to generate data and then they're training on that data with cheaper models. So GPT-IV might end up disrupting, not knowledge workers necessarily, but it might just disru…”
Douwe Kiela Sep 15, 2023 ▶ 20:19 20VC: The Biggest AI Leaders on What Matters More; Model Size or Data Size & Where Does The Value in AI Accrue; to Startups or to Incumbents
20VC What-if
Douwe Kiela: Current AI breakthroughs would not happen without Meta's PyTorch
“So PyTorch really without PyTorch, none of this stuff would be happening right now. And so it's really like fundamental for all of the AI breakthroughs.”
Douwe Kiela Jun 30, 2023 ▶ 4:59 Douwe Kiela: Why Data Size Matters More Than Model Size; Why Open Source Isn't Going to Win | E1032 · 20VC with Harry Stebbings
20VC Prediction Not checkable as stated
Douwe Kiela: Artificial Specialized Intelligence will be achieved much faster than AGI
“And I think you're totally right that that's a much easier problem to solve much quicker and then slowly grow with the capabilities of these models.”
Douwe Kiela Jun 30, 2023 ▶ 15:51 Douwe Kiela: Why Data Size Matters More Than Model Size; Why Open Source Isn't Going to Win | E1032 · 20VC with Harry Stebbings
20VC Insight
Douwe Kiela: Data size matters more than model size for AI performance
“It's just that data size matters even more than model size. And I think the Lama paper out of Meta really brilliantly showed this. Where if you train a smaller model on more data for longer, then you get a better model. So you get more bang for your buck if yo…”
Douwe Kiela Jun 30, 2023 ▶ 16:37 Douwe Kiela: Why Data Size Matters More Than Model Size; Why Open Source Isn't Going to Win | E1032 · 20VC with Harry Stebbings
20VC Prediction Not checkable as stated
Douwe Kiela: GPT-4 will disrupt Mechanical Turk before knowledge workers
“And so, so GPT-IV might end up disrupting, not like knowledge workers necessarily, but it might just disrupt like mechanical Turk and is just a, an annotator on steroids.”
Douwe Kiela Jun 30, 2023 ▶ 21:15 Douwe Kiela: Why Data Size Matters More Than Model Size; Why Open Source Isn't Going to Win | E1032 · 20VC with Harry Stebbings

Show 24statements(42 left)

The other half of the tape: Douwe Kiela's own voice is left out of every number here. Other people bring the name up 3 times in 2 episodes across the shows. every mention, with the transcript →

Who brings them up most Harry Stebbings 2Matt Turck 1

Every mention by year

tap a year for its mentions
001121202320242025episodesmentions
011202320242025episodes it came up in
0010.521202320242025episodesmentions per episode

20VC 2the MAD Podcast 1

One line per show, most statements first. The link opens Douwe's full record on that show: the calibration, argument clarity, speaking style and every statement made there.

ShowRole thereEpsStatementsRecord
20VCLEDGER Research Scientist Director, Google DeepMind 2 +1 44 86% 6/7 full record on 20VC →
MADLEDGER Research Scientist Director, Google DeepMind 1 15 0% 0/2 full record on the MAD Podcast →
SAASTRLEDGER Research Scientist Director, Google DeepMind 1 7 full record on the Official SaaStr Podcast →
Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.