Jun 30, 2023 · 53m · news

Douwe Kiela: Why Data Size Matters More Than Model Size; Why Open Source Isn't Going to Win | E1032 · 20VC with Harry Stebbings

Douwe Kiela · 37m spoken Harry Stebbings · 12m spoken
0:00 / 0:00
▶ Watch on YouTube →

gold bands on the timeline = statements, start to end. Hover to read, click to jump. CC turns on captions

In this episode of 20VC, host Harry Stebbings interviews Douwe Kiela, co-founder and CEO of Contextual AI, who argues that the AI revolution is still in its early stages and explains why data optimization, Retrieval-Augmented Generation (RAG), and Artificial Specialized Intelligence (ASI) represent the future of enterprise AI.

How this conversation actually went

Every chapter scored 0–10 on four independent dynamics. Hover any point for the reasoning behind the score. Harry holds 25.1% of the talking time here. How this is scored →

Harry as informed peer 4.6 Guest teaching 5.4 Guest disagreement 3.1 Harry pushing back 3.6
05100:0015:0030:0045:003:00–5:08 · Harry as informed peer 5/10 Facebook AI Research (FAIR) and the Impact of Open Source Host interrupts to challenge the guest on value capture, questioning whether Meta lost value by over-indexing on open source. Guest politely reframes around Meta's broader corporate mission beyond purely maximizing revenue.5:08–7:54 · Harry as informed peer 3/10 Experience at Hugging Face and Community Building Host asks about Hugging Face's company culture and shares an anecdote about their custom swag policies before transitioning to the founding of Contextual.7:54–9:59 · Harry as informed peer 2/10 LLMs Enterprise Readiness and RAG Architecture Guest details why LLMs fail enterprise requirements and explains how Retrieval Augmented Generation (RAG) decouples memory from generation to solve hallucination and privacy.9:59–12:37 · Harry as informed peer 6/10 Model Transparency and RAG's Attribution Power Host probes model transparency and challenges guest with a counter-quote from Emad Mostaque claiming hallucinations are a feature, not a bug. Guest reframes the quote with nuance.12:37–14:48 · Harry as informed peer 5/10 Model Agnosticism in a Rapidly Evolving Field Host presents a guest thesis that model-agnostic startups will win. Guest agrees short-term but explains Contextual's strategic choice to focus on Specialized Intelligence (ASI) over AGI.14:48–18:35 · Harry as informed peer 3/10 AGI vs. Artificial Specialized Intelligence (ASI) Guest educates the host on model scaling trade-offs, citing Sam Altman and Meta's Llama paper to show data volume matters more than parameter count for optimal performance.18:35–23:08 · Harry as informed peer 6/10 The Role of Proprietary and Pre-trained Data Host cites VC rationale for rejecting AI wrappers lacking proprietary data. Guest breaks down the technical steps to build ChatGPT (Pre-training, SFT, RLHF) and explains data distillation using GPT-4.23:08–27:55 · Harry as informed peer 5/10 Moat Debate: OpenAI vs. Open Source Host raises the famous leaked Google memo claiming OpenAI and Google have no moats. Guest strongly rejects the memo's author as clueless and labels open source parity expectations naive.27:55–33:13 · Harry as informed peer 5/10 Model Security and Next-Generation Cybersecurity Host asks whether AI security tools will be built natively by model providers or third-party security vendors. Guest categorizes the AI landscape using a three-tier pyramid model.33:13–36:38 · Harry as informed peer 4/10 Existential Risk, Self-Interest, and the AI Pause Petition Guest ridicules Elon Musk's AI pause petition as self-serving and calls x-risk narratives incumbent fear-mongering designed to secure regulatory capture.36:38–41:09 · Harry as informed peer 5/10 AI Regulation and the Chasm of Knowledge Host highlights proposed EU rules threatening model usage. Guest criticizes the tech literacy gap in regulatory bodies, citing infamous Congressional hearings, and warns EU over-regulation will destroy innovation.41:09–43:11 · Harry as informed peer 6/10 Enterprise AI Adoption and Data Sovereignty Host cites a major European bank refusing to send data off-premise. Guest explains how Contextual's hybrid architecture keeps data inside the customer's virtual private cloud.43:11–45:35 · Harry as informed peer 5/10 Decoupling the Data Plane and Model Plane Host pushes guest on whether VC valuations for AI startups have become absurd. Guest defends high valuations based on massive potential market payoffs.45:35–46:24 · Harry as informed peer 4/10 Winners and Losers of the Tech Giants Host asks for a breakdown of winner and loser tech giants. Guest lauds Microsoft's Open AI pivot while noting Apple's lack of progress in AI.3:00–5:08 · Guest teaching 3/10 Facebook AI Research (FAIR) and the Impact of Open Source Host interrupts to challenge the guest on value capture, questioning whether Meta lost value by over-indexing on open source. Guest politely reframes around Meta's broader corporate mission beyond purely maximizing revenue.5:08–7:54 · Guest teaching 1/10 Experience at Hugging Face and Community Building Host asks about Hugging Face's company culture and shares an anecdote about their custom swag policies before transitioning to the founding of Contextual.7:54–9:59 · Guest teaching 6/10 LLMs Enterprise Readiness and RAG Architecture Guest details why LLMs fail enterprise requirements and explains how Retrieval Augmented Generation (RAG) decouples memory from generation to solve hallucination and privacy.9:59–12:37 · Guest teaching 6/10 Model Transparency and RAG's Attribution Power Host probes model transparency and challenges guest with a counter-quote from Emad Mostaque claiming hallucinations are a feature, not a bug. Guest reframes the quote with nuance.12:37–14:48 · Guest teaching 5/10 Model Agnosticism in a Rapidly Evolving Field Host presents a guest thesis that model-agnostic startups will win. Guest agrees short-term but explains Contextual's strategic choice to focus on Specialized Intelligence (ASI) over AGI.14:48–18:35 · Guest teaching 7/10 AGI vs. Artificial Specialized Intelligence (ASI) Guest educates the host on model scaling trade-offs, citing Sam Altman and Meta's Llama paper to show data volume matters more than parameter count for optimal performance.18:35–23:08 · Guest teaching 7/10 The Role of Proprietary and Pre-trained Data Host cites VC rationale for rejecting AI wrappers lacking proprietary data. Guest breaks down the technical steps to build ChatGPT (Pre-training, SFT, RLHF) and explains data distillation using GPT-4.23:08–27:55 · Guest teaching 7/10 Moat Debate: OpenAI vs. Open Source Host raises the famous leaked Google memo claiming OpenAI and Google have no moats. Guest strongly rejects the memo's author as clueless and labels open source parity expectations naive.27:55–33:13 · Guest teaching 6/10 Model Security and Next-Generation Cybersecurity Host asks whether AI security tools will be built natively by model providers or third-party security vendors. Guest categorizes the AI landscape using a three-tier pyramid model.33:13–36:38 · Guest teaching 6/10 Existential Risk, Self-Interest, and the AI Pause Petition Guest ridicules Elon Musk's AI pause petition as self-serving and calls x-risk narratives incumbent fear-mongering designed to secure regulatory capture.36:38–41:09 · Guest teaching 5/10 AI Regulation and the Chasm of Knowledge Host highlights proposed EU rules threatening model usage. Guest criticizes the tech literacy gap in regulatory bodies, citing infamous Congressional hearings, and warns EU over-regulation will destroy innovation.41:09–43:11 · Guest teaching 5/10 Enterprise AI Adoption and Data Sovereignty Host cites a major European bank refusing to send data off-premise. Guest explains how Contextual's hybrid architecture keeps data inside the customer's virtual private cloud.43:11–45:35 · Guest teaching 6/10 Decoupling the Data Plane and Model Plane Host pushes guest on whether VC valuations for AI startups have become absurd. Guest defends high valuations based on massive potential market payoffs.45:35–46:24 · Guest teaching 5/10 Winners and Losers of the Tech Giants Host asks for a breakdown of winner and loser tech giants. Guest lauds Microsoft's Open AI pivot while noting Apple's lack of progress in AI.3:00–5:08 · Guest disagreement 2/10 Facebook AI Research (FAIR) and the Impact of Open Source Host interrupts to challenge the guest on value capture, questioning whether Meta lost value by over-indexing on open source. Guest politely reframes around Meta's broader corporate mission beyond purely maximizing revenue.5:08–7:54 · Guest disagreement 1/10 Experience at Hugging Face and Community Building Host asks about Hugging Face's company culture and shares an anecdote about their custom swag policies before transitioning to the founding of Contextual.7:54–9:59 · Guest disagreement 1/10 LLMs Enterprise Readiness and RAG Architecture Guest details why LLMs fail enterprise requirements and explains how Retrieval Augmented Generation (RAG) decouples memory from generation to solve hallucination and privacy.9:59–12:37 · Guest disagreement 3/10 Model Transparency and RAG's Attribution Power Host probes model transparency and challenges guest with a counter-quote from Emad Mostaque claiming hallucinations are a feature, not a bug. Guest reframes the quote with nuance.12:37–14:48 · Guest disagreement 2/10 Model Agnosticism in a Rapidly Evolving Field Host presents a guest thesis that model-agnostic startups will win. Guest agrees short-term but explains Contextual's strategic choice to focus on Specialized Intelligence (ASI) over AGI.14:48–18:35 · Guest disagreement 2/10 AGI vs. Artificial Specialized Intelligence (ASI) Guest educates the host on model scaling trade-offs, citing Sam Altman and Meta's Llama paper to show data volume matters more than parameter count for optimal performance.18:35–23:08 · Guest disagreement 2/10 The Role of Proprietary and Pre-trained Data Host cites VC rationale for rejecting AI wrappers lacking proprietary data. Guest breaks down the technical steps to build ChatGPT (Pre-training, SFT, RLHF) and explains data distillation using GPT-4.23:08–27:55 · Guest disagreement 7/10 Moat Debate: OpenAI vs. Open Source Host raises the famous leaked Google memo claiming OpenAI and Google have no moats. Guest strongly rejects the memo's author as clueless and labels open source parity expectations naive.27:55–33:13 · Guest disagreement 2/10 Model Security and Next-Generation Cybersecurity Host asks whether AI security tools will be built natively by model providers or third-party security vendors. Guest categorizes the AI landscape using a three-tier pyramid model.33:13–36:38 · Guest disagreement 8/10 Existential Risk, Self-Interest, and the AI Pause Petition Guest ridicules Elon Musk's AI pause petition as self-serving and calls x-risk narratives incumbent fear-mongering designed to secure regulatory capture.36:38–41:09 · Guest disagreement 5/10 AI Regulation and the Chasm of Knowledge Host highlights proposed EU rules threatening model usage. Guest criticizes the tech literacy gap in regulatory bodies, citing infamous Congressional hearings, and warns EU over-regulation will destroy innovation.41:09–43:11 · Guest disagreement 3/10 Enterprise AI Adoption and Data Sovereignty Host cites a major European bank refusing to send data off-premise. Guest explains how Contextual's hybrid architecture keeps data inside the customer's virtual private cloud.43:11–45:35 · Guest disagreement 3/10 Decoupling the Data Plane and Model Plane Host pushes guest on whether VC valuations for AI startups have become absurd. Guest defends high valuations based on massive potential market payoffs.45:35–46:24 · Guest disagreement 3/10 Winners and Losers of the Tech Giants Host asks for a breakdown of winner and loser tech giants. Guest lauds Microsoft's Open AI pivot while noting Apple's lack of progress in AI.3:00–5:08 · Harry pushing back 5/10 Facebook AI Research (FAIR) and the Impact of Open Source Host interrupts to challenge the guest on value capture, questioning whether Meta lost value by over-indexing on open source. Guest politely reframes around Meta's broader corporate mission beyond purely maximizing revenue.5:08–7:54 · Harry pushing back 1/10 Experience at Hugging Face and Community Building Host asks about Hugging Face's company culture and shares an anecdote about their custom swag policies before transitioning to the founding of Contextual.7:54–9:59 · Harry pushing back 2/10 LLMs Enterprise Readiness and RAG Architecture Guest details why LLMs fail enterprise requirements and explains how Retrieval Augmented Generation (RAG) decouples memory from generation to solve hallucination and privacy.9:59–12:37 · Harry pushing back 5/10 Model Transparency and RAG's Attribution Power Host probes model transparency and challenges guest with a counter-quote from Emad Mostaque claiming hallucinations are a feature, not a bug. Guest reframes the quote with nuance.12:37–14:48 · Harry pushing back 4/10 Model Agnosticism in a Rapidly Evolving Field Host presents a guest thesis that model-agnostic startups will win. Guest agrees short-term but explains Contextual's strategic choice to focus on Specialized Intelligence (ASI) over AGI.14:48–18:35 · Harry pushing back 3/10 AGI vs. Artificial Specialized Intelligence (ASI) Guest educates the host on model scaling trade-offs, citing Sam Altman and Meta's Llama paper to show data volume matters more than parameter count for optimal performance.18:35–23:08 · Harry pushing back 4/10 The Role of Proprietary and Pre-trained Data Host cites VC rationale for rejecting AI wrappers lacking proprietary data. Guest breaks down the technical steps to build ChatGPT (Pre-training, SFT, RLHF) and explains data distillation using GPT-4.23:08–27:55 · Harry pushing back 5/10 Moat Debate: OpenAI vs. Open Source Host raises the famous leaked Google memo claiming OpenAI and Google have no moats. Guest strongly rejects the memo's author as clueless and labels open source parity expectations naive.27:55–33:13 · Harry pushing back 3/10 Model Security and Next-Generation Cybersecurity Host asks whether AI security tools will be built natively by model providers or third-party security vendors. Guest categorizes the AI landscape using a three-tier pyramid model.33:13–36:38 · Harry pushing back 4/10 Existential Risk, Self-Interest, and the AI Pause Petition Guest ridicules Elon Musk's AI pause petition as self-serving and calls x-risk narratives incumbent fear-mongering designed to secure regulatory capture.36:38–41:09 · Harry pushing back 3/10 AI Regulation and the Chasm of Knowledge Host highlights proposed EU rules threatening model usage. Guest criticizes the tech literacy gap in regulatory bodies, citing infamous Congressional hearings, and warns EU over-regulation will destroy innovation.41:09–43:11 · Harry pushing back 4/10 Enterprise AI Adoption and Data Sovereignty Host cites a major European bank refusing to send data off-premise. Guest explains how Contextual's hybrid architecture keeps data inside the customer's virtual private cloud.43:11–45:35 · Harry pushing back 5/10 Decoupling the Data Plane and Model Plane Host pushes guest on whether VC valuations for AI startups have become absurd. Guest defends high valuations based on massive potential market payoffs.45:35–46:24 · Harry pushing back 2/10 Winners and Losers of the Tech Giants Host asks for a breakdown of winner and loser tech giants. Guest lauds Microsoft's Open AI pivot while noting Apple's lack of progress in AI.

speaking balance: gold is Harry, purple is the guest (3 minute bins)

0:00 · Harry 14% · guest 86%0:00 · Harry 14% · guest 86%3:00 · Harry 31.7% · guest 68.3%3:00 · Harry 31.7% · guest 68.3%6:00 · Harry 29.8% · guest 70.2%6:00 · Harry 29.8% · guest 70.2%9:00 · Harry 29.9% · guest 70.1%9:00 · Harry 29.9% · guest 70.1%12:00 · Harry 28.9% · guest 71.1%12:00 · Harry 28.9% · guest 71.1%15:00 · Harry 31.9% · guest 68.1%15:00 · Harry 31.9% · guest 68.1%18:00 · Harry 30.8% · guest 69.2%18:00 · Harry 30.8% · guest 69.2%21:00 · Harry 13.3% · guest 86.7%21:00 · Harry 13.3% · guest 86.7%24:00 · Harry 22.3% · guest 77.7%24:00 · Harry 22.3% · guest 77.7%27:00 · Harry 24.1% · guest 75.9%27:00 · Harry 24.1% · guest 75.9%30:00 · Harry 27.8% · guest 72.2%30:00 · Harry 27.8% · guest 72.2%33:00 · Harry 7.2% · guest 92.8%33:00 · Harry 7.2% · guest 92.8%36:00 · Harry 25.7% · guest 74.3%36:00 · Harry 25.7% · guest 74.3%39:00 · Harry 47.5% · guest 52.5%39:00 · Harry 47.5% · guest 52.5%42:00 · Harry 29.1% · guest 70.9%42:00 · Harry 29.1% · guest 70.9%45:00 · Harry 26.4% · guest 73.6%45:00 · Harry 26.4% · guest 73.6%48:00 · Harry 14.4% · guest 85.6%48:00 · Harry 14.4% · guest 85.6%51:00 · Harry 17.6% · guest 82.4%51:00 · Harry 17.6% · guest 82.4%
Sharpest disagreement ▶ 33:03 Guest dismisses Elon Musk's petition

Douwe forcefully rejects Musk's 6-month pause petition, sarcastically re-framing it as Musk asking everyone else to pause so he can catch up.

Hardest push from Harry ▶ 4:17 Host challenges open source value capture

Harry directly interrupts Douwe to reject the premise that Meta's open-sourcing strategy was economically wise, arguing value was given away rather than enshrined.

Biggest teaching moment ▶ 21:54 Guest explains 3 steps to build ChatGPT

Douwe clearly and concisely educates the host on the exact technical pipeline required to train instruction-following models: pre-training, supervised fine-tuning, and RLHF.

Harry holds his own ▶ 11:33 Host counters with Emad Mostaque quote

Harry brings deep domain preparation by pitting Stability AI founder Emad Mostaque's claim that hallucinations are a feature against Douwe's claim that they are an enterprise dealbreaker.

the scores for every segment, with the reasoning behind each
ChapterTopicHarry as informed peerGuest teachingGuest disagreementHarry pushing backWhy
Facebook AI Research (FAIR) and the Impact of Open Source 5325 Host interrupts to challenge the guest on value capture, questioning whether Meta lost value by over-indexing on open source. Guest politely reframes around Meta's broader corporate mission beyond purely maximizing revenue.
Experience at Hugging Face and Community Building 3111 Host asks about Hugging Face's company culture and shares an anecdote about their custom swag policies before transitioning to the founding of Contextual.
LLMs Enterprise Readiness and RAG Architecture 2612 Guest details why LLMs fail enterprise requirements and explains how Retrieval Augmented Generation (RAG) decouples memory from generation to solve hallucination and privacy.
Model Transparency and RAG's Attribution Power 6635 Host probes model transparency and challenges guest with a counter-quote from Emad Mostaque claiming hallucinations are a feature, not a bug. Guest reframes the quote with nuance.
Model Agnosticism in a Rapidly Evolving Field 5524 Host presents a guest thesis that model-agnostic startups will win. Guest agrees short-term but explains Contextual's strategic choice to focus on Specialized Intelligence (ASI) over AGI.
AGI vs. Artificial Specialized Intelligence (ASI) 3723 Guest educates the host on model scaling trade-offs, citing Sam Altman and Meta's Llama paper to show data volume matters more than parameter count for optimal performance.
The Role of Proprietary and Pre-trained Data 6724 Host cites VC rationale for rejecting AI wrappers lacking proprietary data. Guest breaks down the technical steps to build ChatGPT (Pre-training, SFT, RLHF) and explains data distillation using GPT-4.
Moat Debate: OpenAI vs. Open Source 5775 Host raises the famous leaked Google memo claiming OpenAI and Google have no moats. Guest strongly rejects the memo's author as clueless and labels open source parity expectations naive.
Model Security and Next-Generation Cybersecurity 5623 Host asks whether AI security tools will be built natively by model providers or third-party security vendors. Guest categorizes the AI landscape using a three-tier pyramid model.
Existential Risk, Self-Interest, and the AI Pause Petition 4684 Guest ridicules Elon Musk's AI pause petition as self-serving and calls x-risk narratives incumbent fear-mongering designed to secure regulatory capture.
AI Regulation and the Chasm of Knowledge 5553 Host highlights proposed EU rules threatening model usage. Guest criticizes the tech literacy gap in regulatory bodies, citing infamous Congressional hearings, and warns EU over-regulation will destroy innovation.
Enterprise AI Adoption and Data Sovereignty 6534 Host cites a major European bank refusing to send data off-premise. Guest explains how Contextual's hybrid architecture keeps data inside the customer's virtual private cloud.
Decoupling the Data Plane and Model Plane 5635 Host pushes guest on whether VC valuations for AI startups have become absurd. Guest defends high valuations based on massive potential market payoffs.
Winners and Losers of the Tech Giants 4532 Host asks for a breakdown of winner and loser tech giants. Guest lauds Microsoft's Open AI pivot while noting Apple's lack of progress in AI.

Statements from this episode (38)

Assertion Supported
Douwe Kiela: LLMs generate inaccurate statements with high confidence
“These models make things up with very high confidence.”
Douwe Kiela Jun 30, 2023 ▶ 8:16
Assertion Supported
Douwe Kiela: Inability to delete LLM data creates GDPR compliance issues
“So we can't really remove information from them, which is kind of tricky from a GDPR perspective.”
Douwe Kiela Jun 30, 2023 ▶ 0:13
Insight
Douwe Kiela: Applied AI research creates more value than theoretical tangents
“Having a very clear Real world application for the research that you're doing makes it much more valuable than going off on a tangent and maybe being a bit too far ahead of the rest of the field.”
Douwe Kiela Jun 30, 2023 ▶ 3:43
What-if
Douwe Kiela: Current AI breakthroughs would not happen without Meta's PyTorch
“So PyTorch really without PyTorch, none of this stuff would be happening right now. And so it's really like fundamental for all of the AI breakthroughs.”
Douwe Kiela Jun 30, 2023 ▶ 4:59
Opinion
Douwe Kiela: Hugging Face excels at marketing and community building
“What really impressed me actually is how good they are just at marketing and branding and community building.”
Douwe Kiela Jun 30, 2023 ▶ 5:58
Assertion Not checkable as stated
Douwe Kiela: Generative AI was not ready for enterprise adoption post-ChatGPT
“We saw this kind of great excitement in the world, but at the same time, a lot of disappointment about it not being quite ready yet for real world adaption in, in enterprises where you actually want to use this technology.”
Douwe Kiela Jun 30, 2023 ▶ 7:20
Opinion
Douwe Kiela: The generative AI market is far from settled
“We think it's still very early innings in the game, so I think a lot of people sometimes think that it's you know, the game has been played, but it's just getting started.”
Douwe Kiela Jun 30, 2023 ▶ 7:45
Assertion Supported
Douwe Kiela: I co-created Retrieval-Augmented Generation (RAG) at Facebook AI
“We're specifically basing it on this retrieval augmented generation, which is something that me and my colleagues at fair came up with in.”
Douwe Kiela Jun 30, 2023 ▶ 9:12
Insight
Douwe Kiela: Decoupling LLM memory reduces hallucinations and enables data deletion
“You decouple the memory from the generative capacity of the large language model, and this allows you to ground the generations from the language model into things you retrieved. In your memory, essentially. So you get much less hallucination. You get attribut…”
Douwe Kiela Jun 30, 2023 ▶ 9:21
Assertion Not checkable as stated
Douwe Kiela: Humans will never fully understand large neural network outputs
“So we're not going to be able to really know why a neural net, what does what it does at the scale that neural networks operate at.”
Douwe Kiela Jun 30, 2023 ▶ 10:33
Insight
Douwe Kiela: AI hallucinations aid creativity but destroy enterprise value
“So I think in some cases it is a feature. If you want to use a language model for creative writing and if you want it to be really, really creative, then you probably want it to hallucinate. So in a way it's a spectrum of groundedness and hallucination where i…”
Douwe Kiela Jun 30, 2023 ▶ 12:01
Prediction Not checkable as stated
Douwe Kiela: Model-agnostic AI startups will have a competitive advantage
“At this particular point in time, probably yes, but just because the field is moving so incredibly quickly. And so I think that in the next year, we're going to see lots of other models coming out. And if you can have a language model, agnostic AI company that…”
Douwe Kiela Jun 30, 2023 ▶ 12:57
Opinion
Douwe Kiela: Anthropic and OpenAI are consumer-facing companies chasing AGI
“So if you look at anthropic and open AI, I think they're really chasing for this idea of AGI and they're relatively consumer facing.”
Douwe Kiela Jun 30, 2023 ▶ 14:14
Opinion
Douwe Kiela: The AI industry lacks a clear definition of AGI
“It's slightly better defined now, but one of the big issues has always been that we don't really know what that term even means.”
Douwe Kiela Jun 30, 2023 ▶ 15:34
Prediction Not checkable as stated
Douwe Kiela: Artificial Specialized Intelligence will be achieved much faster than AGI
“And I think you're totally right that that's a much easier problem to solve much quicker and then slowly grow with the capabilities of these models.”
Douwe Kiela Jun 30, 2023 ▶ 15:51
Insight
Douwe Kiela: Data size matters more than model size for AI performance
“It's just that data size matters even more than model size. And I think the Lama paper out of Meta really brilliantly showed this. Where if you train a smaller model on more data for longer, then you get a better model. So you get more bang for your buck if yo…”
Douwe Kiela Jun 30, 2023 ▶ 16:37
Assertion Supported
Douwe Kiela: Meta's LLaMA model used zero proprietary training data
“So the Lama model was not trained on any proprietary data. It was just trained on open data on the web”
Douwe Kiela Jun 30, 2023 ▶ 18:38
Prediction Not checkable as stated
Douwe Kiela: GPT-4 will disrupt Mechanical Turk before knowledge workers
“And so, so GPT-IV might end up disrupting, not like knowledge workers necessarily, but it might just disrupt like mechanical Turk and is just a, an annotator on steroids.”
Douwe Kiela Jun 30, 2023 ▶ 21:15
Assertion Not checkable as stated
Douwe Kiela: OpenAI built a massive, underutilized data moat from ChatGPT
“And they haven't even really trained as far as I know on the data that comes out of ChatGPT going viral, right? So they had ChatGPT, it went viral. This led to this giant, giant data mode that they haven't even really used, used yet.”
Douwe Kiela Jun 30, 2023 ▶ 23:28
Opinion
Douwe Kiela: The leaked 'no AI moat' Google memo is completely wrong
“I think in terms of data modes and maybe you, you've seen this come by actually, there was this Google memo from an internal Google employee who had written that open AI and Google have no mode. I think for me as a AI researcher, when I read that memo, I was l…”
Douwe Kiela Jun 30, 2023 ▶ 23:45
Insight
Douwe Kiela: Huge market opportunity exists for an AI evaluation rating agency
“I think there's a giant opportunity in the market actually for. A startup or several startups becoming like the Moody's or the SMP sort of you know the folks who evaluate the quality of AI for specific use cases because nobody really knows.”
Douwe Kiela Jun 30, 2023 ▶ 25:32
Assertion Not checkable as stated
Douwe Kiela: GPT-4's coding skills may be inflated by dataset contamination
“Data contamination where a bunch of these language models are trained on the things that they are being evaluated on. So GPT-IV looks like it's an amazing coder, but it might also just be trained on the data that it's evaluated on, which means that it's not ac…”
Douwe Kiela Jun 30, 2023 ▶ 25:54
Prediction Held up
Douwe Kiela: AI model protection will drive a new cybersecurity wave
“Oh, absolutely. Yeah. Yeah. So that's completely going to change everything.”
Douwe Kiela Jun 30, 2023 ▶ 28:09
Prediction Partly held up
Douwe Kiela: Foundation model builders will rely on external AI security audits
“And so that's a, an interesting part of the market, but I don't think that, that the actual foundation model builders like open AI and contextual are going to build that technology in house. It's probably gonna be an external sort of audit.”
Douwe Kiela Jun 30, 2023 ▶ 29:24
Prediction Not checkable as stated
Douwe Kiela: The AI market will be tiered across multiple specialized models
“It's going to be lots of models at different parts different layers of this pyramid being used for different kinds of applications.”
Douwe Kiela Jun 30, 2023 ▶ 31:32
What-if
Douwe Kiela: The open-source AI ecosystem exists solely due to Meta's LLaMA
“This whole flourishing that you see right now of open source models that basically comes from Meta's generosity in giving Lama away for free. And if they hadn't done that, then you wouldn't see that.”
Douwe Kiela Jun 30, 2023 ▶ 32:10
Opinion
Douwe Kiela: Getting struck by lightning is more likely than AI extinction
“Probably the chance of me getting hit by lightning, like right now is much higher than that happening.”
Douwe Kiela Jun 30, 2023 ▶ 34:32
Opinion
Douwe Kiela: Incumbent AI firms promote existential risk to entrench regulatory advantage
“So the people who are pushing this narrative are really the people who are benefiting from this being the narrative. So these are the incumbent AI companies who are, who want to have either the market regulated. Right. Because then they benefit because they ca…”
Douwe Kiela Jun 30, 2023 ▶ 35:01
Prediction Not checkable as stated
Douwe Kiela: EU AI regulation will completely destroy innovation
“What Europe is going to try to do is overregulate everything and just completely destroy innovation.”
Douwe Kiela Jun 30, 2023 ▶ 38:28
Prediction Not checkable as stated
Douwe Kiela: An enterprise AI adoption tidal wave is coming
“I think the tidal wave is coming. There are just big problems that we have to overcome, and these are the things I just talked about, right? So hallucination, attribution compliance up-to-dateness, data privacy, latency and I think the whole field is, is movin…”
Douwe Kiela Jun 30, 2023 ▶ 40:44
Opinion
Douwe Kiela: Large AI startup funding rounds are justified by potential payoffs
“I think some of the rounds were pretty big, but I think it's also justified just because this stuff is really going to change the world. And so one bet, if it's right, has massive payoff.”
Douwe Kiela Jun 30, 2023 ▶ 44:33
Prediction Not checkable as stated
Douwe Kiela: AI hype disillusionment will dry up funding and threaten startups
“At some point there's going to be a disillusionment with the technology and then funding might dry up and then these places are really in trouble.”
Douwe Kiela Jun 30, 2023 ▶ 45:14
Opinion
Douwe Kiela: Microsoft repositioned itself as an AI leader via OpenAI
“So I, I've been very impressed actually by how Microsoft has managed to turn everything around by strategically collaborating with a better AI lab in the shape of open AI. And they've just really turned that into this narrative where Microsoft is an AI leader …”
Douwe Kiela Jun 30, 2023 ▶ 45:35
Opinion
Douwe Kiela: Apple has failed to produce interesting AI innovations
“So far I haven't really seen a lot of interesting things coming out of Apple.”
Douwe Kiela Jun 30, 2023 ▶ 46:04
Assertion Not checkable as stated
Douwe Kiela: AutoGPT does not actually work despite the hype
“And there's an auto GPT thing that is going to change the world, but doesn't actually work.”
Douwe Kiela Jun 30, 2023 ▶ 48:10
Disclosure
Douwe Kiela: I was wrong to underestimate compute scaling laws and OpenAI
“The strongest belief I had that turned out to be wrong is that I really underestimated how important scale is in artificial intelligence. So I, and I think this is really one of the things that open AI has excelled at is that if you throw an order of magnitude…”
Douwe Kiela Jun 30, 2023 ▶ 50:26
Prediction Open · timeframe Jun 2033
Douwe Kiela: AI will displace the majority human workforce within 5-10 years
“If you look at how OpenAI and Anthropic and these places define AGI, it's as systems achieving capabilities that allow them to effectively do the work of humans for the majority of economically valuable human tasks. Then we're not that far away from it. And so…”
Douwe Kiela Jun 30, 2023 ▶ 52:02
Opinion
Douwe Kiela: OpenAI and Anthropic are the Lycos and AltaVista of AI
“So if all the stars align, then OpenAI and Anthropic and all of these places, they had this great first generation technology. So they're kind of like the Lycos and Alta Vista of search engines. And the technology we have is more like PageRank. And that would …”
Douwe Kiela Jun 30, 2023 ▶ 53:20

Shorts cut from this episode

▶ AI hallucinations: bug or feature? 🤖 · 20VC with Harry Steb (@11:48) ▶ ChatGPT's Secret Sauce (allegedly) 🤫 · 20VC with Harry Steb (@19:01) ▶ How To Build Your Own ChatGPT · 20VC with Harry Stebbings (@21:59)
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 1,200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.