Jerry Liu

Co-founder & CEO, LlamaIndex · 1 appearance on the record.

computed by AI from the episodes · how this works → · full disclaimer →

founderexecutiveengineerscientist@jerryjliu0 ↗LinkedIn ↗jerryjliu.github.io ↗

Jerry Liu created GPT Index in 2022 and co-founded LlamaIndex, a data framework and enterprise platform for retrieval-augmented generation applications. Before founding LlamaIndex, he conducted autonomous driving research at Uber ATG, engineered feed ranking systems at Quora, and led machine learning monitoring efforts at Robust Intelligence.

14statements → 4claims → 4claims resolved → 75%fully supported → 3.93/5average certainty → 1.29/5average debate potential → 2said about them ↓

3 supported 0 partly supported 1 contradicted how the 4 claims stand · each chip opens the sources

4 assertions · 3 insights · 7 disclosures · every statement was checked. The predictions and assertions are the 4 claims: statements the public record can support or contradict. 4 are resolved. Everything else (opinions, insights, what ifs, disclosures) can never be settled by the record, so it carries no assessment.

The record, in short

What the tape says about how Jerry argues and how the claims held up. Everything they said, and everything said about them, is in the tabs below.

Their most notable supported claim

Assertion Supported
Liu: LlamaHub features over 100 community-contributed data loaders
“And so these days, like Lama Hub is a very rich repository of, like, you know, the hundred plus, like, different data loaders from all different services and formats, and it's growing, like, every day.”
Jerry Liu Jul 12, 2023 ▶ 15:39 Building LlamaIndex: Jerry Liu on Scaling Retrieval-Augmented AI

Their most notable contradicted claim

Assertion Contradicted
Liu: Uber and Instabase use LlamaIndex for enterprise data applications
“We've seen people build these workflows at different settings from, for instance, like hacks on projects at startups building, for instance, like track GPT, like, plugin over, like, your Slack or your Notion all the way to kind of, like, bigger companies, for …”
Jerry Liu Jul 12, 2023 ▶ 26:58 Building LlamaIndex: Jerry Liu on Scaling Retrieval-Augmented AI

How they sound: speaking style how? →

277 words/min while actually speaking · 57.5 um and uh per 1k words

No argument clarity score for Jerry Liu: only 2 usable question→answer exchanges on raw tape (a fair score needs 8+). We do not score a sample that small. Roundtable and news formats yield far fewer direct exchanges than interviews.

Measured by listening to the audio itself: 6,818 words across 1 episode of raw-level tape, transcribed verbatim with every um and uh kept, each one attributed only where the alignment onto our timed stream is unambiguous. These are measurements of speaking style. We do not rank them: across this corpus, fluency and argument quality are nearly uncorrelated (ρ≈0.2), and smooth talking does not signal clear thinking. How it's measured →

Everything Jerry Liu said on the MAD Podcast that made the record, most notable first. Filter by type, assessment or year in the ledger →

Disclosure
LlamaIndex is not building a vector database, integrates with 12-20 existing ones
“We're not building our own vector database, but we have a rich set of integrations with, you know like 12 to like 20 different vector databases out there these days.”
Jerry Liu Jul 12, 2023 ▶ 9:20 Building LlamaIndex: Jerry Liu on Scaling Retrieval-Augmented AI
Disclosure
LlamaIndex focuses on data infrastructure while LangChain builds broader application frameworks
“Blind train is a great application framework for you to just like get us out of building blocks for a lot of different components, for instance, from like LL modules to prompts to some basic like retrieval and vector database abstractions to like also agent fr…”
Jerry Liu Jul 12, 2023 ▶ 12:23 Building LlamaIndex: Jerry Liu on Scaling Retrieval-Augmented AI
Insight
Liu: LLM data pipelines differ from traditional ETL stacks due to unstructured content comprehension
“If we're building this new age of L-empowered applications, The kind of requirements for the type of data that like you want to load as well as how you want to extract information from that data will be a little bit different than the existing ETL stack. The r…”
Jerry Liu Jul 12, 2023 ▶ 18:57 Building LlamaIndex: Jerry Liu on Scaling Retrieval-Augmented AI
Assertion Contradicted
Liu: Uber and Instabase use LlamaIndex for enterprise data applications
“We've seen people build these workflows at different settings from, for instance, like hacks on projects at startups building, for instance, like track GPT, like, plugin over, like, your Slack or your Notion all the way to kind of, like, bigger companies, for …”
Jerry Liu Jul 12, 2023 ▶ 26:58 Building LlamaIndex: Jerry Liu on Scaling Retrieval-Augmented AI
Insight
Liu: Generating long-form content over custom data remains a hard problem
“Generating something that's like a paragraph is pretty easy for ChatGPT to do. Generating like an entire blog post or essay, especially over your data is a pretty challenging problem.”
Jerry Liu Jul 12, 2023 ▶ 28:37 Building LlamaIndex: Jerry Liu on Scaling Retrieval-Augmented AI
Insight
Jerry Liu: Injecting metadata into text chunks improves LLM retrieval performance
“Second is being able to inject metadata actually is quite important to actually improve retrieval performance of like the downstream application, because like, you know, let's say you're splitting up like a sec, 10 K filing into a bunch of chunks within a sing…”
Jerry Liu Jul 12, 2023 ▶ 8:23 Building LlamaIndex: Jerry Liu on Scaling Retrieval-Augmented AI
Assertion Supported
Liu: LlamaHub features over 100 community-contributed data loaders
“And so these days, like Lama Hub is a very rich repository of, like, you know, the hundred plus, like, different data loaders from all different services and formats, and it's growing, like, every day.”
Jerry Liu Jul 12, 2023 ▶ 15:39 Building LlamaIndex: Jerry Liu on Scaling Retrieval-Augmented AI
Assertion Supported
Liu: LlamaIndex enables data querying in three or four lines of code
“Because in about three or four lines of code, you can load data and just it index it and then query it.”
Jerry Liu Jul 12, 2023 ▶ 21:17 Building LlamaIndex: Jerry Liu on Scaling Retrieval-Augmented AI
Disclosure
LlamaIndex adopted Keras's design philosophy of progressive disclosure of complexity
“A key philosophy that we launched and like zero dot six auto that we're continuing to build towards is this idea that was inspired by Karis, which is this idea of like progressive disclosure of complexity, where in the beginning things are very high level and …”
Jerry Liu Jul 12, 2023 ▶ 24:38 Building LlamaIndex: Jerry Liu on Scaling Retrieval-Augmented AI
Assertion Supported
Liu: OpenBB integrated LlamaIndex as its natural language layer
“OpenBB, which is the open source, like financial analysis tool, and they recently incorporated Lama Index as a natural language layer to help power, like their, basically, open source Bloomberg vibe, right?”
Jerry Liu Jul 12, 2023 ▶ 27:48 Building LlamaIndex: Jerry Liu on Scaling Retrieval-Augmented AI
Disclosure
Jerry Liu: LlamaIndex is building a production-grade enterprise version
“So we're Actually building out and scoping out initial version of what, like kind of production grade Lama index would look like.”
Jerry Liu Jul 12, 2023 ▶ 29:38 Building LlamaIndex: Jerry Liu on Scaling Retrieval-Augmented AI
Disclosure
Jerry Liu: LlamaIndex aims to connect language models with custom data
“So the high level mission is really to connect your language models with your data and basically unlock the capabilities of language models, whether it's like reasoning agent, like planning, or also like question answering inside extraction on top of your data…”
Jerry Liu Jul 12, 2023 ▶ 5:30 Building LlamaIndex: Jerry Liu on Scaling Retrieval-Augmented AI
Disclosure
Jerry Liu: 'Drop everything' reactiveness accounts for 10-15% of LlamaIndex's work
“There's going to be cases where there's going to be things that come out where like, you know, this really is going to be like a drop everything a moment. Like, you know, just stop what you're doing. Like this actually is like P zero. We have to like figure ou…”
Jerry Liu Jul 12, 2023 ▶ 32:33 Building LlamaIndex: Jerry Liu on Scaling Retrieval-Augmented AI
Disclosure
LlamaIndex paused planned work for three days to completely rewrite documentation
“So we basically just like stop what we were doing, spent the past three days just completely rewriting the documentation, and then we just launched that yesterday.”
Jerry Liu Jul 12, 2023 ▶ 33:43 Building LlamaIndex: Jerry Liu on Scaling Retrieval-Augmented AI

The other half of the tape: Jerry Liu's own voice is left out of every number here. Other people bring the name up 2 times in 1 episode on the MAD Podcast. every mention, with the transcript →

Who brings them up most Matt Turck 2

Every mention by year

tap a year for its mentions
0011212023episodesmentions
0112023episodes it came up in
0010.5212023episodesmentions per episode

Appearances (1)

EpisodeDateSpeaking time
Building LlamaIndex: Jerry Liu on Scaling Retrieval-Augmented AI Jul 12, 2023 30m
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.