Jeff Huber

Co-founder & CEO, Chroma · 1 appearance on the record.

computed by AI from the episodes · how this works → · full disclaimer →

founderexecutive

Jeff Huber is the co-founder and CEO of Chroma, an open-source vector database company. He works on AI-native infrastructure and search systems designed to scale with modern context requirements.

16statements → 6claims → 1claims resolved → 3.81/5average certainty → 2/5average debate potential → ≈4.5/5argument clarity, estimated →

1 supported 0 partly supported 0 contradicted 5 not checkable as stated how the 6 claims stand · each chip opens the sources

5 predictions · 1 assertion · 3 opinions · 3 insights · 4 disclosures · every statement was checked. The predictions and assertion are the 6 claims: statements the public record can support or contradict. 1 is resolved, and 5 name no date, number or outcome precise enough to check. Everything else (opinions, insights, what ifs, disclosures) can never be settled by the record, so it carries no assessment.

The record, in short

What the tape says about how Jeff argues and how the claims held up. Everything they said, and everything said about them, is in the tabs below.

Their most notable supported claim

Assertion Supported
Huber: Steering LLMs at embedding layer is not exposed in closed-source models
“So, there's lots of stuff around steering language models at the embedding layer itself, and not using text but this is not yet exposed To at least closed source models, so.”
Jeff Huber Jun 28, 2023 ▶ 19:55 Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Infra

How they sound: speaking style how? →

260 words/min while actually speaking · 42 um and uh per 1k words

No argument clarity score for Jeff Huber: only 7 usable question→answer exchanges on raw tape (a fair score needs 8+). We do not score a sample that small. Roundtable and news formats yield far fewer direct exchanges than interviews. Their coarse estimate from 7 raw tape exchanges is ≈4.5/5, shown at half point precision because the sample is small.

Measured by listening to the audio itself: 4,359 words across 1 episode of raw-level tape, transcribed verbatim with every um and uh kept, each one attributed only where the alignment onto our timed stream is unambiguous. These are measurements of speaking style. We do not rank them: across this corpus, fluency and argument quality are nearly uncorrelated (ρ≈0.2), and smooth talking does not signal clear thinking. How it's measured →

Everything Jeff Huber said on the MAD Podcast that made the record, most notable first. Filter by type, assessment or year in the ledger →

Opinion
Huber: Current SOTA LLMs Lack Reliability for Multi-Agent Workflows
“Now, of course, for those of you that have actually played with technology, I think it's questionable whether the current state of the art Language models, embedding models, et cetera, will give you the reliability you want from, ah, you know, agents working t…”
Jeff Huber Jun 28, 2023 ▶ 18:16 Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Infra
Disclosure
Huber: Chroma is committed to remaining fully open source
“Chroma will always, we are committed to building the ubiquitous open source standard.”
Jeff Huber Jun 28, 2023 ▶ 13:30 Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Infra
Prediction Not checkable as stated
Huber: Vector databases must support both transactional and analytical workloads
“And we think that both certainly transactional has to be the case because it is a online database. It's gonna sit in the loop of applications. Again, you've already seen demos of this happening tonight. But also to make this technology useful for developers, y…”
Jeff Huber Jun 28, 2023 ▶ 16:06 Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Infra
Opinion
Huber: Asking whether vector databases replace classic databases is dumb
“People, there's this, like, you know, big question about, oh, are vector databases gonna replace classic databases? Are these competitive in some way? And I think it's just kind of a dumb question.”
Jeff Huber Jun 28, 2023 ▶ 29:23 Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Infra
Insight
Embedding search and analytics enable developers to improve model reliability
“By looking at embeddings doing embedding search, doing analytics over embedding space you could give developers you know, at the minimum of a divining rod, if not a compass, to be able to improve their models and get to the level of reliability they want to ha…”
Jeff Huber Jun 28, 2023 ▶ 1:47 Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Infra
Insight
Huber: Programmable memory enables reliable LLMs across all use cases
“Chroma's belief is that programmable memory, so developers being able to set terministically Hey, language model, this is the knowledge you should know about, this is the knowledge you should use, these are the tools you should know about, these are the tools …”
Jeff Huber Jun 28, 2023 ▶ 3:19 Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Infra
Prediction Not checkable as stated
Huber: 'Chat Your Data' AI Use Case Will See Mass Adoption
“So I think that use case, even just the Chat Your Data use case, truly will go to the ends of the earth.”
Jeff Huber Jun 28, 2023 ▶ 17:32 Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Infra
Assertion Supported
Huber: Steering LLMs at embedding layer is not exposed in closed-source models
“So, there's lots of stuff around steering language models at the embedding layer itself, and not using text but this is not yet exposed To at least closed source models, so.”
Jeff Huber Jun 28, 2023 ▶ 19:55 Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Infra
Prediction Not checkable as stated
Huber: AI-native databases will be much thicker than traditional databases
“We think the database will be, like, much thicker than it's been before.”
Jeff Huber Jun 28, 2023 ▶ 22:31 Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Infra
Prediction Not checkable as stated
Huber: Multimodal models will run directly inside application code and databases
“There'll be language models running inside the application code, obviously, language models running inside the database as well or large models more broadly, multimodal models will run, you know, everywhere as well.”
Jeff Huber Jun 28, 2023 ▶ 23:02 Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Infra
Prediction Not checkable as stated
Huber: Most enterprises will deploy language models within three years
“I certainly think probably most enterprises, organizations, companies on earth will have brought language models Into the company, probably in pretty meaningful ways. At minimum, the customer service department, the sales department, ops, back end, legal and h…”
Jeff Huber Jun 28, 2023 ▶ 24:26 Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Infra
Opinion
Huber: 'Vector database' is too narrow a term for information retrieval
“I don't actually like the term vector database that much. I think it's sort of narrow. I think the job to be done is information retrieval more broadly, and vector search happens to be a useful tool in our toolbox to doing information retrieval.”
Jeff Huber Jun 28, 2023 ▶ 28:31 Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Infra
Insight
Huber: Vector databases primarily handle un-databased unstructured data
“Vector databases are primarily about unstructured data. It's actually taking data that had no database that knew about it and loading it in for the very first time. Most applications of this stuff is not about taking data that's already in your relational data…”
Jeff Huber Jun 28, 2023 ▶ 29:36 Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Infra
Disclosure
Chroma is developing an open-source distributed version of its database
“We're working on a distributed version of Chroma. So in the same way that Elastic for those of you that are familiar with TextSearch, Elastic picked up Lucene, made it developer-friendly, they made it distributed, Chroma picks up some of these ANN algorithms, …”
Jeff Huber Jun 28, 2023 ▶ 9:32 Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Infra
Disclosure
Huber: Chroma is building density detection for vector space retrievals
“One of the things that we've been working on is this idea of query relevancy or density. So given retrievals, From vector space, given the search. We can say whether it came from a sparse or dense part of the embedding space.”
Jeff Huber Jun 28, 2023 ▶ 25:50 Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Infra
Disclosure
Huber: Chroma will not natively manage full human-in-the-loop workflows
“A database specifically, Chroma specifically, is not gonna do all of this workflow, obviously, and we're not gonna have, like, you know, user code for typing in answers.”
Jeff Huber Jun 28, 2023 ▶ 26:38 Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Infra

Appearances (1)

EpisodeDateSpeaking time
Why Vector Databases Are Exploding: Chroma Co-Founder Jeff Huber on Building AI-Native Inf Jun 28, 2023 20m
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.