“This is sort of the holy grail of database engineering is, why not build a single system that can do both of this? But it ends up just being a lot of compromises. And one, I think one of the first issue is that, hey, each, they say Postgres has a massive ecosystem, right? You want to be using the tools that's built for Postgres. And Spark, for example, had a massive ecosystem. There's a lot of libraries you want to use. If you were to create now a new thing, you don't have an ecosystem. You tend to create a new, smaller proprietary API, and you're lacking both. And it's also very difficult to make it performance-wise to be sort of comparable. On either side. So it ends up being actually sucking on both.”
quote is from the automated transcript, cleaned for reading:
filler sounds and stutters are removed, nothing is rephrased. names can be misheard
(the analysis reads context, assessments check outside sources). how →
More from Reynold Xin
Opinion
Xin: Vector databases should never have been a separate category
“Vector database should have never been a separate category.”
Reynold XinJun 24, 2026▶ 52:32The Agent Cloud: Databricks’ Bet on the Future of AI — Matei Zaharia and Reynold Xin
Insight
Xin: Unifying storage delivers 99% of HTAP database benefits
“HTAP wants to build a single engine for both. We think you can get 99% of what you need by unifying the storage and just have a single storage layer.”
Reynold XinJun 24, 2026▶ 33:16The Agent Cloud: Databricks’ Bet on the Future of AI — Matei Zaharia and Reynold Xin
AssertionSupported
Xin: Transcoding database rows to Parquet speeds object storage writes with zero compromise
“And as a matter of fact, once you transcode the data compresses better. So from those services writing to, for example, S three or other data lake, like object stores, you can actually write them faster because now they are now smaller. So there's no. Overhead…”
Reynold XinJun 24, 2026▶ 36:50The Agent Cloud: Databricks’ Bet on the Future of AI — Matei Zaharia and Reynold Xin
Insight
Xin: Overfitting to Initial Customers Has Far Lower Downside Than Boiling the Ocean
“I think the industry has a sense of, hey, maybe if you overfit to like one or two customers, it's going to be really bad for you. But I think the downside overfitting is much smaller than the upside itself. And if you sort of try to be too ambitious and boil t…”
Reynold XinJun 24, 2026▶ 41:55The Agent Cloud: Databricks’ Bet on the Future of AI — Matei Zaharia and Reynold Xin
PredictionNot checkable as stated
Xin: Much of traditional software will be rewritten with data and agents
“Actually, I think many of the traditional software will be sort of rewritten with this new paradigm, which is just get the data to be there. And then they slap some agent on top.”
Reynold XinJun 24, 2026▶ 1:07:44The Agent Cloud: Databricks’ Bet on the Future of AI — Matei Zaharia and Reynold Xin
AssertionSupported
Xin: Every major analytics database engine in traction is a decade old
“Actually, every single database engine out there, especially on the analytics side, are kind of a decade old. Pretty much everything that had reasonable traction are about a decade old.”
Reynold XinJun 24, 2026▶ 45:08The Agent Cloud: Databricks’ Bet on the Future of AI — Matei Zaharia and Reynold Xin
Made with StarZero
Turn any episode into a week of clips.
This entire site, over 200 episodes transcribed, diarized, checked and made playable,
runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the
moments worth sharing, cuts them, captions them, and reframes them for every feed.
We use essential cookies to make the site work. With your permission we
also use analytics cookies (Google Analytics and Mixpanel) to understand
usage and improve StarZero. See our Cookie Policy.