Insight certainty 4/5 debate potential 3/5

Xin: Single HTAP database engines compromise ecosystem compatibility and performance

Reynold Xin · The Agent Cloud: Databricks’ Bet on the Future of AI — Matei Zaharia and Reynold Xin · Jun 24, 2026 · at 32:30

Databricks co-founder Reynold Xin discusses why historical attempts at Hybrid Transactional/Analytical Processing (HTAP) fail to replace dedicated OLTP and OLAP systems.

0:00 / 0:38exact quote · 38.8s
▶ Watch the full episode on YouTube → 720p mp4 · rendered on demand · StarZero watermark
“This is sort of the holy grail of database engineering is, why not build a single system that can do both of this? But it ends up just being a lot of compromises. And one, I think one of the first issue is that, hey, each, they say Postgres has a massive ecosystem, right? You want to be using the tools that's built for Postgres. And Spark, for example, had a massive ecosystem. There's a lot of libraries you want to use. If you were to create now a new thing, you don't have an ecosystem. You tend to create a new, smaller proprietary API, and you're lacking both. And it's also very difficult to make it performance-wise to be sort of comparable. On either side. So it ends up being actually sucking on both.”

quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →

More from Reynold Xin

Opinion
Xin: Vector databases should never have been a separate category
“Vector database should have never been a separate category.”
Reynold Xin Jun 24, 2026 ▶ 52:32 The Agent Cloud: Databricks’ Bet on the Future of AI — Matei Zaharia and Reynold Xin
Insight
Xin: Unifying storage delivers 99% of HTAP database benefits
“HTAP wants to build a single engine for both. We think you can get 99% of what you need by unifying the storage and just have a single storage layer.”
Reynold Xin Jun 24, 2026 ▶ 33:16 The Agent Cloud: Databricks’ Bet on the Future of AI — Matei Zaharia and Reynold Xin
Assertion Supported
Xin: Transcoding database rows to Parquet speeds object storage writes with zero compromise
“And as a matter of fact, once you transcode the data compresses better. So from those services writing to, for example, S three or other data lake, like object stores, you can actually write them faster because now they are now smaller. So there's no. Overhead…”
Reynold Xin Jun 24, 2026 ▶ 36:50 The Agent Cloud: Databricks’ Bet on the Future of AI — Matei Zaharia and Reynold Xin
Insight
Xin: Overfitting to Initial Customers Has Far Lower Downside Than Boiling the Ocean
“I think the industry has a sense of, hey, maybe if you overfit to like one or two customers, it's going to be really bad for you. But I think the downside overfitting is much smaller than the upside itself. And if you sort of try to be too ambitious and boil t…”
Reynold Xin Jun 24, 2026 ▶ 41:55 The Agent Cloud: Databricks’ Bet on the Future of AI — Matei Zaharia and Reynold Xin
Prediction Not checkable as stated
Xin: Much of traditional software will be rewritten with data and agents
“Actually, I think many of the traditional software will be sort of rewritten with this new paradigm, which is just get the data to be there. And then they slap some agent on top.”
Reynold Xin Jun 24, 2026 ▶ 1:07:44 The Agent Cloud: Databricks’ Bet on the Future of AI — Matei Zaharia and Reynold Xin
Assertion Supported
Xin: Every major analytics database engine in traction is a decade old
“Actually, every single database engine out there, especially on the analytics side, are kind of a decade old. Pretty much everything that had reasonable traction are about a decade old.”
Reynold Xin Jun 24, 2026 ▶ 45:08 The Agent Cloud: Databricks’ Bet on the Future of AI — Matei Zaharia and Reynold Xin
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.