Matei Zaharia
Co-Founder and CTO, Databricks · 1 appearance on the record.
computed by AI from the episodes · how this works → · full disclaimer →
founderexecutiveacademicscientistengineer@matei_zaharia ↗LinkedIn ↗profiles.stanford.edu/matei-zaharia ↗Wikipedia ↗
Matei Zaharia created Apache Spark during his Ph.D. at UC Berkeley and co-founded Databricks. He has also co-developed open-source technologies such as MLflow and Delta Lake, earning the ACM Prize in Computing.
3 supported 0 partly supported 1 contradicted 1 not yet assessed 1 not checkable as stated how the 6 claims stand · each chip opens the sources
6 assertions · 1 opinion · 4 insights · 2 disclosures · every statement was checked. The predictions and assertions are the 6 claims: statements the public record can support or contradict. 4 are resolved, 1 is not yet assessed, and 1 names no date, number or outcome precise enough to check. Everything else (opinions, insights, what ifs, disclosures) can never be settled by the record, so it carries no assessment.
The record, in short
What the tape says about how Matei argues and how the claims held up. Everything they said, and everything said about them, is in the tabs below.
Their most notable supported claim
Their most notable contradicted claim
Expressed certainty vs assessment result
weighted support: a fully supported claim counts one, a partly supported claim counts half. Each filled bar is clickable and opens exactly those claims; "none yet" means nothing said at that certainty level has resolved yet
Argument clarity: do they answer the question? how? →
answered every one of 10 assessed questions directly
This is a score against a rubric. It is not a rank. Every host question → answer exchange is scored with names hidden on directness, coherence, precision and compression, 1–5 each, on meaning alone: disfluencies are ignored, and only raw unedited episodes count. This is the score that measures thought. Every scored exchange, scores shown → · The rubric and its checks →
How they sound: speaking style how? →
242 words/min while actually speaking · 55.3 um and uh per 1k words · 22.2 false starts per 1k · 36.2% of pauses land inside a clause
Measured by listening to the audio itself: 2,982 words across 1 episode of raw-level tape, transcribed verbatim with every um and uh kept, each one attributed only where the alignment onto our timed stream is unambiguous. These are measurements of speaking style. We do not rank them: across this corpus, fluency and argument quality are nearly uncorrelated (ρ≈0.2), and smooth talking does not signal clear thinking. How it's measured →
Everything Matei Zaharia said on the a16z Podcast that made the record, most notable first. Filter by type, assessment or year in the ledger →
The other half of the tape: Matei Zaharia's own voice is left out of every number here. Other people bring the name up 6 times in 3 episodes on the a16z Podcast. every mention, with the transcript →
Who brings them up most Ben Horowitz 3Michael Franklin 1Ion Stoica 1Ali Ghodsi 1
Every mention by year
2025 5 mentions in 2 episodes 3 per episode
Appearances (1)
| Episode | Date | Speaking time |
|---|---|---|
| a16z Podcast | A Conversation With the Inventor of Spark | Jan 2, 2019 | 13m |