Mark Grover

Co-founder & CEO, Stemma · 1 appearance on the record.

computed by AI from the episodes · how this works → · full disclaimer →

Mark Grover is the Co-founder and CEO of Stemma. He participated in a virtual fireside chat at Data Driven NYC in November 2021.

12statements → 2claims → 1claims resolved → 4.08/5average certainty → 2.08/5average debate potential → ≈4.5/5argument clarity, estimated →

1 supported 0 partly supported 0 contradicted 1 not checkable as stated how the 2 claims stand · each chip opens the sources

2 assertions · 2 opinions · 6 insights · 2 disclosures · every statement was checked. The predictions and assertions are the 2 claims: statements the public record can support or contradict. 1 is resolved, and 1 names no date, number or outcome precise enough to check. Everything else (opinions, insights, what ifs, disclosures) can never be settled by the record, so it carries no assessment.

The record, in short

What the tape says about how Mark argues and how the claims held up. Everything they said, and everything said about them, is in the tabs below.

Their most notable supported claim

Assertion Supported
Over 35 enterprise companies use the open-source data catalog Amundsen
“Over 35 companies use it in open source, you know, ING, Instacart, Brexasana.”
Mark Grover Nov 16, 2021 ▶ 17:50 Fireside Chat: Mark Grover (Co-Founder & CEO, Stemma) with Matt Turck (Partner, FirstMark)

How they sound: speaking style how? →

248 words/min while actually speaking · 23.2 um and uh per 1k words

No argument clarity score for Mark Grover: only 6 usable question→answer exchanges on raw tape (a fair score needs 8+). We do not score a sample that small. Roundtable and news formats yield far fewer direct exchanges than interviews. Their coarse estimate from 6 raw tape exchanges is ≈4.5/5, shown at half point precision because the sample is small.

Measured by listening to the audio itself: 4,474 words across 1 episode of raw-level tape, transcribed verbatim with every um and uh kept, each one attributed only where the alignment onto our timed stream is unambiguous. These are measurements of speaking style. We do not rank them: across this corpus, fluency and argument quality are nearly uncorrelated (ρ≈0.2), and smooth talking does not signal clear thinking. How it's measured →

Everything Mark Grover said on the MAD Podcast that made the record, most notable first. Filter by type, assessment or year in the ledger →

Insight
Curated data catalogs take up to three years and quickly become outdated
“The problem with this approach is that a, it takes a very long time to value. It takes You depending on the size of your organization, anywhere from like a year to three years to actually get all this metadata in to then hand it to your users or your complianc…”
Mark Grover Nov 16, 2021 ▶ 5:55 Fireside Chat: Mark Grover (Co-Founder & CEO, Stemma) with Matt Turck (Partner, FirstMark)
Insight
Manual data curation fails for fast-growing, data-intensive companies
“When you have one of one or more of these two criteria met, that system breaks. You cannot rely on curation as the source of discovery, understanding, context about a catalog, and therefore there's a need for what I now call automated data catalogs.”
Mark Grover Nov 16, 2021 ▶ 6:48 Fireside Chat: Mark Grover (Co-Founder & CEO, Stemma) with Matt Turck (Partner, FirstMark)
Opinion
Selling support and services for open-source products is a bad business model
“I find that selling support and services on an open source product is not a great business model.”
Mark Grover Nov 16, 2021 ▶ 24:24 Fireside Chat: Mark Grover (Co-Founder & CEO, Stemma) with Matt Turck (Partner, FirstMark)
Insight
Automated data catalogs narrow scope to make manual curation possible
“An automated data catalog cannot guarantee you that this is a single source of truth, but it can tell you that out of these 200 data sets related to, I don't know, pricing, These 180 are no good for you and your use case because they haven't been updated all t…”
Mark Grover Nov 16, 2021 ▶ 7:06 Fireside Chat: Mark Grover (Co-Founder & CEO, Stemma) with Matt Turck (Partner, FirstMark)
Insight
Analysts spend most of their time answering Slack questions, not modeling
“The expectation is like, oh, they spent all their time and like modeling and that modeling can be analytical or algorithmic. The reality is I feel like they spent all their time on Slack and they're like, Hey, what's the source of truth for this data?”
Mark Grover Nov 16, 2021 ▶ 9:36 Fireside Chat: Mark Grover (Co-Founder & CEO, Stemma) with Matt Turck (Partner, FirstMark)
Insight
Data catalog discovery should be public while underlying access remains gated
“One default stance that we have is that within a certain deployment and sometimes the deployment is for the entire company, certain, sometimes the deployment is for a certain line of business within the company, within a certain deployment, discovery of the da…”
Mark Grover Nov 16, 2021 ▶ 12:11 Fireside Chat: Mark Grover (Co-Founder & CEO, Stemma) with Matt Turck (Partner, FirstMark)
Opinion
Unbundled data tools create management problems that require automated data catalogs
“I, in my opinion, I feel strongly that's the right thing to do and you get the best of breed products, but it creates problems around management and governance that are new and need to be solved in a new way. And that's where like something like a data catalog…”
Mark Grover Nov 16, 2021 ▶ 14:20 Fireside Chat: Mark Grover (Co-Founder & CEO, Stemma) with Matt Turck (Partner, FirstMark)
Disclosure
Stemma struggles to serve command-and-control organizations that broadly restrict data access
“STEMM very clearly is in the latter category, right? We do not do well in serving the command and control style organizations.”
Mark Grover Nov 16, 2021 ▶ 21:14 Fireside Chat: Mark Grover (Co-Founder & CEO, Stemma) with Matt Turck (Partner, FirstMark)
Disclosure
Stemma omits messaging and BI features to force Slack and Looker integration
“But when you grow, you have to integrate with the ecosystem the organization is in, and we, for example, we have no feature to have a conversation in the data catalog, right? And that's intentional because we want to integrate with Slack, and that's why there'…”
Mark Grover Nov 16, 2021 ▶ 22:45 Fireside Chat: Mark Grover (Co-Founder & CEO, Stemma) with Matt Turck (Partner, FirstMark)
Insight
Central data teams should provide enablement tools, not act as bottlenecks
“It is the central team's responsibility to provide that tool. And it's the owner's responsibility to use that tool and then communicate directly to their consumers. So hard for me to say is governance should be managed at domain level. I'm not sure if I'm answ…”
Mark Grover Nov 16, 2021 ▶ 28:17 Fireside Chat: Mark Grover (Co-Founder & CEO, Stemma) with Matt Turck (Partner, FirstMark)
Assertion Not checkable as stated
Amundsen was Lyft's highest-rated internal product for three years
“From that day, which is probably mid 2018 till today, this product is the single highest CSAT scoring internal product, right?”
Mark Grover Nov 16, 2021 ▶ 16:41 Fireside Chat: Mark Grover (Co-Founder & CEO, Stemma) with Matt Turck (Partner, FirstMark)
Assertion Supported
Over 35 enterprise companies use the open-source data catalog Amundsen
“Over 35 companies use it in open source, you know, ING, Instacart, Brexasana.”
Mark Grover Nov 16, 2021 ▶ 17:50 Fireside Chat: Mark Grover (Co-Founder & CEO, Stemma) with Matt Turck (Partner, FirstMark)

Appearances (1)

EpisodeDateSpeaking time
Fireside Chat: Mark Grover (Co-Founder & CEO, Stemma) with Matt Turck (Partner, FirstMark) Nov 16, 2021 22m
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.