Dec 17, 2015 · 35m · mad

A Fireside Chat with MapR CTO M.C. Srivas (Data Driven NYC / FirstMark)

M.C. Srivas · 23m spoken Matt Turck · 3m spoken
0:00 / 0:00
▶ Watch on YouTube →

gold bands on the timeline = statements, start to end. Hover to read, click to jump. CC turns on captions

At a FirstMark Data Driven NYC event, MapR co-founder and CTO M.C. Srivas joins Matt Turck to discuss the founding, architectural innovations, and real-world enterprise applications of MapR's big data platform. He shares insights on big data scale, open-source strategy, and major implementations like India's 1.3-billion-person Aadhaar biometric project.

How this conversation actually went

Every chapter scored 0–10 on four independent dynamics. Hover any point for the reasoning behind the score. Matt holds 10.1% of the talking time here. How this is scored →

Matt as informed peer 2.4 Guest teaching 4.1 Guest disagreement 1.4 Matt pushing back 1.3
05100:0010:0020:0030:001:28–6:03 · Matt as informed peer 2/10 MapR Founding Story and Early Startup Challenges Matt sets up the founding story by noting MapR's funding and headcount, then asks standard questions about hiring and early sales. Srivas explains early struggles, like hiring in a sparse room without chairs and learning that potential customers lie during pre-product interviews.6:03–9:48 · Matt as informed peer 3/10 MapRFS vs HDFS and High-Scale Email Infrastructure Matt asks a direct technical question about how MapRFS differs from HDFS. Srivas educates the room by detailing HDFS's architectural limits with small files and scale, contrasting it with MapRFS handling massive email traffic and multi-temperature storage.9:48–16:05 · Matt as informed peer 1/10 MapRDB and India's Aadhaar Biometric Identification System Matt prompts Srivas to discuss MapRDB, leading to an extended story from Srivas about NoSQL history and India's Aadhaar biometric system. Srivas dominates the narrative detailing how MapRDB powers identity verification for over 900 million citizens.16:05–19:32 · Matt as informed peer 1/10 MapR Streams and Real-Time IoT Data Processing Matt introduces MapR Streams, and Srivas leads an engaging discussion on real-time streaming for IoT. Srivas quizzes the audience and explains the massive data footprint of self-driving cars and defense systems.19:32–24:57 · Matt as informed peer 6/10 Open Source Strategy and Competitive Landscape Matt pushes Srivas on the tension between open source and proprietary models and highlights competitors like Cloudera and Hortonworks. Srivas aggressively reframes the debate, accusing competitors of being disingenuous and exposing single-vendor control over open source projects.24:57–30:27 · Matt as informed peer 3/10 Mainstream Adoption and Future of Big Data Matt asks if Hadoop has crossed the chasm into mainstream enterprise adoption. Srivas explains that flexible schema requirements have made big data a standard budget item for major companies before transitioning to audience questions.30:27–34:27 · Matt as informed peer 1/10 Q&A: Google Technical Leadership and Native File Systems Audience members ask about Google's technical lead and MapR's native file system capabilities outside Hadoop. Srivas provides technical perspective on Google's historic advantage and confirms MapRFS usage via POSIX and NFS interfaces.1:28–6:03 · Guest teaching 3/10 MapR Founding Story and Early Startup Challenges Matt sets up the founding story by noting MapR's funding and headcount, then asks standard questions about hiring and early sales. Srivas explains early struggles, like hiring in a sparse room without chairs and learning that potential customers lie during pre-product interviews.6:03–9:48 · Guest teaching 5/10 MapRFS vs HDFS and High-Scale Email Infrastructure Matt asks a direct technical question about how MapRFS differs from HDFS. Srivas educates the room by detailing HDFS's architectural limits with small files and scale, contrasting it with MapRFS handling massive email traffic and multi-temperature storage.9:48–16:05 · Guest teaching 4/10 MapRDB and India's Aadhaar Biometric Identification System Matt prompts Srivas to discuss MapRDB, leading to an extended story from Srivas about NoSQL history and India's Aadhaar biometric system. Srivas dominates the narrative detailing how MapRDB powers identity verification for over 900 million citizens.16:05–19:32 · Guest teaching 4/10 MapR Streams and Real-Time IoT Data Processing Matt introduces MapR Streams, and Srivas leads an engaging discussion on real-time streaming for IoT. Srivas quizzes the audience and explains the massive data footprint of self-driving cars and defense systems.19:32–24:57 · Guest teaching 6/10 Open Source Strategy and Competitive Landscape Matt pushes Srivas on the tension between open source and proprietary models and highlights competitors like Cloudera and Hortonworks. Srivas aggressively reframes the debate, accusing competitors of being disingenuous and exposing single-vendor control over open source projects.24:57–30:27 · Guest teaching 3/10 Mainstream Adoption and Future of Big Data Matt asks if Hadoop has crossed the chasm into mainstream enterprise adoption. Srivas explains that flexible schema requirements have made big data a standard budget item for major companies before transitioning to audience questions.30:27–34:27 · Guest teaching 4/10 Q&A: Google Technical Leadership and Native File Systems Audience members ask about Google's technical lead and MapR's native file system capabilities outside Hadoop. Srivas provides technical perspective on Google's historic advantage and confirms MapRFS usage via POSIX and NFS interfaces.1:28–6:03 · Guest disagreement 1/10 MapR Founding Story and Early Startup Challenges Matt sets up the founding story by noting MapR's funding and headcount, then asks standard questions about hiring and early sales. Srivas explains early struggles, like hiring in a sparse room without chairs and learning that potential customers lie during pre-product interviews.6:03–9:48 · Guest disagreement 1/10 MapRFS vs HDFS and High-Scale Email Infrastructure Matt asks a direct technical question about how MapRFS differs from HDFS. Srivas educates the room by detailing HDFS's architectural limits with small files and scale, contrasting it with MapRFS handling massive email traffic and multi-temperature storage.9:48–16:05 · Guest disagreement 0/10 MapRDB and India's Aadhaar Biometric Identification System Matt prompts Srivas to discuss MapRDB, leading to an extended story from Srivas about NoSQL history and India's Aadhaar biometric system. Srivas dominates the narrative detailing how MapRDB powers identity verification for over 900 million citizens.16:05–19:32 · Guest disagreement 1/10 MapR Streams and Real-Time IoT Data Processing Matt introduces MapR Streams, and Srivas leads an engaging discussion on real-time streaming for IoT. Srivas quizzes the audience and explains the massive data footprint of self-driving cars and defense systems.19:32–24:57 · Guest disagreement 5/10 Open Source Strategy and Competitive Landscape Matt pushes Srivas on the tension between open source and proprietary models and highlights competitors like Cloudera and Hortonworks. Srivas aggressively reframes the debate, accusing competitors of being disingenuous and exposing single-vendor control over open source projects.24:57–30:27 · Guest disagreement 1/10 Mainstream Adoption and Future of Big Data Matt asks if Hadoop has crossed the chasm into mainstream enterprise adoption. Srivas explains that flexible schema requirements have made big data a standard budget item for major companies before transitioning to audience questions.30:27–34:27 · Guest disagreement 1/10 Q&A: Google Technical Leadership and Native File Systems Audience members ask about Google's technical lead and MapR's native file system capabilities outside Hadoop. Srivas provides technical perspective on Google's historic advantage and confirms MapRFS usage via POSIX and NFS interfaces.1:28–6:03 · Matt pushing back 1/10 MapR Founding Story and Early Startup Challenges Matt sets up the founding story by noting MapR's funding and headcount, then asks standard questions about hiring and early sales. Srivas explains early struggles, like hiring in a sparse room without chairs and learning that potential customers lie during pre-product interviews.6:03–9:48 · Matt pushing back 1/10 MapRFS vs HDFS and High-Scale Email Infrastructure Matt asks a direct technical question about how MapRFS differs from HDFS. Srivas educates the room by detailing HDFS's architectural limits with small files and scale, contrasting it with MapRFS handling massive email traffic and multi-temperature storage.9:48–16:05 · Matt pushing back 0/10 MapRDB and India's Aadhaar Biometric Identification System Matt prompts Srivas to discuss MapRDB, leading to an extended story from Srivas about NoSQL history and India's Aadhaar biometric system. Srivas dominates the narrative detailing how MapRDB powers identity verification for over 900 million citizens.16:05–19:32 · Matt pushing back 0/10 MapR Streams and Real-Time IoT Data Processing Matt introduces MapR Streams, and Srivas leads an engaging discussion on real-time streaming for IoT. Srivas quizzes the audience and explains the massive data footprint of self-driving cars and defense systems.19:32–24:57 · Matt pushing back 5/10 Open Source Strategy and Competitive Landscape Matt pushes Srivas on the tension between open source and proprietary models and highlights competitors like Cloudera and Hortonworks. Srivas aggressively reframes the debate, accusing competitors of being disingenuous and exposing single-vendor control over open source projects.24:57–30:27 · Matt pushing back 1/10 Mainstream Adoption and Future of Big Data Matt asks if Hadoop has crossed the chasm into mainstream enterprise adoption. Srivas explains that flexible schema requirements have made big data a standard budget item for major companies before transitioning to audience questions.30:27–34:27 · Matt pushing back 1/10 Q&A: Google Technical Leadership and Native File Systems Audience members ask about Google's technical lead and MapR's native file system capabilities outside Hadoop. Srivas provides technical perspective on Google's historic advantage and confirms MapRFS usage via POSIX and NFS interfaces.

speaking balance: gold is Matt, purple is the guest (3 minute bins)

0:00 · Matt 46.3% · guest 53.7%0:00 · Matt 46.3% · guest 53.7%3:00 · Matt 9.7% · guest 90.3%3:00 · Matt 9.7% · guest 90.3%6:00 · Matt 13.3% · guest 86.7%6:00 · Matt 13.3% · guest 86.7%9:00 · Matt 3% · guest 97%9:00 · Matt 3% · guest 97%12:00 · Matt 0% · guest 100%12:00 · Matt 0% · guest 100%15:00 · Matt 2.7% · guest 97.3%15:00 · Matt 2.7% · guest 97.3%18:00 · Matt 9% · guest 91%18:00 · Matt 9% · guest 91%21:00 · Matt 17.4% · guest 82.6%21:00 · Matt 17.4% · guest 82.6%24:00 · Matt 14.8% · guest 85.2%24:00 · Matt 14.8% · guest 85.2%27:00 · Matt 0% · guest 100%27:00 · Matt 0% · guest 100%30:00 · Matt 0% · guest 100%30:00 · Matt 0% · guest 100%33:00 · Matt 4.7% · guest 95.3%33:00 · Matt 4.7% · guest 95.3%
Sharpest disagreement ▶ 22:49 Exposing open source hypocrisies

Srivas forcefully rejects the open-source framing promoted by competitors like Cloudera and Hortonworks, calling their claims disingenuous and exposing how vendors privately control open-source projects.

Hardest push from Matt ▶ 21:29 Challenging MapR's competitive positioning

Matt cites Cloudera founder Mike Olson and Hortonworks' open-source strategy to directly challenge Srivas on how MapR differentiates in a ruthlessly competitive market.

Biggest teaching moment ▶ 8:29 Explaining HDFS architectural limits

Srivas clearly educates the room by detailing exact numerical limitations of HDFS clusters regarding file counts and latency, illustrating why native architecture was required for high-scale applications.

Matt holds his own ▶ 21:29 Framing the open-source competitive landscape

Matt demonstrates high domain expertise by invoking specific rival executives and distinct open-source philosophical stances to force Srivas into defending MapR's architecture.

the scores for every segment, with the reasoning behind each
ChapterTopicMatt as informed peerGuest teachingGuest disagreementMatt pushing backWhy
MapR Founding Story and Early Startup Challenges 2311 Matt sets up the founding story by noting MapR's funding and headcount, then asks standard questions about hiring and early sales. Srivas explains early struggles, like hiring in a sparse room without chairs and learning that potential customers lie during pre-product interviews.
MapRFS vs HDFS and High-Scale Email Infrastructure 3511 Matt asks a direct technical question about how MapRFS differs from HDFS. Srivas educates the room by detailing HDFS's architectural limits with small files and scale, contrasting it with MapRFS handling massive email traffic and multi-temperature storage.
MapRDB and India's Aadhaar Biometric Identification System 1400 Matt prompts Srivas to discuss MapRDB, leading to an extended story from Srivas about NoSQL history and India's Aadhaar biometric system. Srivas dominates the narrative detailing how MapRDB powers identity verification for over 900 million citizens.
MapR Streams and Real-Time IoT Data Processing 1410 Matt introduces MapR Streams, and Srivas leads an engaging discussion on real-time streaming for IoT. Srivas quizzes the audience and explains the massive data footprint of self-driving cars and defense systems.
Open Source Strategy and Competitive Landscape 6655 Matt pushes Srivas on the tension between open source and proprietary models and highlights competitors like Cloudera and Hortonworks. Srivas aggressively reframes the debate, accusing competitors of being disingenuous and exposing single-vendor control over open source projects.
Mainstream Adoption and Future of Big Data 3311 Matt asks if Hadoop has crossed the chasm into mainstream enterprise adoption. Srivas explains that flexible schema requirements have made big data a standard budget item for major companies before transitioning to audience questions.
Q&A: Google Technical Leadership and Native File Systems 1411 Audience members ask about Google's technical lead and MapR's native file system capabilities outside Hadoop. Srivas provides technical perspective on Google's historic advantage and confirms MapRFS usage via POSIX and NFS interfaces.

Statements from this episode (20)

Assertion Not checkable as stated
Turck: MapR raised a $110 million funding round from Google Capital
“Most recently it was a hundred and ten million round from Google capital.”
Matt Turck Dec 17, 2015 ▶ 0:38
Assertion Not checkable as stated
Srivas: MapR has approximately 400 employees
“It's more like 400.”
M.C. Srivas Dec 17, 2015 ▶ 0:50
Insight
Srivas: Early founders must intentionally avoid hiring people like themselves
“And you have to try to hire people who are not like you. I mean, basically, that's the first advice. Because you tend to be attracted to people who are like you, and in the beginning you really want a big variety of people, and not from the same DNA.”
M.C. Srivas Dec 17, 2015 ▶ 4:11
Disclosure
Srivas interviewed 50 Hadoop-using companies before founding MapR
“You know, well, before we started Mapper, I spoke to, like, about 40 or 50 people who were using Hadoop. 50 companies.”
M.C. Srivas Dec 17, 2015 ▶ 4:43
Insight
Srivas: First funding rounds build product; later rounds fund pure sales
“Because the way this venture money flows is that you get the first bunch of money for building the product, and not again. After that, you get the money for selling it.”
M.C. Srivas Dec 17, 2015 ▶ 5:37
Assertion Not checkable as stated
Srivas: MapR powers a global email provider with 1.5 billion accounts
“We are now, from a file system perspective, we are powering one of the largest email providers in the world. Ok, ah, when I say largest email providers, everybody of you probably is on it, ah, one and a half billion email accounts.”
M.C. Srivas Dec 17, 2015 ▶ 6:31
Assertion Supported
Srivas: A single Apache HDFS cluster handles roughly 100 million files
“HDFS, a single cluster, can do about a hundred million files.”
M.C. Srivas Dec 17, 2015 ▶ 8:41
Prediction Partly held up
Srivas: Aadhaar will take 2.5 years to achieve full enrollment in India
“At the rate of one and a half million a day, they're gonna go for another two and a half years before everybody's on.”
M.C. Srivas Dec 17, 2015 ▶ 13:03
Assertion Partly supported
Srivas: India's Aadhaar biometric database runs on MapR across four datacenters
“Now, this is running on MapRDB in four data centers worldwide, in across India.”
M.C. Srivas Dec 17, 2015 ▶ 13:09
Disclosure
Srivas: MapR provided its database software to India's Aadhaar project for free
“We gave it for free, by the way, right, because we thought it was very important.”
M.C. Srivas Dec 17, 2015 ▶ 13:30
Assertion Not checkable as stated
Srivas: A single self-driving car produces one terabyte of data hourly
“So it's a single self-driving car produces one terabyte per hour.”
M.C. Srivas Dec 17, 2015 ▶ 16:29
Prediction Not checkable as stated
Srivas: Next scale of big data requires thousands of edge processing clusters
“I think the real next scale in, in Stefan's graph is you need thousands of clusters that do edge processing everywhere.”
M.C. Srivas Dec 17, 2015 ▶ 18:35
Opinion
Srivas: Cloudera and Hortonworks are disingenuous about their open source control
“I think that's a bit disingenuous on Hortonworks and Cloudrass part to say that Because they would like to, so the, here's the dirty little truth about open source software, open source vendors, right? I love open source as long as it's my open source and not …”
M.C. Srivas Dec 17, 2015 ▶ 22:49
Opinion
Srivas: Open source does not guarantee high software quality
“It's not clear that open source equals good quality yet.”
M.C. Srivas Dec 17, 2015 ▶ 24:00
Disclosure
Srivas: MapR has reached 1,000 customers in four years
“We have about a thousand customers now, right? In four years.”
M.C. Srivas Dec 17, 2015 ▶ 25:28
Assertion Not checkable as stated
Srivas: Almost every enterprise now has a Hadoop budget item
“Every company now has a Hadoop budget item. Almost every company.”
M.C. Srivas Dec 17, 2015 ▶ 25:50
Insight
Srivas: The era of rigid database schemas is over
“The old schema, rigid schema is gone. I mean, if you're still waiting, thinking that it's gonna come back, it's not. It's over.”
M.C. Srivas Dec 17, 2015 ▶ 26:34
Assertion Supported
Srivas: India's Aadhaar biometric system has registered 930 million people
“There's nine hundred and thirty million people online right now.”
M.C. Srivas Dec 17, 2015 ▶ 27:33
Disclosure
Srivas: Multiple governments in Africa, Asia, and South America approached MapR
“So I have been approached by the governments, a lot of governments in Africa, a lot of governments in Asia to do this, or even in South America.”
M.C. Srivas Dec 17, 2015 ▶ 29:40
Insight
Srivas: Startups should avoid emulating Google or Facebook engineering infrastructure
“Technology wise, they are more advanced, but there's only one Google and one Facebook, right, and everybody else is not like that, so don't try to do that. I mean, as a startup, that's very extreme.”
M.C. Srivas Dec 17, 2015 ▶ 31:46
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.