Apache Spark, every mention

20 scenes (2015), the whole family · ← back to Apache Spark

tap a year for its mentions
003086015201420152016201720182019202020212022202320242025episodesmentions
0815201420152016201720182019202020212022202320242025episodes it came up in
0047.5815201420152016201720182019202020212022202320242025episodesmentions per episode

every year 2015 anyone Matt Turck 59Haoyuan Li 13Ali Ghodsi 11Stefan Groschupf 10Praveen Murugesan 10Julien Le Dem 9Tobi Knaup 6Matt Housley 6Christopher Nguyen 6Prat Moghe 5

Verbatim, from the transcripts: the passages where Apache Spark comes up

loading…

B2B Big Data Challenges, Nick Mehta, Gainsight (Data Driven NYC / FirstMark Capital) Dec 17, 2015 · 1 mention

  • ▶ 21:00 Nick Mehta Like, by the way, we're going to do, um, Spark, and we'll probably, maybe after the sales pitch, I'll do MapR and Datamir and everything else I do, right?

A Fireside Chat with MapR CTO M.C. Srivas (Data Driven NYC / FirstMark) Dec 17, 2015 · 2 mentions

  • ▶ 20:58 M.C. Srivas We introduced JSON to Hadoop and Spark and in MapReduce and in Hive and everywhere. 2 times in the scene

The Acceleration of Innovation in Big Data w/ Stefan Groschupf, Datameer Dec 17, 2015 · 10 mentions

  • ▶ 10:12 Stefan Groschupf Um, who thinks Spark is disruptive? 6 times in the scene
  • ▶ 11:51 Stefan Groschupf Oh yeah, we need spark, we need real time, we need this, this, this, but the, the, the thought, the think process, the thought process, I really believe needs to have happen upside down, bottom up, because it's, think about how difficult… 3 times in the scene
  • ▶ 15:36 Stefan Groschupf And finally, Yarn is very batch-centric, where again, now we kind of make it work with Spark and, and, and other things now, but, uh, Mesosphere really has a very flexible scheduling mechanism.

Black Boxes and Unicorns - DataRobot CEO Jeremy Achin Nov 23, 2015 · 1 mention

  • ▶ 18:30 Jeremy Achin Um, so we'll, we'll compete algorithms from R, from Python, from H-to-O, Spark, um, and just kind of compete them all against each other and see what works best for yours, right?

Liz Crawford, Birchbox // Data Science & Analytics at Birchbox (Hosted by FirstMark Capital) Oct 21, 2015 · 2 mentions

  • ▶ 4:49 Liz Crawford So we don't expect our data scientists to be able to produce our new spark infrastructure, for example.
  • ▶ 13:02 Liz Crawford We didn't start out with all of them on a, on Spark, for example.

Ramana Rao, Livefyre // Real-Time Social Engagement (Hosted by FirstMark Capital) Oct 21, 2015 · 2 mentions

Joseph Essas, OpenTable // Mining Diner Talk (Hosted by FirstMark Capital) Jun 19, 2015 · 2 mentions

  • ▶ 2:32 Joseph Essas Uh, all of our events flowing through Kafka, they've been populated into Cassandra, which then we run Spark instances that kind of model on top of the data. 2 times in the scene

Ion Stoica, Databricks // Creating Apache Spark // Data Driven NYC (FirstMark Capital) Apr 2, 2015 · 37 mentions

  • ▶ 0:19 Matt Turck So, uh, so as a quick intro, you are one of the co-creators of Apache Spark, uh, which is a unified framework for building, uh, data pipelines, and also one of the most active, um, open source projects in, in big data. 2 times in the scene
  • ▶ 5:45 Ion Stoica Um, the next project was Spark, and Spark was, ah, you know, we targeted first some workloads which are not covered by Hadoop, and from all this experience I mentioned earlier, we look at iterative, iterative computations to support…
  • ▶ 6:15 Matt Turck Uh, so, uh, you know, to the sort of uninitiated, uh, the, you know, Hadoop was the big thing, uh, for the last two to three years, and then Spark sort of appeared on the scene. 4 times in the scene
  • ▶ 7:24 Matt Turck Yes, so perhaps more specifically, can you go through the, the, the key advantages of Spark over MapReduce? 6 times in the scene
  • ▶ 9:37 Ion Stoica So there's a key here, you know, one way to look at Spark is that think about 3 times in the scene
  • ▶ 12:07 Matt Turck And to take just one of the, one of the things you mentioned, um, so maybe contrast Spark streaming with Storm, which also does micro-batching. 4 times in the scene
  • ▶ 15:30 Ion Stoica And it has, ah, a few things to make it easy, ah, in addition to Spark.
  • ▶ 16:31 Matt Turck What is the, uh, relationship between Databricks and the Spark community specifically? 6 times in the scene
  • ▶ 19:06 unnamed speaker I just wanted to get your thought on that and whether you guys were considering that for Spark. 2 times in the scene
  • ▶ 20:44 Matt Turck You've officially been appointed Spark Questionnaire. 8 times in the scene
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.