Apache Spark, every mention
8 scenes (2017), the whole family · ← back to Apache Spark
every year 2017 anyone Matt Turck 59Haoyuan Li 13Ali Ghodsi 11Stefan Groschupf 10Praveen Murugesan 10Julien Le Dem 9Tobi Knaup 6Matt Housley 6Christopher Nguyen 6Prat Moghe 5
Verbatim, from the transcripts: the passages where Apache Spark comes up
Where Should Machines Go to Learn? // Auren Hoffman, SafeGraph (FirstMark's Data Driven)
- ▶ 3:41 Auren Hoffman You've got some sort of, like, Spark infrastructure streaming the data in with Kafka.
- ▶ 11:43 Auren Hoffman So, these are companies like Palantir, they're the BI tools, even things like, you know, Hadoop or Spark, um, you know, basically any, most of these companies are basically, let me take your own data and help you make better decisions with…
Three Loops of Analytics Efficiency // Sean Kandel, Trifacta (FirstMark's Data Driven)
- ▶ 6:41 Sean Kandel Uh, so probably still today the most common is using kind of hand coding tools, um, so programming languages, Python, Spark,
Big Data as a Service // Prat Moghe, Cazena (FirstMark's Data Driven)
- ▶ 4:52 Prat Moghe It's Hadoop, it's Spark, it's Python, it's R, and it's like every three months there's a new open source project around it.
- ▶ 10:17 Prat Moghe You're running Spark.
- ▶ 13:22 Prat Moghe Or, I'm running a data engineering job on Spark, and I got certain SLA, and I have certain price points, and go run it for me.
- ▶ 13:52 Prat Moghe It could, they could be SQL, they could be Spark, um, and we picked certain, uh, tools here. 2 times in the scene
Project Jupyter // Jason Grout & Sylvain Corlay, Bloomberg (FirstMark's Data Driven)
- ▶ 18:18 unnamed speaker So as Spark starts to kind of work its tentacles into every corner of the big data ecosystem, I'm seeing a lot of interest in projects like the Apache Zeppelin project as compared to, to Jupiter. 5 times in the scene