Apache Hive, every mention
5 scenes (2016) · ← back to Apache Hive
every year 2016 anyone Matt Turck 7Praveen Murugesan 5Dave Burgess 4Justin Borgman 3Todd Papaioannou 2Stefan Groschupf 2Mike Driscoll 2M.C. Srivas 2Ion Stoica 2Ben Rogojan (Seattle Data Guy) 2
Verbatim, from the transcripts: the passages where Apache Hive comes up
Making Big Data Accessible Using the Cloud // Ashish Thusoo, Qubole [FirstMark's Data Driven]
- ▶ 3:07 Ashish Thusoo So, um, you know, big data, essentially the emergence of these new systems, we hear about systems like Hadoop, Spark, Hive, and so on and so forth.
The Uber Big Data Story // Praveen Murugesan, Uber (Data Driven NYC / FirstMark)
- ▶ 4:41 Praveen Murugesan And, ah, on top of HDFS, we basically have, like, Spark, and, ah, Presto, and Hive.
- ▶ 7:18 Praveen Murugesan So, a few things, like I talked about, strict schema management, so we actually built, like, a central schema repository which is used for schema management, and then, ah, we unlocked, like, a whole bunch of new tools with, like, data on…
- ▶ 15:21 Praveen Murugesan and then it's provided as, like, a Hive UDF. 3 times in the scene
A Virtual Distributed Storage System // Haoyuan Li, Alluxio (Hosted by FirstMark)
- ▶ 12:41 Haoyuan Li They run a streaming workload on top of it, and they also use Aluxio to share the data efficiently between streaming processing and their batch processing, and use Hive in this particular use case.