Uber transitioned from ETL into Vertica to EL into Hadoop
Praveen Murugesan · The Uber Big Data Story // Praveen Murugesan, Uber (Data Driven NYC / FirstMark) · Sep 30, 2016 · at 6:19
Praveen Murugesan, Ride Experience Lead at Uber, explains the architectural transition of Uber's data pipeline from traditional ETL to a Hadoop data lake.
“We went from an ETL model, where we scraped from, like, the original source, transformed the data and loaded to Vertica, to, like, just an EL model, where we just, like, just copy the data as soon as possible into, like, Hadoop, and all the transformation can basically just happen as secondary steps on Hadoop.”
quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →