Uber engineers frequently crashed Kafka clusters with unthrottled Spark executor writes
Praveen Murugesan · The Uber Big Data Story // Praveen Murugesan, Uber (Data Driven NYC / FirstMark) · Sep 30, 2016 · at 10:21
Praveen Murugesan, Ride Experience Lead at Uber, discusses architectural challenges in Uber's big data infrastructure during a presentation at Data Driven NYC.
“Kafka was, in general, like, a nice way where people used to pipe the results of, like, their Spark jobs. But often cases, what they do is, like, they hit Kafka hard and bring Kafka down because they're trying to, like, actually send data from, like, hundred executors or so on in Spark.”
quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →