Jun 12, 2019 · 23m · mad
What's Next for Open-Source Time Series Data? // Paul Dix, Influx Data (FirstMark's Data Driven NYC)
gold bands on the timeline = statements, start to end. Hover to read, click to jump. CC turns on captions
In this Data Driven NYC keynote, InfluxData CTO and Founder Paul Dix explores the unique database requirements of time series data and introduces InfluxDB 2.0 alongside Flux, a novel functional scripting language designed to unify data collection, querying, and stream processing.
How this conversation actually went
Every chapter scored 0–10 on four independent dynamics. Hover any point for the reasoning behind the score. Matt holds 1.4% of the talking time here. How this is scored →
speaking balance: gold is Matt, purple is the guest (3 minute bins)
Paul forcefully and humorously dismisses adopting existing languages like Lisp or JavaScript for Flux, arguing that if Paul Graham and Rich Hickey could not make Lisp popular, nobody will.
Hardest push from Matt ▶ 19:38 Host Caps Talk Duration for Q&AHost Matt Turck steps in to end the formal presentation due to time limits and redirects the remaining time to quick audience questions.
Biggest teaching moment ▶ 20:10 Explaining Dual Inverted Index ArchitecturePaul educates an audience member on how InfluxDB combines a columnar data store for time series values with an inverted index mapping metadata tag key-value pairs.
Matt holds his own ▶ 19:38 Host Contextualizes Company ProgressHost Matt Turck highlights the massive progress made by InfluxData since Paul last presented at the Data Driven NYC event.
the scores for every segment, with the reasoning behind each
| Chapter | Topic | Matt as informed peer | Guest teaching | Guest disagreement | Matt pushing back | Why |
|---|---|---|---|---|---|---|
| Defining Time Series Data Across Key Use Cases | 0 | 0 | 0 | 0 | Paul Dix provides an introductory presentation defining time series data and distinguishing between regular metrics and irregular event streams. Because this is a solo keynote segment, all host-side dynamic scores are zero. | |
| Unique Database Workload Characteristics of Time Series | 0 | 0 | 0 | 0 | Paul explains database workload characteristics unique to time series, such as high write throughput, large range scans, and retention eviction. Host interaction scores remain zero for this monologue segment. | |
| The Origin and Evolution of the TICK Stack | 0 | 0 | 0 | 0 | Paul traces the origin of InfluxDB and the emergence of Telegraf, Kapacitor, and Chronograf to form the TICK stack. This is a solo technical talk with no host participation. | |
| InfluxDB Line Protocol Schema and Nanosecond Precision | 0 | 0 | 0 | 0 | Paul details the Line Protocol schema, support for varied field types, and nanosecond timestamp precision. Host interaction scores are zero during this monologue presentation. | |
| The Unified Architecture of InfluxDB 2.0 | 0 | 0 | 0 | 0 | Paul describes consolidating the separate TICK stack components into a unified InfluxDB 2.0 platform. Host scores remain zero as this is part of the main presentation. | |
| Client Libraries and Modular Visualization Tools | 0 | 0 | 0 | 0 | Paul covers official client libraries, UI JavaScript components, and the Flux engine pluggable parser architecture. Host scores are zero due to the solo monologue format. | |
| Flux Code Demonstration and Serverless Execution Platform | 0 | 0 | 0 | 0 | Paul demonstrates Flux syntax, functional pipe operators, and its capability as a serverless execution engine inside the database. Host-side scores are zero for this monologue. | |
| InfluxDB 2.0 Distribution Options and Keynote Conclusion | 0 | 2 | 2 | 0 | Host Matt Turck opens the floor to audience questions regarding indexing and language design rationale. Paul playfully dismisses Lisp and alternative options while answering the audience, maintaining full conversational authority. |