Trillions of rows, explorable in real time.
The analytics team needed to explore time-series data at a scale where the usual answer — pre-aggregate everything and accept the loss of resolution — destroyed the signal they were looking for.
We built a high-density visualization tool combining Rust-based downsampling with Databricks SQL and Apache Arrow zero-copy memory buffers. Queries adapt to the analyst’s zoom level: aggregated views by default, with SQL pushdown retrieving original high-resolution data on demand. Larger-than-memory datasets stay interactive instead of being flattened into something smaller and less useful.
“Petabyte Pitstops with Mercedes, Databricks SQL and Plotly Resampler” · Data + AI Summit 2024