Skip to content
brickster.ai
All videos
newsDatabricks·July 19, 2022

Opening the Floodgates: Enabling Fast, Unmediated End User Access to Trillion-Row Datasets with SQL

Description

Spreadsheets revolutionized IT by giving end users the ability to create their own analytics. Providing direct end user access to trillion-row datasets generated in financial markets or digital marketing is much harder. New SQL data warehouses like ClickHouse and Druid can provide fixed latency with constant cost on very large datasets, which opens up new possibilities. Our talk walks through recent experience on analytic apps developed by ClickHouse users that enable end users like market traders to develop their own analytics directly off raw data. We’ll cover the following topics. 1. Characteristics of new open source column databases and how they enable low-latency analytics at constant cost. 2. Idiomatic ways to validate new apps by building MVPs that support a wide range of queries on source data including storing source JSON, schema design, applying compression on columns, and building indexes for needle-in-a-haystack queries. 3. Incrementally identifying hotspots and applying easy optimizations to bring query performance into line with long term latency and cost requirements. 4. Methods of building accessible interfaces, including traditional dashboards, imitating e

Description from YouTube. Full content on the video page.

More from Databricks