DBScholar

Back to papers

Terabyte-Scale Analytics in the Blink of an Eye

Summary: Prototype GPU-cluster SQL engine applying ML/HPC practices (group communication, high-bandwidth memory, fast interconnects) to bound achievable speedups for distributed analytics. Shows a realistic >60× improvement, running all 22 TPC-H queries at 1TB almost instantaneously. (summarized by gpt-5-mini on Mar 13 2026)

Paper ID
h4648334be94b4f7f
Venue
VLDB
Year
2026
Pagerank
5.9219616e-05
Overall Rank
5,991 | 59.74%
DOI
10.14778/3773749.3773754
PDF
Download (CC BY-NC-ND 4.0)

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{wu_vldb26,
        title = {{Terabyte-Scale Analytics in the Blink of an Eye}},
        author = {Wu, Bowen and Cui, Wei and Curino, Carlo and Interlandi, Matteo and Sen, Rathijit},
        journal = {PVLDB},
        series = {{VLDB} '26},
        volume = {19},
        number = {2},
        pages = {141--155},
        doi = {10.14778/3773749.3773754},
        url = {https://doi.org/10.14778/3773749.3773754},
        year = {2026}
}

Incoming Citations (Sorted by Pagerank)

Showing 7 of 7 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 38 of 38 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
52 The Snowflake Elastic Data Warehouse 2016 SIGMOD 0.00041210636
681 Amazon Redshift Re-invented 2022 SIGMOD 0.0001482366
775 The Yin and Yang of Processing Data Warehousing Queries on GPU Devices 2013 VLDB 0.0001407924
1,028 High-Speed Query Processing over High-Speed Networks 2016 VLDB 0.00012416417
1,135 Dremel: A Decade of Interactive SQL Analysis at Web Scale 2020 VLDB 0.00011886308
1,206 NUMA-aware algorithms: the case of data shuffling 2013 CIDR 0.00011536099
1,269 A Study of the Fundamental Performance Characteristics of GPUs and CPUs for Database Analytics 2020 SIGMOD 0.00011254742
1,487 Photon: A Fast Query Engine for Lakehouse Systems 2022 SIGMOD 0.00010516813
1,748 HetExchange: Encapsulating heterogeneous CPU-GPU parallelism in JIT compiled engines 2019 VLDB 9.7289385e-05
1,790 POLARIS: The Distributed SQL Engine in Azure Synapse 2020 VLDB 9.6227952e-05
2,287 Pump Up the Volume: Processing Large Data on GPUs with Fast Interconnects 2020 SIGMOD 8.691301e-05
2,460 Query Processing on Tensor Computation Runtimes 2022 VLDB 8.4308406e-05
3,055 MG-Join: A Scalable Join for Massively Parallel Multi-GPU Architectures 2021 SIGMOD 7.6983202e-05
3,628 Triton Join: Efficiently Scaling to a Large Join State on GPUs with Fast Interconnects 2022 SIGMOD 7.1490678e-05
3,657 Tile-based Lightweight Integer Compression in GPU 2022 SIGMOD 7.1227107e-05
3,888 Orchestrating Data Placement and Query Execution in Heterogeneous CPU-GPU DBMS 2022 VLDB 6.942437e-05
3,955 Tensors: An abstraction for general data processing 2021 VLDB 6.8994712e-05
3,974 TCUDB: Accelerating Database with Tensor Processors 2022 SIGMOD 6.8848857e-05
3,986 Designing an Open Framework for Query Optimization and Compilation 2022 VLDB 6.8697828e-05
4,484 GPU Database Systems Characterization and Optimization 2024 VLDB 6.5776765e-05
4,992 GOLAP: A GPU-in-Data-Path Architecture for High-Speed OLAP 2024 SIGMOD 6.3217714e-05
5,376 Distributed GPU Joins on Fast RDMA-capable Networks 2023 SIGMOD 6.1555648e-05
5,543 Vortex: Overcoming Memory Capacity Limitations in GPU-Accelerated Large-Scale Data Analytics 2025 VLDB 6.0873423e-05
5,616 BOSS - An Architecture for Database Kernel Composition 2024 VLDB 6.0623704e-05
5,778 The Tensor Data Platform: Towards an AI-centric Database System 2023 CIDR 5.9953658e-05
5,891 Unified Query Optimization in the Fabric Data Warehouse 2024 SIGMOD 5.9542027e-05
6,153 Efficiently Processing Joins and Grouped Aggregations on GPUs 2025 SIGMOD 5.8669913e-05
6,417 Evaluating Multi-GPU Sorting with Modern Interconnects 2022 SIGMOD 5.7900099e-05
6,589 Maximus: A Modular Accelerated Query Engine for Data Analytics on Heterogeneous Systems 2025 SIGMOD 5.7401837e-05
6,671 Powerful GPUs or Fast Interconnects: Analyzing Relational Workloads on Modern GPUs 2025 VLDB 5.7125893e-05
7,455 Modularis: Modular Relational Analytics over Heterogeneous Distributed Platforms 2021 VLDB 5.5210654e-05
8,242 Scaling your Hybrid CPU-GPU DBMS to Multiple GPUs 2024 VLDB 5.3684987e-05
8,280 Scaling GPU-Accelerated Databases beyond GPU Memory Size 2025 VLDB 5.3616863e-05
9,401 Themis: A GPU-accelerated Relational Query Execution Engine 2025 VLDB 5.1837077e-05
9,438 Efficiently Joining Large Relations on Multi-GPU Systems 2025 VLDB 5.1761941e-05
9,870 GPU Acceleration of SQL Analytics on Compressed Data 2026 VLDB 5.115241e-05
10,025 Share the Tensor Tea: How Databases can Leverage the Machine Learning Ecosystem 2022 VLDB 5.091949e-05
10,139 GpJSON: High-performance JSON Data Processing on GPUs 2025 VLDB 5.0725068e-05
Previous Page 1 / 1 Next

Semantically Similar Papers