DBScholar

Back to papers

Terabyte-Scale Analytics in the Blink of an Eye

Summary: Prototype GPU-cluster SQL engine applying ML/HPC practices (group communication, high-bandwidth memory, fast interconnects) to bound achievable speedups for distributed analytics. Shows a realistic >60× improvement, running all 22 TPC-H queries at 1TB almost instantaneously. (summarized by gpt-5-mini on Mar 13 2026)

Paper ID
14475
Venue
VLDB
Year
2026
Pagerank
5.5181056e-05
Overall Rank
7,959 | 45.40%
DOI
10.14778/3773749.3773754

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{wu_vldb26,
        title = {{Terabyte-Scale Analytics in the Blink of an Eye}},
        author = {Wu, Bowen and Cui, Wei and Curino, Carlo and Interlandi, Matteo and Sen, Rathijit},
        journal = {PVLDB},
        series = {{VLDB} '26},
        volume = {19},
        number = {2},
        pages = {141--155},
        doi = {10.14778/3773749.3773754},
        url = {https://doi.org/10.14778/3773749.3773754},
        year = {2026}
}

Incoming Citations (Sorted by Pagerank)

Showing 3 of 3 citing papers.

Rank Citing Paper Year Venue Pagerank
10,119 Data Movement-Aware GPU Sharing for Data-Intensive Systems 2026 CIDR 5.093636e-05
10,123 Cloudspecs: Cloud Hardware Evolution Through the Looking Glass 2026 CIDR 5.093636e-05
10,580 GPU Acceleration of SQL Analytics on Compressed Data 2026 VLDB 5.093636e-05
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 38 of 38 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
66 The Snowflake Elastic Data Warehouse 2016 SIGMOD 0.00038561587
818 Amazon Redshift Re-invented 2022 SIGMOD 0.00013822916
823 The Yin and Yang of Processing Data Warehousing Queries on GPU Devices 2013 VLDB 0.00013792901
1,076 High-Speed Query Processing over High-Speed Networks 2016 VLDB 0.00012270109
1,203 NUMA-aware algorithms: the case of data shuffling 2013 CIDR 0.00011671169
1,453 Dremel: A Decade of Interactive SQL Analysis at Web Scale 2020 VLDB 0.00010742227
1,466 A Study of the Fundamental Performance Characteristics of GPUs and CPUs for Database Analytics 2020 SIGMOD 0.0001068941
1,824 Photon: A Fast Query Engine for Lakehouse Systems 2022 SIGMOD 9.6734544e-05
1,925 POLARIS: The Distributed SQL Engine in Azure Synapse 2020 VLDB 9.4764398e-05
1,977 HetExchange: Encapsulating heterogeneous CPU-GPU parallelism in JIT compiled engines 2019 VLDB 9.3641101e-05
2,566 Pump Up the Volume: Processing Large Data on GPUs with Fast Interconnects 2020 SIGMOD 8.4116562e-05
2,823 Query Processing on Tensor Computation Runtimes 2022 VLDB 8.0893814e-05
3,227 MG-Join: A Scalable Join for Massively Parallel Multi-GPU Architectures 2021 SIGMOD 7.6217889e-05
3,874 Tensors: An abstraction for general data processing 2021 VLDB 7.0561161e-05
4,089 TCUDB: Accelerating Database with Tensor Processors 2022 SIGMOD 6.9096857e-05
4,144 Tile-based Lightweight Integer Compression in GPU 2022 SIGMOD 6.8744592e-05
4,215 Designing an Open Framework for Query Optimization and Compilation 2022 VLDB 6.8275676e-05
4,243 Orchestrating Data Placement and Query Execution in Heterogeneous CPU-GPU DBMS 2022 VLDB 6.8073248e-05
4,535 Triton Join: Efficiently Scaling to a Large Join State on GPUs with Fast Interconnects 2022 SIGMOD 6.6419266e-05
5,422 GPU Database Systems Characterization and Optimization 2024 VLDB 6.2245373e-05
5,587 Distributed GPU Joins on Fast RDMA-capable Networks 2023 SIGMOD 6.1596139e-05
5,778 GOLAP: A GPU-in-Data-Path Architecture for High-Speed OLAP 2024 SIGMOD 6.0918646e-05
5,953 The Tensor Data Platform: Towards an AI-centric Database System 2023 CIDR 6.0309873e-05
6,065 Vortex: Overcoming Memory Capacity Limitations in GPU-Accelerated Large-Scale Data Analytics 2025 VLDB 5.9895288e-05
6,774 Evaluating Multi-GPU Sorting with Modern Interconnects 2022 SIGMOD 5.7778738e-05
7,055 Unified Query Optimization in the Fabric Data Warehouse 2024 SIGMOD 5.7170734e-05
7,058 BOSS - An Architecture for Database Kernel Composition 2024 VLDB 5.7154673e-05
7,432 Powerful GPUs or Fast Interconnects: Analyzing Relational Workloads on Modern GPUs 2025 VLDB 5.6199783e-05
7,591 Efficiently Processing Joins and Grouped Aggregations on GPUs 2025 SIGMOD 5.5900315e-05
7,899 Modularis: Modular Relational Analytics over Heterogeneous Distributed Platforms 2021 VLDB 5.5195553e-05
8,053 Maximus: A Modular Accelerated Query Engine for Data Analytics on Heterogeneous Systems 2025 SIGMOD 5.4990605e-05
8,808 Scaling your Hybrid CPU-GPU DBMS to Multiple GPUs 2024 VLDB 5.3661351e-05
9,280 Themis: A GPU-accelerated Relational Query Execution Engine 2025 VLDB 5.2933689e-05
9,333 Efficiently Joining Large Relations on Multi-GPU Systems 2025 VLDB 5.2887551e-05
9,835 Share the Tensor Tea: How Databases can Leverage the Machine Learning Ecosystem 2022 VLDB 5.2112879e-05
9,989 GpJSON: High-performance JSON Data Processing on GPUs 2025 VLDB 5.1826377e-05
10,580 GPU Acceleration of SQL Analytics on Compressed Data 2026 VLDB 5.093636e-05
10,985 Scaling GPU-Accelerated Databases beyond GPU Memory Size 2025 VLDB 5.093636e-05
Previous Page 1 / 1 Next

Semantically Similar Papers