DBScholar

Back to papers

Scaling GPU-Accelerated Databases beyond GPU Memory Size

Summary: Hybrid CPU–GPU query processing scales beyond GPU memory: CPU-side filtering minimizes PCIe transfers, while GPUs execute joins and other compute-intensive operators. Handles 1TB TPC-H (SF1000) on one 80GB A100, outperforming CPU-only systems in speed and cost. (summarized by gpt-5.6-luna on Jul 24 2026)

Paper ID
14251
Venue
VLDB
Year
2025
Pagerank
5.093636e-05
Overall Rank
10,985 | 24.64%
DOI
10.14778/3749646.3749710

Incoming Non-self Citations Over Time

No non-self incoming citations found for this paper in this database.

Authors

BibTeX Citation

@article{li_vldb25,
        title = {{Scaling GPU-Accelerated Databases beyond GPU Memory Size}},
        author = {Li, Yinan and Ding, Bailu and Wei, Ziyun and Maas, Lukas M. and Al-Ghosien, Momin and Blanas, Spyros and Bruno, Nicolas and Curino, Carlo and Interlandi, Matteo and Peeper, Craig and Rajan, Kaushik and Chaudhuri, Surajit and Gehrke, Johannes},
        journal = {PVLDB},
        series = {{VLDB} '25},
        volume = {18},
        number = {11},
        pages = {4518--4531},
        doi = {10.14778/3749646.3749710},
        url = {https://doi.org/10.14778/3749646.3749710},
        year = {2025}
}

Incoming Citations (Sorted by Pagerank)

Showing 4 of 4 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 40 of 40 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
12 C-Store: A Column-oriented DBMS 2005 VLDB 0.00069513174
60 Integrating Compression and Execution in Column-Oriented Database Systems 2006 SIGMOD 0.0003955489
216 SIMD-Scan: Ultra Fast in-Memory Table Scan using on-Chip Vector Processing Units 2009 VLDB 0.00024498128
242 A Performance Evaluation of Four Parallel Join Algorithms in a Shared-Nothing Multiprocessor Environment 1989 SIGMOD 0.00023604323
293 Implementing Database Operations Using SIMD Instructions 2002 SIGMOD 0.00022259273
344 On the Power of Magic 1987 PODS 0.00020659405
631 Relational Joins on Graphics Processors 2008 SIGMOD 0.00015591241
634 Rethinking SIMD Vectorization for In-Memory Databases 2015 SIGMOD 0.00015533814
823 The Yin and Yang of Processing Data Warehousing Queries on GPU Devices 2013 VLDB 0.00013792901
870 BitWeaving: Fast Scans for Main Memory Data Processing 2013 SIGMOD 0.0001350293
912 Hardware-Oblivious Parallelism for In-Memory Column-Stores 2013 VLDB 0.00013269804
1,450 Row-wise Parallel Predicate Evaluation 2008 VLDB 0.00010746714
1,466 A Study of the Fundamental Performance Characteristics of GPUs and CPUs for Database Analytics 2020 SIGMOD 0.0001068941
1,469 Voodoo - A Vector Algebra for Portable Database Performance on Modern Hardware 2016 VLDB 0.00010678751
1,574 Pipelined Query Processing in Coprocessor Environments 2018 SIGMOD 0.00010321274
1,761 ByteSlice: Pushing the Envelop of Main Memory Data Processing with a New Storage Layout 2015 SIGMOD 9.8154969e-05
1,977 HetExchange: Encapsulating heterogeneous CPU-GPU parallelism in JIT compiled engines 2019 VLDB 9.3641101e-05
2,232 Database Compression on Graphics Processors 2010 VLDB 8.8970926e-05
2,667 A Memory Bandwidth-Efficient Hybrid Radix Sort on GPUs 2017 SIGMOD 8.2756346e-05
2,823 Query Processing on Tensor Computation Runtimes 2022 VLDB 8.0893814e-05
3,134 Efficient Join Algorithms For Large Database Tables in a Multi-GPU Environment 2021 VLDB 7.7231028e-05
3,227 MG-Join: A Scalable Join for Massively Parallel Multi-GPU Architectures 2021 SIGMOD 7.6217889e-05
3,471 In-Cache Query Co-Processing on Coupled CPU-GPU Architectures 2015 VLDB 7.3885861e-05
3,571 Looking Ahead Makes Query Plans Robust: Making the Initial Case with In-Memory Star Schema Data Warehouse Workloads 2017 VLDB 7.2991953e-05
3,592 The FastLanes Compression Layout: Decoding >100 Billion Integers per Second with Scalar Code 2023 VLDB 7.2774889e-05
4,144 Tile-based Lightweight Integer Compression in GPU 2022 SIGMOD 6.8744592e-05
4,243 Orchestrating Data Placement and Query Execution in Heterogeneous CPU-GPU DBMS 2022 VLDB 6.8073248e-05
4,553 Predicate Transfer: Efficient Pre-Filtering on Multi-Join Queries 2024 CIDR 6.6346951e-05
4,849 Bitvector-aware Query Optimization for Decision Support Queries 2020 SIGMOD 6.4803828e-05
5,014 Applying Hash Filters to Improving the Execution of Bushy Trees 1993 VLDB 6.4002123e-05
5,308 Towards a Hybrid Design for Fast Query Processing in DB2 with BLU Acceleration Using Graphical Processing Units: A Technology Demonstration 2016 SIGMOD 6.2720812e-05
5,422 GPU Database Systems Characterization and Optimization 2024 VLDB 6.2245373e-05
5,587 Distributed GPU Joins on Fast RDMA-capable Networks 2023 SIGMOD 6.1596139e-05
5,807 Improving Execution Efficiency of Just-in-time Compilation based Query Processing on GPUs 2021 VLDB 6.0805143e-05
6,481 MotherDuck: DuckDB in the cloud and in the client 2024 CIDR 5.8669399e-05
7,139 Selection Pushdown in Column Stores using Bit Manipulation Instructions 2023 SIGMOD 5.6932481e-05
7,660 Pruning in Snowflake: Working Smarter, Not Harder 2025 SIGMOD 5.5736132e-05
8,439 New Query Optimization Techniques in the Spark Engine of Azure Synapse 2022 VLDB 5.4248071e-05
8,808 Scaling your Hybrid CPU-GPU DBMS to Multiple GPUs 2024 VLDB 5.3661351e-05
9,835 Share the Tensor Tea: How Databases can Leverage the Machine Learning Ecosystem 2022 VLDB 5.2112879e-05
Previous Page 1 / 1 Next

Semantically Similar Papers