DBScholar

Back to papers

Pump Up the Volume: Processing Large Data on GPUs with Fast Interconnects

Summary: NVLink 2.0-based interconnect eliminates CPU-GPU transfer bottlenecks, enabling large-scale in-memory processing on GPUs. Demonstrates scalable no-partitioning hash join beyond GPU memory with up to 18x speedup vs PCIe 3.0 and 7.3x vs optimized CPU. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
hc899165b1dbc1f24
Venue
SIGMOD
Year
2020
Pagerank
8.691301e-05
Overall Rank
2,287 | 84.64%
DOI
10.1145/3318464.3389705

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{lutz_sigmod20,
        title = {{Pump Up the Volume: Processing Large Data on GPUs with Fast Interconnects}},
        author = {Lutz, Clemens and Breß, Sebastian and Zeuch, Steffen and Rabl, Tilmann and Markl, Volker},
        series = {{SIGMOD} '20},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/3318464.3389705},
        url = {https://dl.acm.org/doi/10.1145/3318464.3389705},
        year = {2020}
}

Incoming Citations (Sorted by Pagerank)

Showing 33 of 33 citing papers.

Rank Citing Paper Year Venue Pagerank
2,460 Query Processing on Tensor Computation Runtimes 2022 VLDB 8.4308406e-05
3,628 Triton Join: Efficiently Scaling to a Large Join State on GPUs with Fast Interconnects 2022 SIGMOD 7.1490678e-05
3,657 Tile-based Lightweight Integer Compression in GPU 2022 SIGMOD 7.1227107e-05
3,862 GaccO - A GPU-accelerated OLTP DBMS 2022 SIGMOD 6.9623352e-05
3,888 Orchestrating Data Placement and Query Execution in Heterogeneous CPU-GPU DBMS 2022 VLDB 6.942437e-05
3,974 TCUDB: Accelerating Database with Tensor Processors 2022 SIGMOD 6.8848857e-05
4,484 GPU Database Systems Characterization and Optimization 2024 VLDB 6.5776765e-05
4,647 Improving Execution Efficiency of Just-in-time Compilation based Query Processing on GPUs 2021 VLDB 6.4850635e-05
4,946 CXL and the Return of Scale-Up Database Engines 2024 VLDB 6.3412158e-05
5,376 Distributed GPU Joins on Fast RDMA-capable Networks 2023 SIGMOD 6.1555648e-05
5,543 Vortex: Overcoming Memory Capacity Limitations in GPU-Accelerated Large-Scale Data Analytics 2025 VLDB 6.0873423e-05
5,616 BOSS - An Architecture for Database Kernel Composition 2024 VLDB 6.0623704e-05
5,991 Terabyte-Scale Analytics in the Blink of an Eye 2026 VLDB 5.9219616e-05
6,153 Efficiently Processing Joins and Grouped Aggregations on GPUs 2025 SIGMOD 5.8669913e-05
6,417 Evaluating Multi-GPU Sorting with Modern Interconnects 2022 SIGMOD 5.7900099e-05
6,671 Powerful GPUs or Fast Interconnects: Analyzing Relational Workloads on Modern GPUs 2025 VLDB 5.7125893e-05
6,883 DAPHNE: An Open and Extensible System Infrastructure for Integrated Data Analysis Pipelines 2022 CIDR 5.6552024e-05
7,325 An Examination of CXL Memory Use Cases for In-Memory Database Management Systems using SAP HANA 2024 VLDB 5.5508106e-05
7,528 Workload Placement on Heterogeneous CPU-GPU Systems 2024 VLDB 5.499936e-05
7,808 Analyzing Vectorized Hash Tables Across CPU Architectures 2023 VLDB 5.448023e-05
7,927 NOCAP: Near-Optimal Correlation-Aware Partitioning Joins 2023 SIGMOD 5.4234567e-05
7,968 Parallelizing Intra-Window Join on Multicores: An Experimental Study 2021 SIGMOD 5.4139853e-05
8,242 Scaling your Hybrid CPU-GPU DBMS to Multiple GPUs 2024 VLDB 5.3684987e-05
8,760 Zero-sided RDMA: Network-driven Data Shuffling for Disaggregated Heterogeneous Cloud DBMSs 2024 SIGMOD 5.2822161e-05
9,438 Efficiently Joining Large Relations on Multi-GPU Systems 2025 VLDB 5.1761941e-05
9,570 H-Rocks: CPU-GPU accelerated Heterogeneous RocksDB on Persistent Memory 2025 SIGMOD 5.154741e-05
9,870 GPU Acceleration of SQL Analytics on Compressed Data 2026 VLDB 5.115241e-05
10,484 L3: A GPU-Native Co-Designed Data Format for Learned Lossless Lightweight Compression 2026 SIGMOD 4.9769913e-05
10,612 TQEx: Tensor-based Query Engine Enhanced by Bridging the Gap 2026 SIGMOD 4.9769913e-05
10,795 MGI: A Communication Framework for Data Processing in Massive GPU Infrastructures 2026 VLDB 4.9769913e-05
10,811 PystachIO: Efficient Distributed GPU Query Processing with PyTorch over Fast Networks & Fast Storage 2026 VLDB 4.9769913e-05
10,932 TQP++: Bridging ML Compilers and Analytical Query Processing on GPUs 2026 VLDB 4.9769913e-05
11,573 Accelerating Merkle Patricia Trie with GPU 2024 VLDB 4.9769913e-05
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 31 of 31 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
210 Sort vs. Hash Revisited: Fast Join Implementation on Modern Multi-Core CPUs 2009 VLDB 0.00024844328
215 Morsel-Driven Parallelism: A NUMA-Aware Query Evaluation Framework for the Many-Core Age 2014 SIGMOD 0.00024589307
277 FAST: Fast Architecture Sensitive Tree Search on Modern CPUs and GPUs 2010 SIGMOD 0.00022320139
362 Design and Evaluation of Main Memory Hash Join Algorithms for Multi-core CPUs 2011 SIGMOD 0.00019999596
616 Relational Joins on Graphics Processors 2008 SIGMOD 0.00015554627
775 The Yin and Yang of Processing Data Warehousing Queries on GPU Devices 2013 VLDB 0.0001407924
857 Hardware-Oblivious Parallelism for In-Memory Column-Stores 2013 VLDB 0.00013422539
890 Rack-Scale In-Memory Join Processing using RDMA 2015 SIGMOD 0.00013241413
1,206 NUMA-aware algorithms: the case of data shuffling 2013 CIDR 0.00011536099
1,267 An Experimental Comparison of Thirteen Relational Equi-Joins in Main Memory 2016 SIGMOD 0.00011265987
1,454 Fast Computation of Database Operations using Graphics Processors 2004 SIGMOD 0.00010596726
1,534 Pipelined Query Processing in Coprocessor Environments 2018 SIGMOD 0.00010327147
1,542 HippogriffDB: Balancing I/O and GPU Bandwidth in Big Data Analytics 2016 VLDB 0.0001030858
1,614 Compressed Linear Algebra for Large-Scale Machine Learning 2016 VLDB 0.00010067153
1,748 HetExchange: Encapsulating heterogeneous CPU-GPU parallelism in JIT compiled engines 2019 VLDB 9.7289385e-05
2,124 Revisiting Co-Processing for Hash Joins on the Coupled CPU-GPU Architecture 2013 VLDB 9.0041425e-05
2,187 Database Compression on Graphics Processors 2010 VLDB 8.891124e-05
2,487 A Memory Bandwidth-Efficient Hybrid Radix Sort on GPUs 2017 SIGMOD 8.3939994e-05
2,601 Robust Query Processing in Co-Processor-accelerated Databases 2016 SIGMOD 8.2325291e-05
3,108 SABER: Window-Based Hybrid Stream Processing for Heterogeneous Architectures 2016 SIGMOD 7.6392673e-05
3,374 A Hybrid B+-tree as Solution for In-Memory Indexing on CPU-GPU Heterogeneous Computing Platforms 2016 SIGMOD 7.3606793e-05
3,491 In-Cache Query Co-Processing on Coupled CPU-GPU Architectures 2015 VLDB 7.2586387e-05
3,739 In-RDBMS Hardware Acceleration of Advanced Analytics 2018 VLDB 7.0595382e-05
3,766 CROSSBOW: Scaling Deep Learning with Small Batch Sizes on Multi-GPU Servers 2019 VLDB 7.0360981e-05
3,775 Hardware-conscious Query Processing in GPU-accelerated Analytical Engines 2019 CIDR 7.0261504e-05
3,934 The Case For Heterogeneous HTAP 2017 CIDR 6.9145781e-05
4,200 Adaptive Work Placement for Query Processing on Heterogeneous Computing Resources 2017 VLDB 6.7378729e-05
6,249 ColumnML: Column-Store Machine Learning with On-The-Fly Data Transformation 2019 VLDB 5.8343283e-05
6,518 GPU-accelerated data management under the test of time 2020 CIDR 5.7561332e-05
6,662 A Morsel-Driven Query Execution Engine for Heterogeneous Multi-Cores 2019 VLDB 5.7155423e-05
7,285 Lowering the Latency of Data Processing Pipelines Through FPGA based Hardware Acceleration 2020 VLDB 5.5639701e-05
Previous Page 1 / 1 Next

Semantically Similar Papers