Back to papers
Revisiting Co-Processing for Hash Joins on the Coupled CPU-GPU Architecture
Summary: Revisits hash joins on coupled CPU-GPU architectures, exploiting on-chip cache reuse and fine-grained co-processing with and without partitioning. Extends cost models to auto-guide design choices; on AMD APUs, achieves 53% CPU-only, 35% GPU-only, and 28% conventional co-processing gains.
(summarized by gpt-5-nano on Feb 09 2026)
- Paper ID
- 10751
- Venue
- VLDB
- Year
- 2013
- Pagerank
- 8.599693e-05
- Overall Rank
- 2,523 | 82.47%
- DOI
-
-
Incoming Non-self Citations Over Time
Incoming Citations (Sorted by Pagerank)
Showing 26 of 26 citing papers.
| Rank |
Citing Paper |
Year |
Venue |
Pagerank |
| 2,044 |
A Study of the Fundamental Performance Characteristics of GPUs and CPUs for Database Analytics |
2020 |
SIGMOD |
9.6963999e-05 |
| 2,072 |
HippogriffDB: Balancing I/O and GPU Bandwidth in Big Data Analytics |
2016 |
VLDB |
9.6300019e-05 |
| 3,307 |
Robust Query Processing in Co-Processor-accelerated Databases |
2016 |
SIGMOD |
7.2391191e-05 |
| 3,328 |
Pump Up the Volume: Processing Large Data on GPUs with Fast Interconnects |
2020 |
SIGMOD |
7.2136181e-05 |
| 3,471 |
GPL: A GPU-based Pipelined Query Processing Engine |
2016 |
SIGMOD |
7.0628019e-05 |
| 3,771 |
SABER: Window-Based Hybrid Stream Processing for Heterogeneous Architectures |
2016 |
SIGMOD |
6.7738535e-05 |
| 3,899 |
Efficient Join Algorithms For Large Database Tables in a Multi-GPU Environment |
2021 |
VLDB |
6.6513982e-05 |
| 3,994 |
Improving Main Memory Hash Joins on Intel Xeon Phi Processors: An Experimental Approach |
2015 |
VLDB |
6.5476357e-05 |
| 4,000 |
MG-Join: A Scalable Join for Massively Parallel Multi-GPU Architectures |
2021 |
SIGMOD |
6.5419402e-05 |
| 4,089 |
In-Cache Query Co-Processing on Coupled CPU-GPU Architectures |
2015 |
VLDB |
6.4559891e-05 |
| 4,359 |
Hardware-conscious Query Processing in GPU-accelerated Analytical Engines |
2019 |
CIDR |
6.2493951e-05 |
| 4,678 |
OmniDB: Towards Portable and Efficient Query Processing on Parallel CPU/GPU Architectures |
2013 |
VLDB |
5.9988623e-05 |
| 5,018 |
Orchestrating Data Placement and Query Execution in Heterogeneous CPU-GPU DBMS |
2022 |
VLDB |
5.7503878e-05 |
| 5,039 |
Tile-based Lightweight Integer Compression in GPU |
2022 |
SIGMOD |
5.7369993e-05 |
| 5,089 |
TCUDB: Accelerating Database with Tensor Processors |
2022 |
SIGMOD |
5.7017353e-05 |
| 5,251 |
Triton Join: Efficiently Scaling to a Large Join State on GPUs with Fast Interconnects |
2022 |
SIGMOD |
5.6003972e-05 |
| 5,826 |
Towards a Hybrid Design for Fast Query Processing in DB2 with BLU Acceleration Using Graphical Processing Units: A Technology Demonstration |
2016 |
SIGMOD |
5.3116062e-05 |
| 6,220 |
Distributed GPU Joins on Fast RDMA-capable Networks |
2023 |
SIGMOD |
5.1446966e-05 |
| 6,367 |
Improving Execution Efficiency of Just-in-time Compilation based Query Processing on GPUs |
2021 |
VLDB |
5.0887599e-05 |
| 6,963 |
A Morsel-Driven Query Execution Engine for Heterogeneous Multi-Cores |
2019 |
VLDB |
4.8769125e-05 |
| 7,038 |
Demonstrating Efficient Query Processing in Heterogeneous Environments |
2014 |
SIGMOD |
4.8500241e-05 |
| 7,570 |
Powerful GPUs or Fast Interconnects: Analyzing Relational Workloads on Modern GPUs |
2025 |
VLDB |
4.7039167e-05 |
| 7,752 |
Efficiently Processing Joins and Grouped Aggregations on GPUs |
2025 |
SIGMOD |
4.6558737e-05 |
| 8,846 |
Scaling your Hybrid CPU-GPU DBMS to Multiple GPUs |
2024 |
VLDB |
4.432948e-05 |
| 10,253 |
Scalable GPU Acceleration of Scalar Functions in Analytical Databases: Compilation, Benchmarking, and Optimization |
2026 |
VLDB |
4.1905499e-05 |
| 11,023 |
Accelerating Merkle Patricia Trie with GPU |
2024 |
VLDB |
4.1905499e-05 |
Outgoing Citations (Sorted by Pagerank)
Showing 18 of 18 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
Semantically Similar Papers