Efficient Top-K Query Processing on Massively Parallel Hardware
Summary: GPU-based top-k algorithms for massively parallel data analytics, including a novel bitonic top-k with up to 15x speedups over sort for k ≤ 256. A cost model predicts relative performance across algorithms and matches measurements on modern GPUs. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Anil Shanbhag (Massachusetts Institute of Technology)
- 2. Holger Pirk (Imperial College London)
- 3. Samuel Madden (Massachusetts Institute of Technology)
BibTeX Citation
@inproceedings{shanbhag_sigmod18,
title = {{Efficient Top-K Query Processing on Massively Parallel Hardware}},
author = {Shanbhag, Anil and Pirk, Holger and Madden, Samuel},
series = {{SIGMOD} '18},
booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
publisher = {Association for Computing Machinery},
doi = {10.1145/3183713.3183735},
url = {https://dl.acm.org/doi/10.1145/3183713.3183735},
year = {2018}
}
Incoming Citations (Sorted by Pagerank)
Showing 4 of 4 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 7,058 | BOSS - An Architecture for Database Kernel Composition | 2024 | VLDB | 5.7154673e-05 |
| 9,172 | External Merge Sort for Top-K Queries: Eager input filtering guided by histograms | 2020 | SIGMOD | 5.3092396e-05 |
| 11,348 | Cache-Efficient Top-k Aggregation over High Cardinality Large Datasets | 2024 | VLDB | 5.093636e-05 |
| 11,562 | MinMax Sampling: A Near-optimal Global Summary for Aggregation in the Wide Area | 2022 | SIGMOD | 5.093636e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 8 of 8 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 305 | GPUTeraSort: High Performance Graphics Co-processor Sorting for Large Database Management | 2006 | SIGMOD | 0.00021872796 |
| 712 | Efficient Implementation of Sorting on Multi-Core SIMD CPU Architecture | 2008 | VLDB | 0.0001468812 |
| 823 | The Yin and Yang of Processing Data Warehousing Queries on GPU Devices | 2013 | VLDB | 0.00013792901 |
| 912 | Hardware-Oblivious Parallelism for In-Memory Column-Stores | 2013 | VLDB | 0.00013269804 |
| 1,469 | Voodoo - A Vector Algebra for Portable Database Performance on Modern Hardware | 2016 | VLDB | 0.00010678751 |
| 2,232 | Database Compression on Graphics Processors | 2010 | VLDB | 8.8970926e-05 |
| 2,667 | A Memory Bandwidth-Efficient Hybrid Radix Sort on GPUs | 2017 | SIGMOD | 8.2756346e-05 |
| 2,713 | Robust Query Processing in Co-Processor-accelerated Databases | 2016 | SIGMOD | 8.2122726e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 5,233 | Progressive Top-k Subarray Query Processing in Array Databases | 2019 | VLDB |
| 2 | 305 | GPUTeraSort: High Performance Graphics Co-processor Sorting for Large Database Management | 2006 | SIGMOD |
| 3 | 7,959 | Terabyte-Scale Analytics in the Blink of an Eye | 2026 | VLDB |
| 4 | 678 | Fast Sort on CPUs and GPUs: A Case for Bandwidth Oblivious SIMD Sort | 2010 | SIGMOD |
| 5 | 4,225 | Realtime Top-k Personalized PageRank over Large Graphs on GPUs | 2020 | VLDB |
| 6 | 7,591 | Efficiently Processing Joins and Grouped Aggregations on GPUs | 2025 | SIGMOD |
| 7 | 5,422 | GPU Database Systems Characterization and Optimization | 2024 | VLDB |
| 8 | 9,172 | External Merge Sort for Top-K Queries: Eager input filtering guided by histograms | 2020 | SIGMOD |
| 9 | 8,281 | Efficient Top-K Processing Over Query-Dependent Functions | 2008 | VLDB |
| 10 | 2,667 | A Memory Bandwidth-Efficient Hybrid Radix Sort on GPUs | 2017 | SIGMOD |