DBScholar

Back to papers

NUMA-aware algorithms: the case of data shuffling

Summary: Demonstrates that NUMA effects critically impact data shuffling on multi-socket multicore servers, with naive shuffling up to 3× slower than NUMA-aware variants. Achieves top performance using thread binding, NUMA-aware thread allocation, and relaxed global coordination, arguing such algorithmic redesign is essential as socket counts and memory heterogeneity grow. (summarized by gpt-5-mini on Feb 09 2026)

Paper ID
186
Venue
CIDR
Year
2013
Pagerank
0.00011671169
Overall Rank
1,203 | 91.75%
DOI
-

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{li_cidr13,
        address = {Amsterdam, Netherlands},
        series = {{CIDR} '13},
        title = {{NUMA-aware algorithms: the case of data shuffling}},
        booktitle = {Proceedings of the {Conference} on {Innovative} {Data} {Systems} {Research}},
        author = {Li, Yinan and Pandis, Ippokratis and Mueller, Rene and Raman, Vijayshankar and Lohman, Guy},
        year = {2013}
}

Incoming Citations (Sorted by Pagerank)

Showing 27 of 27 citing papers.

Rank Citing Paper Year Venue Pagerank
165 DB2 with BLU Acceleration: So Much More than Just a Column Store 2013 VLDB 0.00027693424
241 Morsel-Driven Parallelism: A NUMA-Aware Query Evaluation Framework for the Many-Core Age 2014 SIGMOD 0.00023654664
252 Multi-Core, Main-Memory Joins: Sort vs. Hash Revisited 2014 VLDB 0.00023242719
959 Memory-Efficient Hash Joins 2015 VLDB 0.00012953588
1,076 High-Speed Query Processing over High-Speed Networks 2016 VLDB 0.00012270109
1,150 DimmWitted: A Study of Main-Memory Statistical Analytics 2014 VLDB 0.00011943462
1,177 A Comprehensive Study of Main-Memory Partitioning and its Application to Large-Scale Comparison- and Radix-Sort 2014 SIGMOD 0.00011808761
1,784 Lambada: Interactive Data Analytics on Cold Data Using Serverless Cloud Infrastructure 2020 SIGMOD 9.7726335e-05
2,140 Revisiting Co-Processing for Hash Joins on the Coupled CPU-GPU Architecture 2013 VLDB 9.0991487e-05
2,566 Pump Up the Volume: Processing Large Data on GPUs with Fast Interconnects 2020 SIGMOD 8.4116562e-05
3,964 Hyper Dimension Shuffle: Efficient Data Repartition at Petabyte Scale in SCOPE 2019 VLDB 6.9855158e-05
4,001 Scaling Up Concurrent Main-Memory Column-Store Scans: Towards Adaptive NUMA-aware Data and Task Placement 2015 VLDB 6.9663191e-05
4,085 Deployment of Query Plans on Multicores 2015 VLDB 6.9149518e-05
4,607 Adaptive NUMA-aware data placement and task scheduling for analytical workloads in main-memory column-stores 2017 VLDB 6.6088324e-05
4,748 Taming Subgraph Isomorphism for RDF Query Processing 2015 VLDB 6.5251089e-05
5,206 BriskStream: Scaling Data Stream Processing on Shared-Memory Multicore Architectures 2019 SIGMOD 6.3180444e-05
5,310 Low-Latency Handshake Join 2014 VLDB 6.2702891e-05
6,349 Grizzly: Efficient Stream Processing Through Adaptive Query Compilation 2020 SIGMOD 5.9049304e-05
7,959 Terabyte-Scale Analytics in the Blink of an Eye 2026 VLDB 5.5181056e-05
8,072 Operational Analytics Data Management Systems 2016 VLDB 5.4933919e-05
8,234 The Case for Learned In-Memory Joins 2023 VLDB 5.460955e-05
8,609 CXL Memory Performance for In-Memory Data Processing 2025 VLDB 5.4023412e-05
9,288 How to Stop Under-Utilization and Love Multicores 2014 SIGMOD 5.2912652e-05
9,986 Thriving in the No Man’s Land between Compilers and Databases 2019 CIDR 5.1832796e-05
10,479 P-MOSS: Scheduling Main-Memory Indexes Over NUMA Servers Using Next Token Prediction 2026 SIGMOD 5.093636e-05
11,360 Templating Shuffles 2023 CIDR 5.093636e-05
12,260 Next Generation Data Analytics at IBM Research 2013 VLDB 5.093636e-05
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 9 of 9 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Previous Page 1 / 1 Next

Semantically Similar Papers