DBScholar

Back to papers

Multi-Core, Main-Memory Joins: Sort vs. Hash Revisited

Summary: Extensive multi-core/NUMA experiments overturn claims that SIMD and NUMA favor sort-merge: optimized radix-hash join remains clearly faster, except at very large scale. Provides state-of-the-art implementations and hardware insights for parallel operators. (summarized by gpt-5.6-luna on Jul 24 2026)

Paper ID
h8b5d1916826b4eb5
Venue
VLDB
Year
2014
Pagerank
0.00023143736
Overall Rank
251 | 98.32%
DOI
10.14778/2732232.2732236

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{balkesen_vldb14,
        title = {{Multi-Core, Main-Memory Joins: Sort vs. Hash Revisited}},
        author = {Balkesen, Cagri and Alonso, Gustavo and Teubner, Jens and Özsu, M. Tamer},
        journal = {PVLDB},
        series = {{VLDB} '14},
        volume = {7},
        number = {1},
        pages = {85--96},
        doi = {10.14778/2732232.2732236},
        url = {https://doi.org/10.14778/2732232.2732236},
        year = {2014}
}

Incoming Citations (Sorted by Pagerank)

Showing 32 of 82 citing papers.

Rank Citing Paper Year Venue Pagerank
7,495 Fast Multi-Column Sorting in Main-Memory Column-Stores 2016 SIGMOD 5.5091753e-05
7,814 Analyzing Vectorized Hash Tables Across CPU Architectures 2023 VLDB 5.4482093e-05
7,866 Building Advanced SQL Analytics From Low-Level Plan Operators 2021 SIGMOD 5.4365876e-05
7,923 NOCAP: Near-Optimal Correlation-Aware Partitioning Joins 2023 SIGMOD 5.4260253e-05
7,956 Main Memory Adaptive Denormalization 2016 SIGMOD 5.4184993e-05
7,964 Parallelizing Intra-Window Join on Multicores: An Experimental Study 2021 SIGMOD 5.4165494e-05
8,241 Pea Hash: A Performant Extendible Adaptive Hashing Index 2023 SIGMOD 5.370431e-05
8,403 The Case for Learned In-Memory Joins 2023 VLDB 5.3389852e-05
8,493 SPRINTER: A Fast n-ary Join Query Processing Method for Complex OLAP Queries 2020 SIGMOD 5.3310013e-05
8,628 Inferray: fast in-memory RDF inference 2016 VLDB 5.2995392e-05
8,673 Adaptive Code Generation for Data-Intensive Analytics 2021 VLDB 5.2913671e-05
8,796 GPH: An Efficient and Effective Perfect Hashing Scheme for GPU Architectures 2025 SIGMOD 5.2748571e-05
9,003 Entropy-Learned Hashing: Constant Time Hashing with Controllable Uniformity 2022 SIGMOD 5.2392472e-05
9,064 A Design Space Exploration and Evaluation for Main-Memory Hash Joins in Storage Class Memory 2023 VLDB 5.2283159e-05
9,201 An Application-Specific Instruction Set for Accelerating Set-Oriented Database Primitives 2014 SIGMOD 5.2083769e-05
9,429 Efficiently Joining Large Relations on Multi-GPU Systems 2025 VLDB 5.1786456e-05
9,456 How to Stop Under-Utilization and Love Multicores 2014 SIGMOD 5.1735054e-05
9,553 Engineering High-Performance Database Engines 2014 VLDB 5.1585591e-05
10,174 Thriving in the No Man’s Land between Compilers and Databases 2019 CIDR 5.0681899e-05
10,312 RAPID: In-Memory Analytical Query Processing Engine with Extreme Performance per Watt 2018 SIGMOD 5.0386264e-05
10,601 TQEx: Tensor-based Query Engine Enhanced by Bridging the Gap 2026 SIGMOD 4.9793485e-05
10,666 P-MOSS: Scheduling Main-Memory Indexes Over NUMA Servers Using Next Token Prediction 2026 SIGMOD 4.9793485e-05
10,946 Why We Created Yet Another Memory Framework: Understanding MGA’s Role in Next-Gen Database Systems 2026 VLDB 4.9793485e-05
11,191 Nested Parquet Is Flat, Why Not Use It? How To Scan Nested Data With On-the-Fly Key Generation and Joins 2025 SIGMOD 4.9793485e-05
11,536 Enabling Adaptive Sampling for Intra-Window Join: Simultaneously Optimizing Quantity and Quality 2024 SIGMOD 4.9793485e-05
11,545 SPID-Join: A Skew-resistant Processing-in-DIMM Join Algorithm Exploiting the Bank- and Rank-level Parallelisms of DIMMs 2024 SIGMOD 4.9793485e-05
11,666 Cache-Efficient Top-k Aggregation over High Cardinality Large Datasets 2024 VLDB 4.9793485e-05
11,750 Cracking-Like Join for Trusted Execution Environments 2023 VLDB 4.9793485e-05
11,865 Scaling Equi-Joins 2022 SIGMOD 4.9793485e-05
12,328 A Study of Sorting Algorithms on Approximate Memory 2016 SIGMOD 4.9793485e-05
12,339 Efficient Query Processing on Many-core Architectures: A Case Study with Intel Xeon Phi Processor 2016 SIGMOD 4.9793485e-05
12,461 Palette: Enabling Scalable Analytics for Big-Memory, Multicore Machines 2014 SIGMOD 4.9793485e-05
Previous Page 2 / 2 Next

Outgoing Citations (Sorted by Pagerank)

Showing 13 of 13 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Previous Page 1 / 1 Next

Semantically Similar Papers