Back to papers
Fast Multi-Column Sorting in Main-Memory Column-Stores
Summary: Proposes code massaging, a bit-level cross-column reordering technique to reduce the number of sorting rounds for multi-column ORDER BY/GROUP BY in main-memory column-stores. Delivers up to 4.7x (TPC-H), 4.7x (TPC-H skew), 4x (TPC-DS), and 3.2x (real workloads) speedups by increasing SIMD parallelism.
(summarized by gpt-5-nano on Feb 09 2026)
- Paper ID
- 5233
- Venue
- SIGMOD
- Year
- 2016
- Pagerank
- 4.8289712e-05
- Overall Rank
- 7,095 | 50.70%
- DOI
-
10.1145/2882903.2915205
Incoming Non-self Citations Over Time
Incoming Citations (Sorted by Pagerank)
Showing 3 of 3 citing papers.
Outgoing Citations (Sorted by Pagerank)
Showing 22 of 22 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank |
Cited Paper |
Year |
Venue |
Pagerank |
| 20 |
C-Store: A Column-oriented DBMS |
2005 |
VLDB |
0.00086163998 |
| 35 |
MonetDB/X100: Hyper-Pipelining Query Execution |
2005 |
CIDR |
0.00076209479 |
| 350 |
Sort vs. Hash Revisited: Fast Join Implementation on Modern Multi-Core CPUs |
2009 |
VLDB |
0.00026368305 |
| 403 |
Multi-Core, Main-Memory Joins: Sort vs. Hash Revisited |
2014 |
VLDB |
0.00024176677 |
| 584 |
Massively Parallel Sort-Merge Joins in Main Memory Multi-Core Database Systems |
2012 |
VLDB |
0.00019700451 |
| 594 |
HYRISE—A Main Memory Hybrid Storage Engine |
2011 |
VLDB |
0.00019515008 |
| 754 |
Database Architecture Evolution: Mammals Flourished long before Dinosaurs became Extinct |
2009 |
VLDB |
0.00017093267 |
| 866 |
Profiling, What-if Analysis, and Cost-based Optimization of MapReduce Programs |
2011 |
VLDB |
0.00015771189 |
| 932 |
Fast Sort on CPUs and GPUs: A Case for Bandwidth Oblivious SIMD Sort |
2010 |
SIGMOD |
0.00015227954 |
| 944 |
Efficient Implementation of Sorting on Multi-Core SIMD CPU Architecture |
2008 |
VLDB |
0.0001512998 |
| 959 |
Rethinking SIMD Vectorization for In-Memory Databases |
2015 |
SIGMOD |
0.00015034808 |
| 1,134 |
Dictionary-based Order-preserving String Compression for Main Memory Column Stores |
2009 |
SIGMOD |
0.00013751593 |
| 1,267 |
BitWeaving: Fast Scans for Main Memory Data Processing |
2013 |
SIGMOD |
0.00012917585 |
| 1,610 |
A Comprehensive Study of Main-Memory Partitioning and its Application to Large-Scale Comparison- and Radix-Sort |
2014 |
SIGMOD |
0.00011155922 |
| 1,622 |
Row-wise Parallel Predicate Evaluation |
2008 |
VLDB |
0.00011104582 |
| 1,726 |
Fast Updates on Read-Optimized Databases Using Multi-Core CPUs |
2012 |
VLDB |
0.00010734339 |
| 2,390 |
ByteSlice: Pushing the Envelop of Main Memory Data Processing with a New Storage Layout |
2015 |
SIGMOD |
8.9006978e-05 |
| 2,409 |
WideTable: An Accelerator for Analytical Data Processing |
2014 |
VLDB |
8.867144e-05 |
| 2,890 |
Database Compression on Graphics Processors |
2010 |
VLDB |
7.9586083e-05 |
| 4,048 |
PARADIS: An Efficient Parallel Algorithm for In-place Radix Sort |
2015 |
VLDB |
6.4970736e-05 |
| 4,649 |
SIMD- and Cache-Friendly Algorithm for Sorting an Array of Structures |
2015 |
VLDB |
6.0171025e-05 |
| 5,541 |
A Padded Encoding Scheme to Accelerate Scans by Leveraging Skew |
2015 |
SIGMOD |
5.4501856e-05 |
Semantically Similar Papers