DBScholar

Back to papers

Optimizing Block Skipping for High-Dimensional Data with Learned Adaptive Curve

Summary: Learned adaptive curve for SMA block skipping in high-dimensional data via an attention-based network and end-to-end training. Scales to 1000 columns with a 2.8x block-skipping improvement over static space-filling curves on real Spark workloads. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
h57a59ace520d0107
Venue
SIGMOD
Year
2025
Pagerank
4.9769913e-05
Overall Rank
11,122 | 25.25%
DOI
10.1145/3709710

Incoming Non-self Citations Over Time

No non-self incoming citations found for this paper in this database.

Authors

BibTeX Citation

@inproceedings{chen_sigmod25,
        title = {{Optimizing Block Skipping for High-Dimensional Data with Learned Adaptive Curve}},
        author = {Chen, Xu and Liu, Shuncheng and Yuan, Tong and Ye, Tao and Zeng, Kai and Su, Han and Zheng, Kai},
        series = {{SIGMOD} '25},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/3709710},
        url = {https://dl.acm.org/doi/10.1145/3709710},
        year = {2025}
}

Incoming Citations (Sorted by Pagerank)

Showing 0 of 0 citing papers.

Rank Citing Paper Year Venue Pagerank
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 24 of 24 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
2 R-Trees: A Dynamic Index Structure For Spatial Searching 1984 SIGMOD 0.0019923528
23 Spark SQL: Relational Data Processing in Spark 2015 SIGMOD 0.00055384955
40 The Case for Learned Index Structures 2018 SIGMOD 0.00046363107
85 Learned Cardinalities: Estimating Correlated Joins with Deep Learning 2019 CIDR 0.00035876108
163 DB2 with BLU Acceleration: So Much More than Just a Column Store 2013 VLDB 0.00027480091
216 Small Materialized Aggregates: A Light Weight Index Structure for Data Warehousing 1998 VLDB 0.00024485637
363 Linear Clustering of Objects with Multiple Attributes 1990 SIGMOD 0.00019980483
406 Deep Unsupervised Cardinality Estimation 2020 VLDB 0.00019050182
460 Delta Lake: High-Performance ACID Table Storage over Cloud Object Stores 2020 VLDB 0.00017850215
869 Learning Multi-dimensional Indexes 2020 SIGMOD 0.00013363241
1,036 Fine-grained Partitioning for Aggressive Data Skipping 2014 SIGMOD 0.00012372946
1,128 Qd-tree: Learning Data Layouts for Big Data Analytics 2020 SIGMOD 0.00011901941
1,188 Tsunami: A Learned Multi-dimensional Index for Correlated Data and Skewed Workloads 2021 VLDB 0.00011598149
1,877 Effectively Learning Spatial Indices 2020 VLDB 9.4498401e-05
2,326 Correlation Maps: A Compressed Access Method for Exploiting Soft Functional Dependencies 2009 VLDB 8.628257e-05
2,765 Instance-Optimized Data Layouts for Cloud Analytics Workloads 2021 SIGMOD 8.0401855e-05
3,083 Skipping-oriented Partitioning for Columnar Layouts 2017 VLDB 7.6620866e-05
3,763 Dimensions Based Data Clustering and Zone Maps 2017 VLDB 7.0389893e-05
4,240 LEON: A New Framework for ML-Aided Query Optimization 2023 VLDB 6.7064546e-05
6,027 Automated Multidimensional Data Layouts in Amazon Redshift 2024 SIGMOD 5.9089922e-05
6,470 QUILTS: Multidimensional Partitioning Framework Based on Query-Aware and Skew-Tolerant Space-Filling Curves 2017 SIGMOD 5.7718975e-05
6,597 LMSFC: A Novel Multidimensional Index based on Learned Monotonic Space Filling Curves 2023 VLDB 5.7386302e-05
8,186 Towards Designing and Learning Piecewise Space-Filling Curves 2023 VLDB 5.3800965e-05
9,214 BASE: Bridging the Gap between Cost and Latency for Query Optimization 2023 VLDB 5.2054849e-05
Previous Page 1 / 1 Next

Semantically Similar Papers