DBScholar

Back to papers

Optimizing Block Skipping for High-Dimensional Data with Learned Adaptive Curve

Summary: Learned adaptive curve for SMA block skipping in high-dimensional data via an attention-based network and end-to-end training. Scales to 1000 columns with a 2.8x block-skipping improvement over static space-filling curves on real Spark workloads. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
h57a59ace520d0107
Venue
SIGMOD
Year
2025
Pagerank
4.9793485e-05
Overall Rank
11,113 | 25.29%
DOI
10.1145/3709710

Incoming Non-self Citations Over Time

No non-self incoming citations found for this paper in this database.

Authors

BibTeX Citation

@inproceedings{chen_sigmod25,
        title = {{Optimizing Block Skipping for High-Dimensional Data with Learned Adaptive Curve}},
        author = {Chen, Xu and Liu, Shuncheng and Yuan, Tong and Ye, Tao and Zeng, Kai and Su, Han and Zheng, Kai},
        series = {{SIGMOD} '25},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/3709710},
        url = {https://dl.acm.org/doi/10.1145/3709710},
        year = {2025}
}

Incoming Citations (Sorted by Pagerank)

Showing 0 of 0 citing papers.

Rank Citing Paper Year Venue Pagerank
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 24 of 24 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
2 R-Trees: A Dynamic Index Structure For Spatial Searching 1984 SIGMOD 0.001992968
23 Spark SQL: Relational Data Processing in Spark 2015 SIGMOD 0.00055406774
40 The Case for Learned Index Structures 2018 SIGMOD 0.00046284649
85 Learned Cardinalities: Estimating Correlated Joins with Deep Learning 2019 CIDR 0.00035864347
163 DB2 with BLU Acceleration: So Much More than Just a Column Store 2013 VLDB 0.0002749118
216 Small Materialized Aggregates: A Light Weight Index Structure for Data Warehousing 1998 VLDB 0.00024485024
364 Linear Clustering of Objects with Multiple Attributes 1990 SIGMOD 0.00019976696
406 Deep Unsupervised Cardinality Estimation 2020 VLDB 0.00019045544
459 Delta Lake: High-Performance ACID Table Storage over Cloud Object Stores 2020 VLDB 0.00017856221
868 Learning Multi-dimensional Indexes 2020 SIGMOD 0.00013354403
1,036 Fine-grained Partitioning for Aggressive Data Skipping 2014 SIGMOD 0.00012377471
1,132 Qd-tree: Learning Data Layouts for Big Data Analytics 2020 SIGMOD 0.00011898257
1,191 Tsunami: A Learned Multi-dimensional Index for Correlated Data and Skewed Workloads 2021 VLDB 0.00011590153
1,878 Effectively Learning Spatial Indices 2020 VLDB 9.4451309e-05
2,329 Correlation Maps: A Compressed Access Method for Exploiting Soft Functional Dependencies 2009 VLDB 8.6292256e-05
2,765 Instance-Optimized Data Layouts for Cloud Analytics Workloads 2021 SIGMOD 8.0439015e-05
3,081 Skipping-oriented Partitioning for Columnar Layouts 2017 VLDB 7.6653727e-05
3,762 Dimensions Based Data Clustering and Zone Maps 2017 VLDB 7.0421949e-05
4,258 LEON: A New Framework for ML-Aided Query Optimization 2023 VLDB 6.6994722e-05
6,209 Automated Multidimensional Data Layouts in Amazon Redshift 2024 SIGMOD 5.8495489e-05
6,468 QUILTS: Multidimensional Partitioning Framework Based on Query-Aware and Skew-Tolerant Space-Filling Curves 2017 SIGMOD 5.7745756e-05
6,595 LMSFC: A Novel Multidimensional Index based on Learned Monotonic Space Filling Curves 2023 VLDB 5.7413481e-05
8,179 Towards Designing and Learning Piecewise Space-Filling Curves 2023 VLDB 5.3826446e-05
9,276 BASE: Bridging the Gap between Cost and Latency for Query Optimization 2023 VLDB 5.204289e-05
Previous Page 1 / 1 Next

Semantically Similar Papers