DBScholar

Back to papers

Optimizing Block Skipping for High-Dimensional Data with Learned Adaptive Curve

Summary: Learned adaptive curve for SMA block skipping in high-dimensional data via an attention-based network and end-to-end training. Scales to 1000 columns with a 2.8x block-skipping improvement over static space-filling curves on real Spark workloads. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
7115
Venue
SIGMOD
Year
2025
Pagerank
5.093636e-05
Overall Rank
10,672 | 26.79%
DOI
10.1145/3709710

Incoming Non-self Citations Over Time

No non-self incoming citations found for this paper in this database.

Authors

BibTeX Citation

@inproceedings{chen_sigmod25,
        title = {{Optimizing Block Skipping for High-Dimensional Data with Learned Adaptive Curve}},
        author = {Chen, Xu and Liu, Shuncheng and Yuan, Tong and Ye, Tao and Zeng, Kai and Su, Han and Zheng, Kai},
        series = {{SIGMOD} '25},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/3709710},
        url = {https://dl.acm.org/doi/10.1145/3709710},
        year = {2025}
}

Incoming Citations (Sorted by Pagerank)

Showing 0 of 0 citing papers.

Rank Citing Paper Year Venue Pagerank
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 24 of 24 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
2 R-Trees: A Dynamic Index Structure For Spatial Searching 1984 SIGMOD 0.0020210012
24 Spark SQL: Relational Data Processing in Spark 2015 SIGMOD 0.00054865648
43 The Case for Learned Index Structures 2018 SIGMOD 0.00046060254
84 Learned Cardinalities: Estimating Correlated Joins with Deep Learning 2019 CIDR 0.00035838391
165 DB2 with BLU Acceleration: So Much More than Just a Column Store 2013 VLDB 0.00027693424
227 Small Materialized Aggregates: A Light Weight Index Structure for Data Warehousing 1998 VLDB 0.00023958508
354 Linear Clustering of Objects with Multiple Attributes 1990 SIGMOD 0.00020355578
401 Deep Unsupervised Cardinality Estimation 2020 VLDB 0.00019092557
520 Delta Lake: High-Performance ACID Table Storage over Cloud Object Stores 2020 VLDB 0.00017136828
873 Learning Multi-dimensional Indexes 2020 SIGMOD 0.00013481915
1,044 Fine-grained Partitioning for Aggressive Data Skipping 2014 SIGMOD 0.0001244236
1,135 Qd-tree: Learning Data Layouts for Big Data Analytics 2020 SIGMOD 0.00012032847
1,174 Tsunami: A Learned Multi-dimensional Index for Correlated Data and Skewed Workloads 2021 VLDB 0.00011817414
1,840 Effectively Learning Spatial Indices 2020 VLDB 9.6404567e-05
2,302 Correlation Maps: A Compressed Access Method for Exploiting Soft Functional Dependencies 2009 VLDB 8.7808696e-05
3,035 Instance-Optimized Data Layouts for Cloud Analytics Workloads 2021 SIGMOD 7.8297746e-05
3,106 Skipping-oriented Partitioning for Columnar Layouts 2017 VLDB 7.7515666e-05
4,329 Dimensions Based Data Clustering and Zone Maps 2017 VLDB 6.75682e-05
4,434 LEON: A New Framework for ML-Aided Query Optimization 2023 VLDB 6.7079088e-05
6,335 QUILTS: Multidimensional Partitioning Framework Based on Query-Aware and Skew-Tolerant Space-Filling Curves 2017 SIGMOD 5.9090498e-05
6,459 LMSFC: A Novel Multidimensional Index based on Learned Monotonic Space Filling Curves 2023 VLDB 5.8727182e-05
7,465 Automated Multidimensional Data Layouts in Amazon Redshift 2024 SIGMOD 5.6108826e-05
8,003 Towards Designing and Learning Piecewise Space-Filling Curves 2023 VLDB 5.5084586e-05
9,123 BASE: Bridging the Gap between Cost and Latency for Query Optimization 2023 VLDB 5.3193264e-05
Previous Page 1 / 1 Next

Semantically Similar Papers