DBScholar

Back to papers

LeaFi: Data Series Indexes on Steroids with Learned Filters

Summary: LeaFi uses learned filters to boost pruning in tree-based data-series indexes. Models predict tight node-wise distance lower bounds for pruning, with train-time index building and query-time calibration to meet per-query recall targets; up to 20x pruning, 32x search speed at 99% recall.— (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
h71fe1c80267bbb3e
Venue
SIGMOD
Year
2025
Pagerank
5.1571823e-05
Overall Rank
9,563 | 35.71%
DOI
10.1145/3709701

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{wang_sigmod25,
        title = {{LeaFi: Data Series Indexes on Steroids with Learned Filters}},
        author = {Wang, Qitong and Ileana, Ioana and Palpanas, Themis},
        series = {{SIGMOD} '25},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/3709701},
        url = {https://dl.acm.org/doi/10.1145/3709701},
        year = {2025}
}

Incoming Citations (Sorted by Pagerank)

Showing 4 of 4 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 30 of 30 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
40 The Case for Learned Index Structures 2018 SIGMOD 0.00046284649
42 Fast Subsequence Matching in Time-Series Databases 1994 SIGMOD 0.00045773967
193 Locally Adaptive Dimensionality Reduction for Indexing Large Time Series Databases 2001 SIGMOD 0.00025648171
298 Query-Aware Locality-Sensitive Hashing for Approximate Nearest Neighbor Search 2016 VLDB 0.00021833987
839 Improving Approximate Nearest Neighbor Search through Learned Adaptive Early Termination 2020 SIGMOD 0.00013547412
868 Learning Multi-dimensional Indexes 2020 SIGMOD 0.00013354403
915 Querying and Mining of Time Series Data: Experimental Comparison of Representations and Distance Measures 2008 VLDB 0.00013099703
1,095 A Data-adaptive and Dynamic Segmentation Index for Whole Matching on Time Series 2013 VLDB 0.00012050625
1,191 Tsunami: A Learned Multi-dimensional Index for Correlated Data and Skewed Workloads 2021 VLDB 0.00011590153
1,579 k-Shape: Efficient and Accurate Clustering of Time Series 2015 SIGMOD 0.00010186397
1,937 TSB-UAD: An End-to-End Benchmark Suite for Univariate Time-Series Anomaly Detection 2022 VLDB 9.3387043e-05
2,265 ELPIS: Graph-Based Similarity Search for Scalable Data Science 2023 VLDB 8.7238222e-05
2,342 Learned Cardinality Estimation: An In-depth Study 2022 SIGMOD 8.6060437e-05
2,730 Return of the Lernaean Hydra: Experimental Evaluation of Data Series Approximate Similarity Search 2020 VLDB 8.0861221e-05
2,846 FactorJoin: A New Cardinality Estimation Framework for Join Queries 2023 SIGMOD 7.9453616e-05
2,871 Graph-Based Vector Search: An Experimental Evaluation of the State-of-the-Art 2025 SIGMOD 7.9229231e-05
2,872 The Lernaean Hydra of Data Series Similarity Search: An Experimental Evaluation of the State of the Art 2019 VLDB 7.9228762e-05
2,908 AI Meets Database: AI4DB and DB4AI 2021 SIGMOD 7.8742664e-05
3,487 LOGER: A Learned Optimizer towards Generating Efficient and Robust Query Execution Plans 2023 VLDB 7.263041e-05
3,513 Scalable, Variable-Length Similarity Search in Data Series: The ULISSE Approach 2018 VLDB 7.2438813e-05
3,563 Auto-WLM: Machine Learning Enhanced Workload Management in Amazon Redshift 2023 SIGMOD 7.2042148e-05
4,304 Data Series Progressive Similarity Search with Probabilistic Quality Guarantees 2020 SIGMOD 6.6771697e-05
4,680 Coconut: A Scalable Bottom-Up Approach for Building Data Series Indexes 2018 VLDB 6.4722037e-05
4,688 Learned Cardinality Estimation for Similarity Queries 2021 SIGMOD 6.4697463e-05
4,986 Indexing for Interactive Exploration of Big Data Series 2014 SIGMOD 6.3264608e-05
5,313 Hercules Against Data Series Similarity Search 2022 VLDB 6.1846986e-05
5,527 DET-LSH: A Locality-Sensitive Hashing Scheme with Dynamic Encoding Tree for Approximate Nearest Neighbor Search 2024 VLDB 6.0936219e-05
6,609 Dumpy: A Compact and Adaptive Index for Large Data Series Collections 2023 SIGMOD 5.7354155e-05
9,440 Odyssey: A Journey in the Land of Distributed Data Series Similarity Search 2023 VLDB 5.1771634e-05
9,577 iEDeaL: A Deep Learning Framework for Detecting Highly Imbalanced Interictal Epileptiform Discharges 2023 VLDB 5.1571823e-05
Previous Page 1 / 1 Next

Semantically Similar Papers