Database Paper Browser

Back to papers

LeaFi: Data Series Indexes on Steroids with Learned Filters

Summary: LeaFi uses learned filters to boost pruning in tree-based data-series indexes. Models predict tight node-wise distance lower bounds for pruning, with train-time index building and query-time calibration to meet per-query recall targets; up to 20x pruning, 32x search speed at 99% recall.— (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
7047
Venue
SIGMOD
Year
2025
Pagerank
4.3648789e-05
Overall Rank
9,233 | 35.84%
DOI
10.1145/3709701

Incoming Non-self Citations Over Time

Authors

Incoming Citations (Sorted by Pagerank)

Showing 3 of 3 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 30 of 30 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
65 Fast Subsequence Matching in Time-Series Databases 1994 SIGMOD 0.00061977022
101 The Case for Learned Index Structures 2018 SIGMOD 0.00049778866
243 Locally Adaptive Dimensionality Reduction for Indexing Large Time Series Databases 2001 SIGMOD 0.00031052437
596 Query-Aware Locality-Sensitive Hashing for Approximate Nearest Neighbor Search 2016 VLDB 0.00019455943
1,157 A Data-adaptive and Dynamic Segmentation Index for Whole Matching on Time Series 2013 VLDB 0.00013600695
1,164 Querying and Mining of Time Series Data: Experimental Comparison of Representations and Distance Measures 2008 VLDB 0.00013572951
1,347 Improving Approximate Nearest Neighbor Search through Learned Adaptive Early Termination 2020 SIGMOD 0.00012463441
1,464 Learning Multi-dimensional Indexes 2020 SIGMOD 0.0001184772
1,510 k-Shape: Efficient and Accurate Clustering of Time Series 2015 SIGMOD 0.00011588558
1,887 Tsunami: A Learned Multi-dimensional Index for Correlated Data and Skewed Workloads 2021 VLDB 0.00010201938
2,381 TSB-UAD: An End-to-End Benchmark Suite for Univariate Time-Series Anomaly Detection 2022 VLDB 8.9241557e-05
3,199 Return of the Lernaean Hydra: Experimental Evaluation of Data Series Approximate Similarity Search 2020 VLDB 7.3999833e-05
3,269 Learned Cardinality Estimation: An In-depth Study 2022 SIGMOD 7.3026051e-05
3,403 ELPIS: Graph-Based Similarity Search for Scalable Data Science 2023 VLDB 7.1338786e-05
3,466 AI Meets Database: AI4DB and DB4AI 2021 SIGMOD 7.0645718e-05
3,544 Scalable, Variable-Length Similarity Search in Data Series: The ULISSE Approach 2018 VLDB 6.98759e-05
3,629 The Lernaean Hydra of Data Series Similarity Search: An Experimental Evaluation of the State of the Art 2019 VLDB 6.8997167e-05
3,992 FactorJoin: A New Cardinality Estimation Framework for Join Queries 2023 SIGMOD 6.5519369e-05
4,464 LOGER: A Learned Optimizer towards Generating Efficient and Robust Query Execution Plans 2023 VLDB 6.1552798e-05
4,524 Data Series Progressive Similarity Search with Probabilistic Quality Guarantees 2020 SIGMOD 6.1091797e-05
4,592 Auto-WLM: Machine Learning Enhanced Workload Management in Amazon Redshift 2023 SIGMOD 6.056004e-05
4,622 Graph-Based Vector Search: An Experimental Evaluation of the State-of-the-Art 2025 SIGMOD 6.0356382e-05
4,751 Indexing for Interactive Exploration of Big Data Series 2014 SIGMOD 5.9411478e-05
5,156 Coconut: A Scalable Bottom-Up Approach for Building Data Series Indexes 2018 VLDB 5.6534878e-05
5,477 Learned Cardinality Estimation for Similarity Queries 2021 SIGMOD 5.4856699e-05
5,747 Hercules Against Data Series Similarity Search 2022 VLDB 5.3427166e-05
6,375 DET-LSH: A Locality-Sensitive Hashing Scheme with Dynamic Encoding Tree for Approximate Nearest Neighbor Search 2024 VLDB 5.0868008e-05
7,090 Dumpy: A Compact and Adaptive Index for Large Data Series Collections 2023 SIGMOD 4.8318862e-05
9,208 Odyssey: A Journey in the Land of Distributed Data Series Similarity Search 2023 VLDB 4.3693005e-05
9,254 iEDeaL: A Deep Learning Framework for Detecting Highly Imbalanced Interictal Epileptiform Discharges 2023 VLDB 4.3648789e-05
Previous Page 1 / 1 Next

Semantically Similar Papers