DBScholar

Back to papers

Dumpy: A Compact and Adaptive Index for Large Data Series Collections

Summary: Dumpy is a compact, adaptive multi-ary index for large data-series collections, enabling fast index building and high-accuracy search. By addressing iSAX limitations—proximity-compactness trade-offs and skew—via adaptive node splitting and Dumpy-Fuzzy duplication, it achieves better efficiency, scalability, and accuracy. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
h1587a0b84be3e747
Venue
SIGMOD
Year
2023
Pagerank
5.7354155e-05
Overall Rank
6,609 | 55.57%
DOI
10.1145/3588965

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{wang_sigmod23,
        title = {{Dumpy: A Compact and Adaptive Index for Large Data Series Collections}},
        author = {Wang, Zeyu and Wang, Qitong and Wang, Peng and Palpanas, Themis and Wang, Wei},
        series = {{SIGMOD} '23},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/3588965},
        url = {https://dl.acm.org/doi/10.1145/3588965},
        year = {2023}
}

Incoming Citations (Sorted by Pagerank)

Showing 16 of 16 citing papers.

Rank Citing Paper Year Venue Pagerank
804 RaBitQ: Quantizing High-Dimensional Vectors with a Theoretical Error Bound for Approximate Nearest Neighbor Search 2024 SIGMOD 0.00013832333
2,265 ELPIS: Graph-Based Similarity Search for Scalable Data Science 2023 VLDB 8.7238222e-05
4,598 Steiner-Hardness: A Query Hardness Measure for Graph-Based ANN Indexes 2024 VLDB 6.5086378e-05
5,527 DET-LSH: A Locality-Sensitive Hashing Scheme with Dynamic Encoding Tree for Approximate Nearest Neighbor Search 2024 VLDB 6.0936219e-05
7,264 Subspace Collision: An Efficient and Accurate Framework for High-dimensional Approximate Nearest Neighbor Search 2025 SIGMOD 5.5703557e-05
7,432 Cracking Vector Search Indexes 2025 VLDB 5.5292193e-05
9,136 DARTH: Declarative Recall Through Early Termination for Approximate Nearest Neighbor Search 2026 SIGMOD 5.2224279e-05
9,440 Odyssey: A Journey in the Land of Distributed Data Series Similarity Search 2023 VLDB 5.1771634e-05
9,563 LeaFi: Data Series Indexes on Steroids with Learned Filters 2025 SIGMOD 5.1571823e-05
9,577 iEDeaL: A Deep Learning Framework for Detecting Highly Imbalanced Interictal Epileptiform Discharges 2023 VLDB 5.1571823e-05
9,910 Cardinality Estimation for Similarity Search on High-Dimensional Data Objects: The Impact of Reference Objects 2025 VLDB 5.1103839e-05
10,121 DIDS: Double Indices and Double Summarizations for Fast Similarity Search 2024 VLDB 5.0771316e-05
10,509 TaCo: Data-adaptive and Query-aware Subspace Collision for High-dimensional Approximate Nearest Neighbor Search 2026 SIGMOD 4.9793485e-05
10,900 ANNiE: A Learned Query Cost Estimator for Graph-Based Approximate Nearest Neighbor Search 2026 VLDB 4.9793485e-05
11,457 Representative Time Series Discovery for Data Exploration 2025 VLDB 4.9793485e-05
11,569 CIVET: Exploring Compact Index for Variable-Length Subsequence Matching on Time Series 2024 VLDB 4.9793485e-05
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 20 of 20 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
74 Fast Approximate Nearest Neighbor Search With The Navigating Spreading-out Graph 2019 VLDB 0.00037091678
298 Query-Aware Locality-Sensitive Hashing for Approximate Nearest Neighbor Search 2016 VLDB 0.00021833987
345 A Comprehensive Survey and Experimental Comparison of Graph-Based Approximate Nearest Neighbor Search 2021 VLDB 0.00020445545
562 SRS: Solving c-Approximate Nearest Neighbor Queries in High Dimensional Euclidean Space with a Tiny Index 2015 VLDB 0.00016335405
650 HD-Index: Pushing the Scalability-Accuracy Boundary for Approximate kNN Search in High-Dimensional Spaces 2018 VLDB 0.00015149775
915 Querying and Mining of Time Series Data: Experimental Comparison of Representations and Distance Measures 2008 VLDB 0.00013099703
1,095 A Data-adaptive and Dynamic Segmentation Index for Whole Matching on Time Series 2013 VLDB 0.00012050625
1,897 iDEC: Indexable Distance Estimating Codes for Approximate Nearest Neighbor Search 2020 VLDB 9.4092345e-05
2,265 ELPIS: Graph-Based Similarity Search for Scalable Data Science 2023 VLDB 8.7238222e-05
2,375 Series2Graph: Graph-based Subsequence Anomaly Detection for Time Series 2020 VLDB 8.555841e-05
2,730 Return of the Lernaean Hydra: Experimental Evaluation of Data Series Approximate Similarity Search 2020 VLDB 8.0861221e-05
2,872 The Lernaean Hydra of Data Series Similarity Search: An Experimental Evaluation of the State of the Art 2019 VLDB 7.9228762e-05
2,970 A Decade of Progress in Indexing and Mining Large Time Series Databases 2006 VLDB 7.7982307e-05
3,513 Scalable, Variable-Length Similarity Search in Data Series: The ULISSE Approach 2018 VLDB 7.2438813e-05
4,680 Coconut: A Scalable Bottom-Up Approach for Building Data Series Indexes 2018 VLDB 6.4722037e-05
4,986 Indexing for Interactive Exploration of Big Data Series 2014 SIGMOD 6.3264608e-05
5,313 Hercules Against Data Series Similarity Search 2022 VLDB 6.1846986e-05
8,998 ANN Softmax: Acceleration of Extreme Classification Training 2022 VLDB 5.2400492e-05
9,440 Odyssey: A Journey in the Land of Distributed Data Series Similarity Search 2023 VLDB 5.1771634e-05
9,577 iEDeaL: A Deep Learning Framework for Detecting Highly Imbalanced Interictal Epileptiform Discharges 2023 VLDB 5.1571823e-05
Previous Page 1 / 1 Next

Semantically Similar Papers