DBScholar

Back to papers

Dumpy: A Compact and Adaptive Index for Large Data Series Collections

Summary: Dumpy is a compact, adaptive multi-ary index for large data-series collections, enabling fast index building and high-accuracy search. By addressing iSAX limitations—proximity-compactness trade-offs and skew—via adaptive node splitting and Dumpy-Fuzzy duplication, it achieves better efficiency, scalability, and accuracy. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
6676
Venue
SIGMOD
Year
2023
Pagerank
5.8011086e-05
Overall Rank
6,691 | 54.10%
DOI
10.1145/3588965

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{wang_sigmod23,
        title = {{Dumpy: A Compact and Adaptive Index for Large Data Series Collections}},
        author = {Wang, Zeyu and Wang, Qitong and Wang, Peng and Palpanas, Themis and Wang, Wei},
        series = {{SIGMOD} '23},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/3588965},
        url = {https://dl.acm.org/doi/10.1145/3588965},
        year = {2023}
}

Incoming Citations (Sorted by Pagerank)

Showing 15 of 15 citing papers.

Rank Citing Paper Year Venue Pagerank
1,357 RaBitQ: Quantizing High-Dimensional Vectors with a Theoretical Error Bound for Approximate Nearest Neighbor Search 2024 SIGMOD 0.00011043994
2,534 ELPIS: Graph-Based Similarity Search for Scalable Data Science 2023 VLDB 8.4561875e-05
5,800 DET-LSH: A Locality-Sensitive Hashing Scheme with Dynamic Encoding Tree for Approximate Nearest Neighbor Search 2024 VLDB 6.0850924e-05
6,148 Steiner-Hardness: A Query Hardness Measure for Graph-Based ANN Indexes 2024 VLDB 5.9593368e-05
7,145 Subspace Collision: An Efficient and Accurate Framework for High-dimensional Approximate Nearest Neighbor Search 2025 SIGMOD 5.6908957e-05
8,898 Cracking Vector Search Indexes 2025 VLDB 5.3495662e-05
9,278 Odyssey: A Journey in the Land of Distributed Data Series Similarity Search 2023 VLDB 5.2936898e-05
9,360 DARTH: Declarative Recall Through Early Termination for Approximate Nearest Neighbor Search 2026 SIGMOD 5.2819088e-05
9,377 LeaFi: Data Series Indexes on Steroids with Learned Filters 2025 SIGMOD 5.2755515e-05
9,394 iEDeaL: A Deep Learning Framework for Detecting Highly Imbalanced Interictal Epileptiform Discharges 2023 VLDB 5.2755515e-05
9,964 DIDS: Double Indices and Double Summarizations for Fast Similarity Search 2024 VLDB 5.18753e-05
10,297 TaCo: Data-adaptive and Query-aware Subspace Collision for High-dimensional Approximate Nearest Neighbor Search 2026 SIGMOD 5.093636e-05
11,059 Cardinality Estimation for Similarity Search on High-Dimensional Data Objects: The Impact of Reference Objects 2025 VLDB 5.093636e-05
11,107 Representative Time Series Discovery for Data Exploration 2025 VLDB 5.093636e-05
11,233 CIVET: Exploring Compact Index for Variable-Length Subsequence Matching on Time Series 2024 VLDB 5.093636e-05
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 20 of 20 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
93 Fast Approximate Nearest Neighbor Search With The Navigating Spreading-out Graph 2019 VLDB 0.00034701237
332 Query-Aware Locality-Sensitive Hashing for Approximate Nearest Neighbor Search 2016 VLDB 0.00020920444
398 A Comprehensive Survey and Experimental Comparison of Graph-Based Approximate Nearest Neighbor Search 2021 VLDB 0.00019194947
580 SRS: Solving c-Approximate Nearest Neighbor Queries in High Dimensional Euclidean Space with a Tiny Index 2015 VLDB 0.00016157635
705 HD-Index: Pushing the Scalability-Accuracy Boundary for Approximate kNN Search in High-Dimensional Spaces 2018 VLDB 0.00014829964
898 Querying and Mining of Time Series Data: Experimental Comparison of Representations and Distance Measures 2008 VLDB 0.00013339042
1,084 A Data-adaptive and Dynamic Segmentation Index for Whole Matching on Time Series 2013 VLDB 0.00012256753
2,010 iDEC: Indexable Distance Estimating Codes for Approximate Nearest Neighbor Search 2020 VLDB 9.3085202e-05
2,346 Series2Graph: Graph-based Subsequence Anomaly Detection for Time Series 2020 VLDB 8.7168932e-05
2,534 ELPIS: Graph-Based Similarity Search for Scalable Data Science 2023 VLDB 8.4561875e-05
2,734 Return of the Lernaean Hydra: Experimental Evaluation of Data Series Approximate Similarity Search 2020 VLDB 8.190416e-05
2,928 A Decade of Progress in Indexing and Mining Large Time Series Databases 2006 VLDB 7.9526656e-05
3,037 The Lernaean Hydra of Data Series Similarity Search: An Experimental Evaluation of the State of the Art 2019 VLDB 7.8275859e-05
3,486 Scalable, Variable-Length Similarity Search in Data Series: The ULISSE Approach 2018 VLDB 7.3696676e-05
4,665 Coconut: A Scalable Bottom-Up Approach for Building Data Series Indexes 2018 VLDB 6.5780693e-05
4,883 Indexing for Interactive Exploration of Big Data Series 2014 SIGMOD 6.4648124e-05
5,231 Hercules Against Data Series Similarity Search 2022 VLDB 6.3068064e-05
8,855 ANN Softmax: Acceleration of Extreme Classification Training 2022 VLDB 5.3573227e-05
9,278 Odyssey: A Journey in the Land of Distributed Data Series Similarity Search 2023 VLDB 5.2936898e-05
9,394 iEDeaL: A Deep Learning Framework for Detecting Highly Imbalanced Interictal Epileptiform Discharges 2023 VLDB 5.2755515e-05
Previous Page 1 / 1 Next

Semantically Similar Papers