DBScholar

Back to papers

Dumpy: A Compact and Adaptive Index for Large Data Series Collections

Summary: Dumpy is a compact, adaptive multi-ary index for large data-series collections, enabling fast index building and high-accuracy search. By addressing iSAX limitations—proximity-compactness trade-offs and skew—via adaptive node splitting and Dumpy-Fuzzy duplication, it achieves better efficiency, scalability, and accuracy. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
h1587a0b84be3e747
Venue
SIGMOD
Year
2023
Pagerank
5.732933e-05
Overall Rank
6,612 | 55.57%
DOI
10.1145/3588965

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{wang_sigmod23,
        title = {{Dumpy: A Compact and Adaptive Index for Large Data Series Collections}},
        author = {Wang, Zeyu and Wang, Qitong and Wang, Peng and Palpanas, Themis and Wang, Wei},
        series = {{SIGMOD} '23},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/3588965},
        url = {https://dl.acm.org/doi/10.1145/3588965},
        year = {2023}
}

Incoming Citations (Sorted by Pagerank)

Showing 16 of 16 citing papers.

Rank Citing Paper Year Venue Pagerank
803 RaBitQ: Quantizing High-Dimensional Vectors with a Theoretical Error Bound for Approximate Nearest Neighbor Search 2024 SIGMOD 0.00013838349
2,264 ELPIS: Graph-Based Similarity Search for Scalable Data Science 2023 VLDB 8.7286407e-05
4,600 Steiner-Hardness: A Query Hardness Measure for Graph-Based ANN Indexes 2024 VLDB 6.5055567e-05
5,530 DET-LSH: A Locality-Sensitive Hashing Scheme with Dynamic Encoding Tree for Approximate Nearest Neighbor Search 2024 VLDB 6.0907372e-05
7,267 Subspace Collision: An Efficient and Accurate Framework for High-dimensional Approximate Nearest Neighbor Search 2025 SIGMOD 5.5677188e-05
7,435 Cracking Vector Search Indexes 2025 VLDB 5.5266018e-05
9,146 DARTH: Declarative Recall Through Early Termination for Approximate Nearest Neighbor Search 2026 SIGMOD 5.2199557e-05
9,449 Odyssey: A Journey in the Land of Distributed Data Series Similarity Search 2023 VLDB 5.1747125e-05
9,571 LeaFi: Data Series Indexes on Steroids with Learned Filters 2025 SIGMOD 5.154741e-05
9,585 iEDeaL: A Deep Learning Framework for Detecting Highly Imbalanced Interictal Epileptiform Discharges 2023 VLDB 5.154741e-05
9,917 Cardinality Estimation for Similarity Search on High-Dimensional Data Objects: The Impact of Reference Objects 2025 VLDB 5.1079647e-05
10,125 DIDS: Double Indices and Double Summarizations for Fast Similarity Search 2024 VLDB 5.0747281e-05
10,520 TaCo: Data-adaptive and Query-aware Subspace Collision for High-dimensional Approximate Nearest Neighbor Search 2026 SIGMOD 4.9769913e-05
10,909 ANNiE: A Learned Query Cost Estimator for Graph-Based Approximate Nearest Neighbor Search 2026 VLDB 4.9769913e-05
11,463 Representative Time Series Discovery for Data Exploration 2025 VLDB 4.9769913e-05
11,575 CIVET: Exploring Compact Index for Variable-Length Subsequence Matching on Time Series 2024 VLDB 4.9769913e-05
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 20 of 20 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
74 Fast Approximate Nearest Neighbor Search With The Navigating Spreading-out Graph 2019 VLDB 0.00037145866
297 Query-Aware Locality-Sensitive Hashing for Approximate Nearest Neighbor Search 2016 VLDB 0.00021849337
344 A Comprehensive Survey and Experimental Comparison of Graph-Based Approximate Nearest Neighbor Search 2021 VLDB 0.00020455839
562 SRS: Solving c-Approximate Nearest Neighbor Queries in High Dimensional Euclidean Space with a Tiny Index 2015 VLDB 0.00016350316
648 HD-Index: Pushing the Scalability-Accuracy Boundary for Approximate kNN Search in High-Dimensional Spaces 2018 VLDB 0.00015156941
916 Querying and Mining of Time Series Data: Experimental Comparison of Representations and Distance Measures 2008 VLDB 0.00013093656
1,094 A Data-adaptive and Dynamic Segmentation Index for Whole Matching on Time Series 2013 VLDB 0.00012053453
1,894 iDEC: Indexable Distance Estimating Codes for Approximate Nearest Neighbor Search 2020 VLDB 9.4129766e-05
2,264 ELPIS: Graph-Based Similarity Search for Scalable Data Science 2023 VLDB 8.7286407e-05
2,376 Series2Graph: Graph-based Subsequence Anomaly Detection for Time Series 2020 VLDB 8.5517908e-05
2,730 Return of the Lernaean Hydra: Experimental Evaluation of Data Series Approximate Similarity Search 2020 VLDB 8.0842027e-05
2,870 The Lernaean Hydra of Data Series Similarity Search: An Experimental Evaluation of the State of the Art 2019 VLDB 7.919641e-05
2,971 A Decade of Progress in Indexing and Mining Large Time Series Databases 2006 VLDB 7.7949928e-05
3,513 Scalable, Variable-Length Similarity Search in Data Series: The ULISSE Approach 2018 VLDB 7.2404521e-05
4,685 Coconut: A Scalable Bottom-Up Approach for Building Data Series Indexes 2018 VLDB 6.4692031e-05
4,989 Indexing for Interactive Exploration of Big Data Series 2014 SIGMOD 6.3234684e-05
5,318 Hercules Against Data Series Similarity Search 2022 VLDB 6.1818367e-05
9,006 ANN Softmax: Acceleration of Extreme Classification Training 2022 VLDB 5.2375792e-05
9,449 Odyssey: A Journey in the Land of Distributed Data Series Similarity Search 2023 VLDB 5.1747125e-05
9,585 iEDeaL: A Deep Learning Framework for Detecting Highly Imbalanced Interictal Epileptiform Discharges 2023 VLDB 5.154741e-05
Previous Page 1 / 1 Next

Semantically Similar Papers