YADING: Fast Clustering of Large-Scale Time Series Data
Summary: YADING scales time-series clustering via theoretically bounded sampling, clustering, and assignment, preserving dataset distributions. L1 similarity and multi-density clustering provide robustness to phase shifts/noise, yielding up to 1,000× speedups over DBSCAN/CLARANS. (summarized by gpt-5.6-luna on Jul 24 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Rui Ding (Microsoft)
- 2. Qiang Wang (Microsoft)
- 3. Yingnong Dang (Microsoft)
- 4. Qiang Fu (Microsoft)
- 5. Haidong Zhang (Microsoft)
- 6. Dongmei Zhang (Microsoft)
BibTeX Citation
@article{ding_vldb15,
title = {{YADING: Fast Clustering of Large-Scale Time Series Data}},
author = {Ding, Rui and Wang, Qiang and Dang, Yingnong and Fu, Qiang and Zhang, Haidong and Zhang, Dongmei},
journal = {PVLDB},
series = {{VLDB} '15},
volume = {8},
number = {5},
pages = {473--484},
doi = {10.14778/2735479.2735481},
url = {https://doi.org/10.14778/2735479.2735481},
year = {2015}
}
Incoming Citations (Sorted by Pagerank)
Showing 11 of 11 citing papers.
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 9 of 9 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 4 | The R*-tree: An Efficient and Robust Access Method for Points and Rectangles | 1990 | SIGMOD | 0.001157935 |
| 85 | The X-tree: An Index Structure for High-Dimensional Data | 1996 | VLDB | 0.00035405879 |
| 88 | Efficient and Effective Clustering Methods for Spatial Data Mining | 1994 | VLDB | 0.00035240327 |
| 190 | Locally Adaptive Dimensionality Reduction for Indexing Large Time Series Databases | 2001 | SIGMOD | 0.00026105472 |
| 291 | OPTICS: Ordering Points To Identify the Clustering Structure | 1999 | SIGMOD | 0.00022264197 |
| 351 | CURE: An Efficient Clustering Algorithm for Large Databases | 1998 | SIGMOD | 0.00020424271 |
| 468 | Fast Time Sequence Indexing for Arbitrary Lp Norms | 2000 | VLDB | 0.00017986163 |
| 1,285 | STING: A Statistical Information Grid Approach to Spatial Data Mining | 1997 | VLDB | 0.00011321748 |
| 3,252 | Fast Time-Series Searching with Scaling and Shifting | 1999 | PODS | 7.5951443e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 8,159 | FeatTS: Feature-based Time Series Clustering | 2021 | SIGMOD |
| 2 | 3,626 | Identifying Representative Trends in Massive Time Series Data Sets Using Sketches | 2000 | VLDB |
| 3 | 8,070 | Towards Metric DBSCAN: Exact, Approximate, and Streaming Algorithms | 2024 | SIGMOD |
| 4 | 5,266 | Fast and Scalable Mining of Time Series Motifs with Probabilistic Guarantees | 2022 | VLDB |
| 5 | 3,934 | CRD: Fast Co-clustering on Large Datasets Utilizing Sampling-Based Matrix Decomposition | 2008 | SIGMOD |
| 6 | 5,335 | Clustering Stream Data by Exploring the Evolution of Density Mountain | 2018 | VLDB |
| 7 | 10,667 | In-Database Time Series Clustering | 2025 | SIGMOD |
| 8 | 1,579 | k-Shape: Efficient and Accurate Clustering of Time Series | 2015 | SIGMOD |
| 9 | 907 | A Framework for Clustering Evolving Data Streams | 2003 | VLDB |
| 10 | 10,978 | Time-Series Clustering: A Comprehensive Study of Data Mining, Machine Learning, and Deep Learning Methods | 2025 | VLDB |