YADING: Fast Clustering of Large-Scale Time Series Data
Summary: YADING scales time-series clustering via theoretically bounded sampling, clustering, and assignment, preserving dataset distributions. L1 similarity and multi-density clustering provide robustness to phase shifts/noise, yielding up to 1,000× speedups over DBSCAN/CLARANS. (summarized by gpt-5.6-luna on Jul 24 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Rui Ding (Microsoft)
- 2. Qiang Wang (Microsoft)
- 3. Yingnong Dang (Microsoft)
- 4. Qiang Fu (Microsoft)
- 5. Haidong Zhang (Microsoft)
- 6. Dongmei Zhang (Microsoft)
BibTeX Citation
@article{ding_vldb15,
title = {{YADING: Fast Clustering of Large-Scale Time Series Data}},
author = {Ding, Rui and Wang, Qiang and Dang, Yingnong and Fu, Qiang and Zhang, Haidong and Zhang, Dongmei},
journal = {PVLDB},
series = {{VLDB} '15},
volume = {8},
number = {5},
pages = {473--484},
doi = {10.14778/2735479.2735481},
url = {https://doi.org/10.14778/2735479.2735481},
year = {2015}
}
Incoming Citations (Sorted by Pagerank)
Showing 11 of 11 citing papers.
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 9 of 9 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 4 | The R*-tree: An Efficient and Robust Access Method for Points and Rectangles | 1990 | SIGMOD | 0.0011405675 |
| 90 | The X-tree: An Index Structure for High-Dimensional Data | 1996 | VLDB | 0.00034860244 |
| 94 | Efficient and Effective Clustering Methods for Spatial Data Mining | 1994 | VLDB | 0.00034579889 |
| 193 | Locally Adaptive Dimensionality Reduction for Indexing Large Time Series Databases | 2001 | SIGMOD | 0.00025648171 |
| 300 | OPTICS: Ordering Points To Identify the Clustering Structure | 1999 | SIGMOD | 0.00021810545 |
| 363 | CURE: An Efficient Clustering Algorithm for Large Databases | 1998 | SIGMOD | 0.00019987463 |
| 478 | Fast Time Sequence Indexing for Arbitrary Lp Norms | 2000 | VLDB | 0.00017631293 |
| 1,305 | STING: A Statistical Information Grid Approach to Spatial Data Mining | 1997 | VLDB | 0.00011099619 |
| 3,302 | Fast Time-Series Searching with Scaling and Shifting | 1999 | PODS | 7.4433382e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 8,201 | FeatTS: Feature-based Time Series Clustering | 2021 | SIGMOD |
| 2 | 3,702 | Identifying Representative Trends in Massive Time Series Data Sets Using Sketches | 2000 | VLDB |
| 3 | 8,240 | Towards Metric DBSCAN: Exact, Approximate, and Streaming Algorithms | 2024 | SIGMOD |
| 4 | 5,375 | Fast and Scalable Mining of Time Series Motifs with Probabilistic Guarantees | 2022 | VLDB |
| 5 | 4,005 | CRD: Fast Co-clustering on Large Datasets Utilizing Sampling-Based Matrix Decomposition | 2008 | SIGMOD |
| 6 | 5,440 | Clustering Stream Data by Exploring the Evolution of Density Mountain | 2018 | VLDB |
| 7 | 11,108 | In-Database Time Series Clustering | 2025 | SIGMOD |
| 8 | 1,579 | k-Shape: Efficient and Accurate Clustering of Time Series | 2015 | SIGMOD |
| 9 | 928 | A Framework for Clustering Evolving Data Streams | 2003 | VLDB |
| 10 | 7,094 | Time-Series Clustering: A Comprehensive Study of Data Mining, Machine Learning, and Deep Learning Methods | 2025 | VLDB |