Fast Approximate Correlation for Massive Time-series Data
Summary: Fast all-pair Pearson correlation over massive time-series using DFT and graph partitioning to cut I/O and CPU. Two approximation methods with guarantees: bounded-error similarity and thresholding with no false positives/negatives, plus batch caching; up to 17× faster than prior exact methods. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Abdullah Mueen (University of California Riverside)
- 2. Suman Nath (Microsoft)
- 3. Jie Liu (Microsoft)
BibTeX Citation
@inproceedings{mueen_sigmod10,
title = {{Fast Approximate Correlation for Massive Time-series Data}},
author = {Mueen, Abdullah and Nath, Suman and Liu, Jie},
series = {{SIGMOD} '10},
booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
publisher = {Association for Computing Machinery},
doi = {10.1145/1807167.1807188},
url = {https://dl.acm.org/doi/10.1145/1807167.1807188},
year = {2010}
}
Incoming Citations (Sorted by Pagerank)
Showing 13 of 13 citing papers.
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 10 of 10 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 41 | Fast Subsequence Matching in Time-Series Databases | 1994 | SIGMOD | 0.00046675394 |
| 190 | Locally Adaptive Dimensionality Reduction for Indexing Large Time Series Databases | 2001 | SIGMOD | 0.00026105472 |
| 551 | On Computing Correlated Aggregates Over Continual Data Streams | 2001 | SIGMOD | 0.00016635191 |
| 668 | StatStream: Statistical Monitoring of Thousands of Data Streams in Real Time | 2002 | VLDB | 0.00015166017 |
| 810 | Similarity-Based Queries for Time Series Data | 1997 | SIGMOD | 0.00013874464 |
| 1,252 | Streaming Pattern Discovery in Multiple Time-Series | 2005 | VLDB | 0.00011483752 |
| 2,659 | Identifying Similarities, Periodicities and Bursts for Online Search Queries | 2004 | SIGMOD | 8.286986e-05 |
| 3,391 | BRAID: Stream Mining through Group Lag Correlations | 2005 | SIGMOD | 7.4518238e-05 |
| 4,882 | Managing Massive Time Series Streams with Multi-Scale Compressed Trickles | 2009 | VLDB | 6.4651572e-05 |
| 9,274 | DataGarage: Warehousing Massive Performance Data on Commodity Servers | 2010 | VLDB | 5.2941406e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 8,901 | NLC: Search Correlated Window Pairs on Long Time Series | 2022 | VLDB |
| 2 | 6,780 | Discovering Longest-lasting Correlation in Sequence Databases | 2013 | VLDB |
| 3 | 3,626 | Identifying Representative Trends in Massive Time Series Data Sets Using Sketches | 2000 | VLDB |
| 4 | 10,621 | Scalable Grid-based Computation of Kendall's tau Correlation | 2026 | VLDB |
| 5 | 9,501 | Tracking Set Correlations at Large Scale | 2014 | SIGMOD |
| 6 | 551 | On Computing Correlated Aggregates Over Continual Data Streams | 2001 | SIGMOD |
| 7 | 8,925 | Multivariate Correlations Discovery in Static and Streaming Data | 2022 | VLDB |
| 8 | 668 | StatStream: Statistical Monitoring of Thousands of Data Streams in Real Time | 2002 | VLDB |
| 9 | 11,416 | Correlation Joins over Time Series Data Streams Utilizing Complementary Dimension Reduction and Transformation | 2023 | SIGMOD |
| 10 | 3,920 | Continually Evaluating Similarity-Based Pattern Queries on a Streaming Time Series | 2002 | SIGMOD |