Bellwether Analysis: Predicting Global Aggregates from Local Regions
Summary: Bellwether analysis for predicting global aggregates from local regions by learning from query-defined labels on partially labeled data. It uses small, cost-efficient bellwether subsets to project those labels onto future data without re-labelling the entire corpus. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Bee-Chung Chen (University of Wisconsin)
- 2. Raghu Ramakrishnan (University of Wisconsin; Yahoo)
- 3. Jude W. Shavlik (University of Wisconsin)
- 4. Pradeep Tamma (University of Wisconsin)
BibTeX Citation
@article{chen_vldb06,
title = {{Bellwether Analysis: Predicting Global Aggregates from Local Regions}},
author = {Chen, Bee-Chung and Ramakrishnan, Raghu and Shavlik, Jude W. and Tamma, Pradeep},
journal = {PVLDB},
series = {{VLDB} '06},
pages = {655--666},
year = {2006}
}
Incoming Citations (Sorted by Pagerank)
Showing 1 of 1 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 2,840 | Towards Keyword-Driven Analytical Processing | 2007 | SIGMOD | 7.9505128e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 7 of 7 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 11 | Implementing Data Cubes Efficiently | 1996 | SIGMOD | 0.00071084324 |
| 403 | Bottom-Up Computation of Sparse and Iceberg CUBEs | 1999 | SIGMOD | 0.0001910396 |
| 1,874 | Efficient Computation of Iceberg Cubes with Complex Measures | 2001 | SIGMOD | 9.4579757e-05 |
| 2,031 | RainForest - A Framework for Fast Decision Tree Construction of Large Datasets | 1998 | VLDB | 9.1572841e-05 |
| 2,208 | Multi-Dimensional Regression Analysis of Time-Series Data Streams | 2002 | VLDB | 8.8382452e-05 |
| 6,543 | Prediction Cubes | 2005 | VLDB | 5.7519742e-05 |
| 13,003 | Composite Subset Measures | 2006 | VLDB | 4.9793485e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 6,783 | Database Workload Capacity Planning using Time Series Analysis and Machine Learning | 2020 | SIGMOD |
| 2 | 5,386 | PREDIcT: Towards Predicting the Runtime of Large Scale Iterative Analytics | 2013 | VLDB |
| 3 | 7,114 | Uncertain Centroid based Partitional Clustering of Uncertain Data | 2012 | VLDB |
| 4 | 2,692 | A Model-based Approach to Attributed Graph Clustering | 2012 | SIGMOD |
| 5 | 1,653 | Scalable Techniques for Mining Causal Structures | 1998 | VLDB |
| 6 | 8,276 | Consistent and Flexible Selectivity Estimation for High-Dimensional Data | 2021 | SIGMOD |
| 7 | 281 | Accelerating Machine Learning Inference with Probabilistic Predicates | 2018 | SIGMOD |
| 8 | 11,952 | Grouped Learning: Group-By Model Selection Workloads | 2021 | SIGMOD |
| 9 | 6,143 | Efficient Construction of Approximate Ad-Hoc ML models Through Materialization and Reuse | 2018 | VLDB |
| 10 | 12,145 | Query-Driven Learning for Next Generation Predictive Modeling & Analytics | 2019 | SIGMOD |