New Sampling-Based Summary Statistics for Improving Approximate Query Answers
Summary: Introduces concise samples and counting samples—two sampling-based summary statistics for fast approximate answers. Demonstrates fast incremental maintenance across distributions, outperforming standard sample views in view-size efficiency; enables hot-list query speedups under continuous insertions. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
Incoming Citations (Sorted by Pagerank)
Showing 8 of 58 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 8,521 | Computing A Well-Representative Summary of Conjunctive Query Results | 2024 | PODS | 4.4893996e-05 |
| 8,790 | Database Optimization for the Cloud: Where Costs, Partial Results, and Consumer Choice Meet | 2015 | CIDR | 4.4464049e-05 |
| 9,946 | Distributed Wavelet Thresholding for Maximum Error Metrics | 2016 | SIGMOD | 4.240067e-05 |
| 10,365 | Perfect Sampling in Turnstile Streams Beyond Small Moments | 2025 | PODS | 4.1905499e-05 |
| 10,594 | GREAT: Generalized Reservoir Sampling based Triangle Counting Estimation over Streaming Graphs | 2025 | VLDB | 4.1905499e-05 |
| 11,322 | Truly Perfect Samplers for Data Streams and Sliding Windows | 2022 | PODS | 4.1905499e-05 |
| 11,905 | Capturing the Laws of (Data) Nature | 2015 | CIDR | 4.1905499e-05 |
| 12,352 | Composable, Scalable, and Accurate Weight Summarization of Unaggregated Data Sets | 2009 | VLDB | 4.1905499e-05 |
Outgoing Citations (Sorted by Pagerank)
Showing 10 of 10 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 14 | Online Aggregation | 1997 | SIGMOD | 0.0010813443 |
| 60 | Sampling-Based Estimation of the Number of Distinct Values of an Attribute | 1995 | VLDB | 0.00064450997 |
| 63 | Improved Histograms for Selectivity Estimation of Range Predicates | 1996 | SIGMOD | 0.00063595699 |
| 270 | Fast Incremental Maintenance of Approximate Histograms | 1997 | VLDB | 0.00029648047 |
| 328 | Balancing Histogram Optimality and Practicality for Query Result Size Estimation | 1995 | SIGMOD | 0.00027301497 |
| 352 | Random Sampling from B+ trees | 1989 | VLDB | 0.00026276293 |
| 523 | Recovering Information from Summary Data | 1997 | VLDB | 0.00021101376 |
| 651 | Dynamic Itemset Counting and Implication Rules for Market Basket Data | 1997 | SIGMOD | 0.00018649942 |
| 801 | Universality of Serial Histograms | 1993 | VLDB | 0.00016440976 |
| 3,965 | Random Sampling from Pseudo-Ranked B+ Trees | 1992 | VLDB | 6.5786424e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| Overall Rank | Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 37 | Statistical Estimators for Relational Algebra Expressions | 1988 | PODS | 0.00075597514 |
| 212 | Join Synopses for Approximate Query Answering | 1999 | SIGMOD | 0.00033997204 |
| 2,813 | A Robust, Optimization-Based Approach for Approximate Answering of Aggregate Queries | 2001 | SIGMOD | 8.0816314e-05 |
| 46 | Simple Random Sampling from Relational Databases | 1986 | VLDB | 0.00071588702 |
| 92 | Practical Selectivity Estimation through Adaptive Sampling | 1990 | SIGMOD | 0.00051431888 |
| 360 | Histogram-Based Approximation of Set-Valued Query Answers | 1999 | VLDB | 0.00025768448 |
| 2,583 | Sample + Seek: Approximating Aggregates with Distribution Precision Guarantee | 2016 | SIGMOD | 8.4973431e-05 |
| 8,235 | Experiences with Approximating Queries in Microsoft’s Production Big-Data Clusters | 2019 | VLDB | 4.5481384e-05 |
| 8,602 | Structure-Aware Sampling: Flexible and Accurate Summarization | 2011 | VLDB | 4.4822166e-05 |
| 270 | Fast Incremental Maintenance of Approximate Histograms | 1997 | VLDB | 0.00029648047 |