DBScholar

Back to papers

Congressional Samples for Approximate Answering of Group-By Queries

Summary: Proposes congressional samples, a hybrid of uniform and biased samples, to maximize group-by accuracy under fixed space. One-pass construction with incremental maintenance without accessing the base relation, plus query-rewriting strategies, validated on TPC-D. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
3271
Venue
SIGMOD
Year
2000
Pagerank
0.00016590619
Overall Rank
553 | 96.21%
DOI
10.1145/342009.335450

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{acharya_sigmod00,
        title = {{Congressional Samples for Approximate Answering of Group-By Queries}},
        author = {Acharya, Swarup and Gibbons, Phillip B. and Poosala, Viswanath},
        series = {{SIGMOD} '00},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/342009.335450},
        url = {https://dl.acm.org/doi/10.1145/342009.335450},
        year = {2000}
}

Incoming Citations (Sorted by Pagerank)

Showing 49 of 49 citing papers.

Rank Citing Paper Year Venue Pagerank
26 Models and Issues in Data Stream Systems 2002 PODS 0.00052982574
255 Distinct Sampling for Highly-Accurate Answers to Distinct Values Queries and Event Reports 2001 VLDB 0.00023174541
327 The Aqua Approximate Query Answering System 1999 SIGMOD 0.00021091539
363 Approximate Query Processing: Taming the TeraBytes! A Tutorial 2001 VLDB 0.0002005475
418 Tracking Join and Self-Join Sizes in Limited Storage 1999 PODS 0.00018812821
819 Quickr: Lazily Approximating Complex AdHoc Queries in BigData Clusters 2016 SIGMOD 0.00013815639
909 Dynamic Sample Selection for Approximate Query Processing 2003 SIGMOD 0.00013291205
931 Aqua: A Fast Decision Support System Using Approximate Query Answers 1999 VLDB 0.00013125812
1,108 Approximate Query Processing: No Silver Bullet 2017 SIGMOD 0.00012145154
1,266 Compressing SQL Workloads 2002 SIGMOD 0.00011412078
1,392 Northstar: An Interactive Data Science System 2018 VLDB 0.00010936065
1,634 Rapid Sampling for Visualizations with Ordering Guarantees 2015 VLDB 0.00010163938
1,736 A Sample-and-Clean Framework for Fast and Accurate Query Processing on Dirty Data 2014 SIGMOD 9.8984415e-05
1,962 Sample + Seek: Approximating Aggregates with Distribution Precision Guarantee 2016 SIGMOD 9.3978414e-05
2,271 Online Maintenance of Very Large Random Samples 2004 SIGMOD 8.8254873e-05
2,318 Dwarf: Shrinking the PetaCube 2002 SIGMOD 8.7588859e-05
2,443 Independent Range Sampling 2014 PODS 8.5754434e-05
2,608 A Robust, Optimization-Based Approach for Approximate Answering of Aggregate Queries 2001 SIGMOD 8.347674e-05
2,993 MRI: Meaningful Interpretations of Collaborative Ratings 2011 VLDB 7.8804914e-05
3,042 Continuous Sampling for Online Aggregation Over Multiple Queries 2010 SIGMOD 7.8231049e-05
3,144 Interactive Data Exploration Using Semantic Windows 2014 SIGMOD 7.7134729e-05
3,157 Turbo-Charging Estimate Convergence in DBO 2009 VLDB 7.6911286e-05
3,366 AQP++: Connecting Approximate Query Processing With Aggregate Precomputation for Interactive Analytics 2018 SIGMOD 7.4748604e-05
3,370 Revisiting Reuse for Approximate Query Processing 2017 VLDB 7.4700891e-05
3,570 Davos: A System for Interactive Data-Driven Decision Making 2021 VLDB 7.3008782e-05
3,684 ThalamusDB: Approximate Query Processing on Multi-Modal Data 2024 SIGMOD 7.2033959e-05
3,945 Exploiting Correlations for Expensive Predicate Evaluation 2015 SIGMOD 7.0055154e-05
4,198 Primitives for Workload Summarization and Implications for SQL 2003 VLDB 6.8397659e-05
4,696 Error-bounded Sampling for Analytics on Big Sparse Data 2014 VLDB 6.557612e-05
4,926 Adaptive Sampling for Rapidly Matching Histograms 2018 VLDB 6.4428362e-05
5,222 CliffGuard: A Principled Framework for Finding Robust Database Designs 2015 SIGMOD 6.3103741e-05
5,460 Derby/S: A DBMS for Sample-Based Query Answering 2006 SIGMOD 6.2106757e-05
5,863 Supporting Time-Constrained SQL Queries in Oracle 2007 VLDB 6.0626421e-05
6,206 Combining Aggregation and Sampling (Nearly) Optimally for Approximate Query Processing 2021 SIGMOD 5.9443409e-05
6,562 Query Sampling in DB2 Universal Database 2004 SIGMOD 5.838575e-05
6,654 Robust Estimation With Sampling and Approximate Pre-Aggregation 2003 VLDB 5.8131331e-05
6,827 Sampling Dirty Data for Matching Attributes 2010 SIGMOD 5.7616041e-05
6,902 The Polynomial Complexity of Fully Materialized Coalesced Cubes 2004 VLDB 5.7426524e-05
7,551 Enabling Efficient and General Subpopulation Analytics in Multidimensional Data Streams 2022 VLDB 5.6006414e-05
7,721 Identifying Insufficient Data Coverage in Databases with Multiple Relations 2020 VLDB 5.5587371e-05
8,204 PilotDB: Database-Agnostic Online Approximate Query Processing with A Priori Error Guarantees 2025 SIGMOD 5.4667903e-05
8,492 ShadowAQP: Efficient Approximate Group-by and Join Query via Attribute-oriented Sample Size Allocation and Data Generation 2023 VLDB 5.4145838e-05
8,608 One Size Does Not Fit All: A Bandit-Based Sampler Combination Framework with Theoretical Guarantees 2022 SIGMOD 5.4024561e-05
8,701 Data Driven Approximation with Bounded Resources 2017 VLDB 5.3828806e-05
10,291 SmartRabbit: An Interactive Query Processor 2026 SIGMOD 5.093636e-05
10,515 Sample-based Distinct Cardinality Estimation for Multiple Attributes in Multi-Dataset Queries 2026 VLDB 5.093636e-05
10,760 FAAQP: Fast and Accurate Approximate Query Processing based on Bitmap-augmented Sum-Product Network 2025 SIGMOD 5.093636e-05
11,736 FlashP: An Analytical Pipeline for Real-time Forecasting of Time-Series Relational Data 2021 VLDB 5.093636e-05
12,819 Estimating the Output Cardinality of Partial Preaggregation with a Measure of Clusteredness 2003 VLDB 5.093636e-05
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 10 of 10 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Previous Page 1 / 1 Next

Semantically Similar Papers