DBScholar

Back to papers

Congressional Samples for Approximate Answering of Group-By Queries

Summary: Proposes congressional samples, a hybrid of uniform and biased samples, to maximize group-by accuracy under fixed space. One-pass construction with incremental maintenance without accessing the base relation, plus query-rewriting strategies, validated on TPC-D. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
h938db19be409eaed
Venue
SIGMOD
Year
2000
Pagerank
0.00016297598
Overall Rank
564 | 96.22%
DOI
10.1145/342009.335450

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{acharya_sigmod00,
        title = {{Congressional Samples for Approximate Answering of Group-By Queries}},
        author = {Acharya, Swarup and Gibbons, Phillip B. and Poosala, Viswanath},
        series = {{SIGMOD} '00},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/342009.335450},
        url = {https://dl.acm.org/doi/10.1145/342009.335450},
        year = {2000}
}

Incoming Citations (Sorted by Pagerank)

Showing 49 of 49 citing papers.

Rank Citing Paper Year Venue Pagerank
26 Models and Issues in Data Stream Systems 2002 PODS 0.00052097907
267 Distinct Sampling for Highly-Accurate Answers to Distinct Values Queries and Event Reports 2001 VLDB 0.00022713652
335 The Aqua Approximate Query Answering System 1999 SIGMOD 0.000206533
372 Approximate Query Processing: Taming the TeraBytes! A Tutorial 2001 VLDB 0.0001971778
429 Tracking Join and Self-Join Sizes in Limited Storage 1999 PODS 0.00018445263
841 Quickr: Lazily Approximating Complex AdHoc Queries in BigData Clusters 2016 SIGMOD 0.00013543
930 Dynamic Sample Selection for Approximate Query Processing 2003 SIGMOD 0.00013009255
947 Aqua: A Fast Decision Support System Using Approximate Query Answers 1999 VLDB 0.00012920489
1,061 Approximate Query Processing: No Silver Bullet 2017 SIGMOD 0.00012208639
1,244 Compressing SQL Workloads 2002 SIGMOD 0.00011369155
1,408 Northstar: An Interactive Data Science System 2018 VLDB 0.00010738859
1,662 Rapid Sampling for Visualizations with Ordering Guarantees 2015 VLDB 9.9502569e-05
1,722 A Sample-and-Clean Framework for Fast and Accurate Query Processing on Dirty Data 2014 SIGMOD 9.7921604e-05
2,003 Sample + Seek: Approximating Aggregates with Distribution Precision Guarantee 2016 SIGMOD 9.2071735e-05
2,257 Dwarf: Shrinking the PetaCube 2002 SIGMOD 8.7397925e-05
2,295 ThalamusDB: Approximate Query Processing on Multi-Modal Data 2024 SIGMOD 8.6822982e-05
2,324 Online Maintenance of Very Large Random Samples 2004 SIGMOD 8.6332031e-05
2,495 Independent Range Sampling 2014 PODS 8.3834888e-05
2,654 A Robust, Optimization-Based Approach for Approximate Answering of Aggregate Queries 2001 SIGMOD 8.1670397e-05
3,036 MRI: Meaningful Interpretations of Collaborative Ratings 2011 VLDB 7.729485e-05
3,085 Continuous Sampling for Online Aggregation Over Multiple Queries 2010 SIGMOD 7.6607519e-05
3,203 Interactive Data Exploration Using Semantic Windows 2014 SIGMOD 7.5392994e-05
3,213 Turbo-Charging Estimate Convergence in DBO 2009 VLDB 7.5304969e-05
3,417 Revisiting Reuse for Approximate Query Processing 2017 VLDB 7.3184905e-05
3,424 AQP++: Connecting Approximate Query Processing With Aggregate Precomputation for Interactive Analytics 2018 SIGMOD 7.3084429e-05
3,651 Davos: A System for Interactive Data-Driven Decision Making 2021 VLDB 7.1336875e-05
3,988 Exploiting Correlations for Expensive Predicate Evaluation 2015 SIGMOD 6.8694751e-05
4,253 Primitives for Workload Summarization and Implications for SQL 2003 VLDB 6.6985405e-05
4,795 Error-bounded Sampling for Analytics on Big Sparse Data 2014 VLDB 6.4101044e-05
5,025 Adaptive Sampling for Rapidly Matching Histograms 2018 VLDB 6.3078009e-05
5,310 CliffGuard: A Principled Framework for Finding Robust Database Designs 2015 SIGMOD 6.184026e-05
5,590 Derby/S: A DBMS for Sample-Based Query Answering 2006 SIGMOD 6.0712954e-05
5,983 Supporting Time-Constrained SQL Queries in Oracle 2007 VLDB 5.9241225e-05
6,202 Combining Aggregation and Sampling (Nearly) Optimally for Approximate Query Processing 2021 SIGMOD 5.8494367e-05
6,687 Query Sampling in DB2 Universal Database 2004 SIGMOD 5.7060981e-05
6,786 Robust Estimation With Sampling and Approximate Pre-Aggregation 2003 VLDB 5.6811262e-05
6,968 Sampling Dirty Data for Matching Attributes 2010 SIGMOD 5.6296644e-05
7,046 The Polynomial Complexity of Fully Materialized Coalesced Cubes 2004 VLDB 5.6111702e-05
7,707 Enabling Efficient and General Subpopulation Analytics in Multidimensional Data Streams 2022 VLDB 5.4723863e-05
7,749 PilotDB: Database-Agnostic Online Approximate Query Processing with A Priori Error Guarantees 2025 SIGMOD 5.4585522e-05
7,885 Identifying Insufficient Data Coverage in Databases with Multiple Relations 2020 VLDB 5.4315576e-05
8,667 ShadowAQP: Efficient Approximate Group-by and Join Query via Attribute-oriented Sample Size Allocation and Data Generation 2023 VLDB 5.2905894e-05
8,771 One Size Does Not Fit All: A Bandit-Based Sampler Combination Framework with Theoretical Guarantees 2022 SIGMOD 5.2808131e-05
8,851 Data Driven Approximation with Bounded Resources 2017 VLDB 5.2634608e-05
10,514 SmartRabbit: An Interactive Query Processor 2026 SIGMOD 4.9769913e-05
10,710 Sample-based Distinct Cardinality Estimation for Multiple Attributes in Multi-Dataset Queries 2026 VLDB 4.9769913e-05
11,193 FAAQP: Fast and Accurate Approximate Query Processing based on Bitmap-augmented Sum-Product Network 2025 SIGMOD 4.9769913e-05
12,045 FlashP: An Analytical Pipeline for Real-time Forecasting of Time-Series Relational Data 2021 VLDB 4.9769913e-05
13,115 Estimating the Output Cardinality of Partial Preaggregation with a Measure of Clusteredness 2003 VLDB 4.9769913e-05
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 10 of 10 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Previous Page 1 / 1 Next

Semantically Similar Papers