DBScholar

Back to papers

Quickr: Lazily Approximating Complex AdHoc Queries in BigData Clusters

Summary: Quickr lazily injects samplers into optimized query plans, approximating complex ad-hoc queries without precomputed samples. Its universe sampler supports multi-input joins, while accuracy analysis preserves groups and bounds aggregates; TPC-DS achieves median 2× resource reduction at cluster scale. (summarized by gpt-5.6-luna on Jul 21 2026)

Paper ID
h7ab7605bf38c4c1d
Venue
SIGMOD
Year
2016
Pagerank
0.00013543
Overall Rank
841 | 94.35%
DOI
10.1145/2882903.2882940

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{kandula_sigmod16,
        title = {{Quickr: Lazily Approximating Complex AdHoc Queries in BigData Clusters}},
        author = {Kandula, Srikanth and Shanbhag, Anil and Vitorovic, Aleksandar and Olma, Matthaios and Grandl, Robert and Chaudhuri, Surajit and Ding, Bolin},
        series = {{SIGMOD} '16},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/2882903.2882940},
        url = {https://dl.acm.org/doi/10.1145/2882903.2882940},
        year = {2016}
}

Incoming Citations (Sorted by Pagerank)

Showing 50 of 53 citing papers.

Rank Citing Paper Year Venue Pagerank
772 VerdictDB: Universalizing Approximate Query Processing 2018 SIGMOD 0.0001409096
795 Random Sampling over Joins Revisited 2018 SIGMOD 0.00013934719
1,061 Approximate Query Processing: No Silver Bullet 2017 SIGMOD 0.00012208639
1,432 Towards a Learning Optimizer for Shared Clouds 2019 VLDB 0.00010676754
1,465 Pessimistic Cardinality Estimation: Tighter Upper Bounds for Intermediate Join Cardinalities 2019 SIGMOD 0.00010572023
1,678 Two-Level Sampling for Join Size Estimation 2017 SIGMOD 9.9056116e-05
1,828 DBEst: Revisiting Approximate Query Processing Engines with Machine Learning Models 2019 SIGMOD 9.547768e-05
2,028 Database Learning: Toward a Database that Becomes Smarter Every Time 2017 SIGMOD 9.1584244e-05
3,424 AQP++: Connecting Approximate Query Processing With Aggregate Precomputation for Interactive Analytics 2018 SIGMOD 7.3084429e-05
4,644 Efficient Join Synopsis Maintenance for Data Warehouse 2020 SIGMOD 6.4869417e-05
4,779 Learned Approximate Query Processing: Make it Light, Accurate and Fast 2021 CIDR 6.416435e-05
5,379 At-the-time and Back-in-time Persistent Sketches 2021 SIGMOD 6.1540191e-05
5,829 Joins on Samples: A Theoretical Guide for Practitioners 2020 VLDB 5.9764044e-05
5,877 BlinkML: Efficient Maximum Likelihood Estimation with Probabilistic Guarantees 2019 SIGMOD 5.9599042e-05
6,202 Combining Aggregation and Sampling (Nearly) Optimally for Approximate Query Processing 2021 SIGMOD 5.8494367e-05
6,248 The Cosmos Big Data Platform at Microsoft: Over a Decade of Progress and a Decade to Look Forward 2021 VLDB 5.8345657e-05
6,857 SpareLLM: Automatically Selecting Task-Specific Minimum-Cost Large Language Models under Equivalence Constraint 2025 SIGMOD 5.6621524e-05
7,368 Weighted Distinct Sampling: Cardinality Estimation for SPJ Queries 2021 SIGMOD 5.5392867e-05
7,749 PilotDB: Database-Agnostic Online Approximate Query Processing with A Priori Error Guarantees 2025 SIGMOD 5.4585522e-05
7,885 Identifying Insufficient Data Coverage in Databases with Multiple Relations 2020 VLDB 5.4315576e-05
7,911 A Practical Approach to Groupjoin and Nested Aggregates 2021 VLDB 5.4261869e-05
7,983 Biathlon: Harnessing Model Resilience for Accelerating ML Inference Pipelines 2024 VLDB 5.4111208e-05
8,172 Fast and Reliable Missing Data Contingency Analysis with Predicate-Constraints 2020 SIGMOD 5.3824434e-05
8,289 Experiences with Approximating Queries in Microsoft’s Production Big-Data Clusters 2019 VLDB 5.360349e-05
8,336 LAQy: Efficient and Reusable Query Approximations via Lazy Sampling 2023 SIGMOD 5.3508072e-05
8,367 Probabilistic Database Summarization for Interactive Data Exploration 2017 VLDB 5.344711e-05
8,501 SPRINTER: A Fast n-ary Join Query Processing Method for Complex OLAP Queries 2020 SIGMOD 5.3285575e-05
8,667 ShadowAQP: Efficient Approximate Group-by and Join Query via Attribute-oriented Sample Size Allocation and Data Generation 2023 VLDB 5.2905894e-05
8,771 One Size Does Not Fit All: A Bandit-Based Sampler Combination Framework with Theoretical Guarantees 2022 SIGMOD 5.2808131e-05
8,851 Data Driven Approximation with Bounded Resources 2017 VLDB 5.2634608e-05
8,953 Practical Dynamic Extension for Sampling Indexes 2023 SIGMOD 5.2517412e-05
9,426 Towards Observability for Production Machine Learning Pipelines 2022 VLDB 5.1779092e-05
9,583 A Step Toward Deep Online Aggregation 2023 SIGMOD 5.154741e-05
9,769 Sapprox: Enabling Efficient and Accurate Approximations on Sub-datasets with Distribution-aware Online Sampling 2017 VLDB 5.1320422e-05
9,972 Secure Sampling for Approximate Multi-party Query Processing 2023 SIGMOD 5.1014161e-05
10,021 The Data Interaction Game 2018 SIGMOD 5.0935279e-05
10,317 AB-tree: Index for Concurrent Random Sampling and Updates 2022 VLDB 5.0363234e-05
10,512 Sketch-based Secure Query Processing for Streaming Data 2026 SIGMOD 4.9769913e-05
10,734 Secure Multi-Party Sampling over Joins 2026 VLDB 4.9769913e-05
11,088 Efficient Approximate Query Processing with Block Sampling 2025 CIDR 4.9769913e-05
11,193 FAAQP: Fast and Accurate Approximate Query Processing based on Bitmap-augmented Sum-Product Network 2025 SIGMOD 4.9769913e-05
11,247 Holistic query Approximation via RL Modeling 2025 VLDB 4.9769913e-05
11,512 PECJ: Stream Window Join on Disorder Data Streams with Proactive Error Compensation 2024 SIGMOD 4.9769913e-05
11,542 Enabling Adaptive Sampling for Intra-Window Join: Simultaneously Optimizing Quantity and Quality 2024 SIGMOD 4.9769913e-05
11,800 Approximate Queries over Concurrent Updates 2023 VLDB 4.9769913e-05
11,938 Accelerating Complex Analytics using Speculation 2021 CIDR 4.9769913e-05
12,011 In the Land of Data Streams where Synopses are Missing, One Framework to Bring Them All 2021 VLDB 4.9769913e-05
12,045 FlashP: An Analytical Pipeline for Real-time Forecasting of Time-Series Relational Data 2021 VLDB 4.9769913e-05
12,058 BitGourmet: Deterministic Approximation via Optimized Bit Selection 2020 CIDR 4.9769913e-05
12,089 Demonstration of BitGourmet: Data Analysis via Deterministic Approximation 2020 SIGMOD 4.9769913e-05
Previous Page 1 / 2 Next

Outgoing Citations (Sorted by Pagerank)

Showing 25 of 25 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
6 Pig Latin: A Not-So-Foreign Language for Data Processing 2008 SIGMOD 0.0010515896
9 Online Aggregation 1997 SIGMOD 0.00076265429
23 Spark SQL: Relational Data Processing in Spark 2015 SIGMOD 0.00055384955
26 Models and Issues in Data Stream Systems 2002 PODS 0.00052097907
30 SCOPE: Easy and Efficient Parallel Processing of Massive Data Sets 2008 VLDB 0.00050475202
31 Hive - A Warehousing Solution Over a Map-Reduce Framework 2009 VLDB 0.00049821554
49 Dremel: Interactive Analysis of Web-Scale Datasets 2010 VLDB 0.0004314366
57 On Random Sampling over Joins 1999 SIGMOD 0.00040095727
124 Approximate Frequency Counts over Data Streams 2002 VLDB 0.00030586757
152 Query Processing, Resource Management, and Approximation in a Data Stream Management System 2003 CIDR 0.00028657752
335 The Aqua Approximate Query Answering System 1999 SIGMOD 0.000206533
518 Random Sampling for Histogram Construction: How much is enough? 1998 SIGMOD 0.00016938992
564 Congressional Samples for Approximate Answering of Group-By Queries 2000 SIGMOD 0.00016297598
707 On Synopses for Distinct-Value Estimation Under Multiset Operations 2007 SIGMOD 0.00014633741
930 Dynamic Sample Selection for Approximate Query Processing 2003 SIGMOD 0.00013009255
1,021 Online Aggregation for Large MapReduce Jobs 2011 VLDB 0.00012437619
1,090 Scalable Approximate Query Processing With The DBO Engine 2007 SIGMOD 0.00012074369
1,428 Knowing When You’re Wrong: Building Fast and Reliable Approximate Query Processing Systems 2014 SIGMOD 0.0001069161
1,608 SciBORQ: Scientific data management with Bounds On Runtime and Quality 2011 CIDR 0.00010085907
1,869 G-OLA: Generalized On-Line Aggregation for Interactive Analysis on Big Data 2015 SIGMOD 9.4749419e-05
1,916 The Analytical Bootstrap: a New Method for Fast Error Estimation in Approximate Query Processing 2014 SIGMOD 9.3822742e-05
2,455 A Sampling Algebra for Aggregate Estimation 2013 VLDB 8.4377251e-05
2,654 A Robust, Optimization-Based Approach for Approximate Answering of Aggregate Queries 2001 SIGMOD 8.1670397e-05
4,795 Error-bounded Sampling for Analytics on Big Sparse Data 2014 VLDB 6.4101044e-05
5,325 Sampling Algorithms in a Stream Operator 2005 SIGMOD 6.1799938e-05
Previous Page 1 / 1 Next

Semantically Similar Papers