DBScholar

Back to papers

Quickr: Lazily Approximating Complex AdHoc Queries in BigData Clusters

Summary: Quickr lazily injects samplers into optimized query plans, approximating complex ad-hoc queries without precomputed samples. Its universe sampler supports multi-input joins, while accuracy analysis preserves groups and bounds aggregates; TPC-DS achieves median 2× resource reduction at cluster scale. (summarized by gpt-5.6-luna on Jul 21 2026)

Paper ID
h7ab7605bf38c4c1d
Venue
SIGMOD
Year
2016
Pagerank
0.0001354605
Overall Rank
840 | 94.36%
DOI
10.1145/2882903.2882940

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{kandula_sigmod16,
        title = {{Quickr: Lazily Approximating Complex AdHoc Queries in BigData Clusters}},
        author = {Kandula, Srikanth and Shanbhag, Anil and Vitorovic, Aleksandar and Olma, Matthaios and Grandl, Robert and Chaudhuri, Surajit and Ding, Bolin},
        series = {{SIGMOD} '16},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/2882903.2882940},
        url = {https://dl.acm.org/doi/10.1145/2882903.2882940},
        year = {2016}
}

Incoming Citations (Sorted by Pagerank)

Showing 50 of 53 citing papers.

Rank Citing Paper Year Venue Pagerank
784 VerdictDB: Universalizing Approximate Query Processing 2018 SIGMOD 0.00014012614
795 Random Sampling over Joins Revisited 2018 SIGMOD 0.00013938779
1,082 Approximate Query Processing: No Silver Bullet 2017 SIGMOD 0.00012122749
1,433 Towards a Learning Optimizer for Shared Clouds 2019 VLDB 0.00010677711
1,465 Pessimistic Cardinality Estimation: Tighter Upper Bounds for Intermediate Join Cardinalities 2019 SIGMOD 0.00010576304
1,678 Two-Level Sampling for Join Size Estimation 2017 SIGMOD 9.9088372e-05
1,829 DBEst: Revisiting Approximate Query Processing Engines with Machine Learning Models 2019 SIGMOD 9.5510333e-05
2,027 Database Learning: Toward a Database that Becomes Smarter Every Time 2017 SIGMOD 9.1618139e-05
3,424 AQP++: Connecting Approximate Query Processing With Aggregate Precomputation for Interactive Analytics 2018 SIGMOD 7.3117029e-05
4,642 Efficient Join Synopsis Maintenance for Data Warehouse 2020 SIGMOD 6.4898745e-05
4,781 Learned Approximate Query Processing: Make it Light, Accurate and Fast 2021 CIDR 6.4162085e-05
5,373 At-the-time and Back-in-time Persistent Sketches 2021 SIGMOD 6.1569337e-05
5,831 Joins on Samples: A Theoretical Guide for Practitioners 2020 VLDB 5.9782109e-05
5,877 BlinkML: Efficient Maximum Likelihood Estimation with Probabilistic Guarantees 2019 SIGMOD 5.9627218e-05
6,221 Combining Aggregation and Sampling (Nearly) Optimally for Approximate Query Processing 2021 SIGMOD 5.8463347e-05
6,245 The Cosmos Big Data Platform at Microsoft: Over a Decade of Progress and a Decade to Look Forward 2021 VLDB 5.837329e-05
6,860 SpareLLM: Automatically Selecting Task-Specific Minimum-Cost Large Language Models under Equivalence Constraint 2025 SIGMOD 5.6613762e-05
7,364 Weighted Distinct Sampling: Cardinality Estimation for SPJ Queries 2021 SIGMOD 5.5418075e-05
7,880 Identifying Insufficient Data Coverage in Databases with Multiple Relations 2020 VLDB 5.43413e-05
7,907 A Practical Approach to Groupjoin and Nested Aggregates 2021 VLDB 5.4287568e-05
7,978 Biathlon: Harnessing Model Resilience for Accelerating ML Inference Pipelines 2024 VLDB 5.4136835e-05
8,166 Fast and Reliable Missing Data Contingency Analysis with Predicate-Constraints 2020 SIGMOD 5.3849926e-05
8,223 PilotDB: Database-Agnostic Online Approximate Query Processing with A Priori Error Guarantees 2025 SIGMOD 5.3751366e-05
8,283 Experiences with Approximating Queries in Microsoft’s Production Big-Data Clusters 2019 VLDB 5.3627138e-05
8,331 LAQy: Efficient and Reusable Query Approximations via Lazy Sampling 2023 SIGMOD 5.3532622e-05
8,363 Probabilistic Database Summarization for Interactive Data Exploration 2017 VLDB 5.3472423e-05
8,493 SPRINTER: A Fast n-ary Join Query Processing Method for Complex OLAP Queries 2020 SIGMOD 5.3310013e-05
8,659 ShadowAQP: Efficient Approximate Group-by and Join Query via Attribute-oriented Sample Size Allocation and Data Generation 2023 VLDB 5.2930951e-05
8,770 One Size Does Not Fit All: A Bandit-Based Sampler Combination Framework with Theoretical Guarantees 2022 SIGMOD 5.2812395e-05
8,846 Data Driven Approximation with Bounded Resources 2017 VLDB 5.2649328e-05
8,944 Practical Dynamic Extension for Sampling Indexes 2023 SIGMOD 5.2542285e-05
9,417 Towards Observability for Production Machine Learning Pipelines 2022 VLDB 5.1803615e-05
9,575 A Step Toward Deep Online Aggregation 2023 SIGMOD 5.1571823e-05
9,764 Sapprox: Enabling Efficient and Accurate Approximations on Sub-datasets with Distribution-aware Online Sampling 2017 VLDB 5.1344728e-05
9,966 Secure Sampling for Approximate Multi-party Query Processing 2023 SIGMOD 5.1038322e-05
10,016 The Data Interaction Game 2018 SIGMOD 5.0959403e-05
10,315 AB-tree: Index for Concurrent Random Sampling and Updates 2022 VLDB 5.0377739e-05
10,501 Sketch-based Secure Query Processing for Streaming Data 2026 SIGMOD 4.9793485e-05
10,724 Secure Multi-Party Sampling over Joins 2026 VLDB 4.9793485e-05
11,079 Efficient Approximate Query Processing with Block Sampling 2025 CIDR 4.9793485e-05
11,184 FAAQP: Fast and Accurate Approximate Query Processing based on Bitmap-augmented Sum-Product Network 2025 SIGMOD 4.9793485e-05
11,239 Holistic query Approximation via RL Modeling 2025 VLDB 4.9793485e-05
11,506 PECJ: Stream Window Join on Disorder Data Streams with Proactive Error Compensation 2024 SIGMOD 4.9793485e-05
11,536 Enabling Adaptive Sampling for Intra-Window Join: Simultaneously Optimizing Quantity and Quality 2024 SIGMOD 4.9793485e-05
11,794 Approximate Queries over Concurrent Updates 2023 VLDB 4.9793485e-05
11,932 Accelerating Complex Analytics using Speculation 2021 CIDR 4.9793485e-05
12,005 In the Land of Data Streams where Synopses are Missing, One Framework to Bring Them All 2021 VLDB 4.9793485e-05
12,039 FlashP: An Analytical Pipeline for Real-time Forecasting of Time-Series Relational Data 2021 VLDB 4.9793485e-05
12,052 BitGourmet: Deterministic Approximation via Optimized Bit Selection 2020 CIDR 4.9793485e-05
12,083 Demonstration of BitGourmet: Data Analysis via Deterministic Approximation 2020 SIGMOD 4.9793485e-05
Previous Page 1 / 2 Next

Outgoing Citations (Sorted by Pagerank)

Showing 25 of 25 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
6 Pig Latin: A Not-So-Foreign Language for Data Processing 2008 SIGMOD 0.001052036
9 Online Aggregation 1997 SIGMOD 0.00076195956
23 Spark SQL: Relational Data Processing in Spark 2015 SIGMOD 0.00055406774
26 Models and Issues in Data Stream Systems 2002 PODS 0.00052121228
30 SCOPE: Easy and Efficient Parallel Processing of Massive Data Sets 2008 VLDB 0.00050495102
31 Hive - A Warehousing Solution Over a Map-Reduce Framework 2009 VLDB 0.00049839909
49 Dremel: Interactive Analysis of Web-Scale Datasets 2010 VLDB 0.00043160717
57 On Random Sampling over Joins 1999 SIGMOD 0.00040108301
124 Approximate Frequency Counts over Data Streams 2002 VLDB 0.00030600691
152 Query Processing, Resource Management, and Approximation in a Data Stream Management System 2003 CIDR 0.0002867034
336 The Aqua Approximate Query Answering System 1999 SIGMOD 0.00020657819
519 Random Sampling for Histogram Construction: How much is enough? 1998 SIGMOD 0.00016942879
564 Congressional Samples for Approximate Answering of Group-By Queries 2000 SIGMOD 0.00016296665
707 On Synopses for Distinct-Value Estimation Under Multiset Operations 2007 SIGMOD 0.00014640173
931 Dynamic Sample Selection for Approximate Query Processing 2003 SIGMOD 0.00013011667
1,022 Online Aggregation for Large MapReduce Jobs 2011 VLDB 0.00012438826
1,090 Scalable Approximate Query Processing With The DBO Engine 2007 SIGMOD 0.00012077577
1,428 Knowing When You’re Wrong: Building Fast and Reliable Approximate Query Processing Systems 2014 SIGMOD 0.00010693831
1,607 SciBORQ: Scientific data management with Bounds On Runtime and Quality 2011 CIDR 0.0001008742
1,868 G-OLA: Generalized On-Line Aggregation for Interactive Analysis on Big Data 2015 SIGMOD 9.4754064e-05
1,916 The Analytical Bootstrap: a New Method for Fast Error Estimation in Approximate Query Processing 2014 SIGMOD 9.3837729e-05
2,456 A Sampling Algebra for Aggregate Estimation 2013 VLDB 8.4377192e-05
2,655 A Robust, Optimization-Based Approach for Approximate Answering of Aggregate Queries 2001 SIGMOD 8.1706092e-05
4,792 Error-bounded Sampling for Analytics on Big Sparse Data 2014 VLDB 6.4130671e-05
5,318 Sampling Algorithms in a Stream Operator 2005 SIGMOD 6.1828476e-05
Previous Page 1 / 1 Next

Semantically Similar Papers