QaaD (Query-as-a-Data): Scalable Execution of Massive Number of Small Queries in Spark
Summary: Query-merging via microRDD converts many small Spark queries into a few larger ones; queries embedded as data enable shared inputs. Dynamic partition sizing minimizes runtime overhead, yielding 10.6x-36.6x speedups over Spark for small queries. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
No non-self incoming citations found for this paper in this database.
Authors
- 1. Yeonsu Park
- 2. Byungchul Tak
- 3. Wook-Shin Han
Incoming Citations (Sorted by Pagerank)
Showing 0 of 0 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 13 of 13 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
Previous
Page 1 / 1
Next
Semantically Similar Papers
| Overall Rank | Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 9,505 | Supporting Scalable Analytics with Latency Constraints | 2015 | VLDB | 4.3300131e-05 |
| 11,951 | A Demonstration of AQWA: Adaptive Query-Workload-Aware Partitioning of Big Spatial Data | 2015 | VLDB | 4.1905499e-05 |
| 8,196 | SparkCruise: Workload Optimization in Managed Spark Clusters at Microsoft | 2021 | VLDB | 4.5568952e-05 |
| 3,536 | Scaling Spark in the Real World: Performance and Usability | 2015 | VLDB | 6.9938207e-05 |
| 9,517 | [Demo] Low-latency Spark Queries on Updatable Data | 2019 | SIGMOD | 4.3294347e-05 |
| 6,683 | Adaptive and Robust Query Execution for Lakehouses at Scale | 2024 | VLDB | 4.9593505e-05 |
| 7,904 | S2RDF: RDF Querying with SPARQL on Spark | 2016 | VLDB | 4.616742e-05 |
| 8,502 | New Query Optimization Techniques in the Spark Engine of Azure Synapse | 2022 | VLDB | 4.491819e-05 |
| 8,585 | A Spark Optimizer for Adaptive, Fine-Grained Parameter Tuning | 2024 | VLDB | 4.4856045e-05 |
| 3,207 | Big Data Analytics with Datalog Queries on Spark | 2016 | SIGMOD | 7.3847098e-05 |