| 25 |
Spark SQL: Relational Data Processing in Spark |
2015 |
SIGMOD |
0.00055280049 |
| 1,041 |
Fine-grained Partitioning for Aggressive Data Skipping |
2014 |
SIGMOD |
0.00012565893 |
| 1,162 |
Simba: Efficient In-Memory Spatial Analytics |
2016 |
SIGMOD |
0.00011985579 |
| 1,345 |
Scuba: Diving into Data at Facebook |
2013 |
VLDB |
0.00011188833 |
| 1,389 |
Knowing When You’re Wrong: Building Fast and Reliable Approximate Query Processing Systems |
2014 |
SIGMOD |
0.00011029933 |
| 1,470 |
From Theory to Practice: Efficient Join Query Evaluation in a Parallel Database System |
2015 |
SIGMOD |
0.00010746951 |
| 1,577 |
Mesa: Geo-Replicated, Near Real-Time, Scalable Data Warehousing |
2014 |
VLDB |
0.00010368033 |
| 1,669 |
Skew in Parallel Query Processing |
2014 |
PODS |
0.00010137031 |
| 1,885 |
SQL-on-Hadoop: Full Circle Back to Shared-Nothing Database Architectures |
2014 |
VLDB |
9.6357703e-05 |
| 2,131 |
Quickstep: A Data Platform Based on the Scaling-Up Approach |
2018 |
VLDB |
9.1824992e-05 |
| 2,360 |
BigDansing: A System for Big Data Cleansing |
2015 |
SIGMOD |
8.7662564e-05 |
| 2,559 |
Towards Scalable Real-time Analytics: An Architecture for Scale-out of OLxP Workloads |
2015 |
VLDB |
8.4826976e-05 |
| 2,585 |
WideTable: An Accelerator for Analytical Data Processing |
2014 |
VLDB |
8.4540632e-05 |
| 2,804 |
HAWQ: A Massively Parallel Processing SQL Engine in Hadoop |
2014 |
SIGMOD |
8.1680595e-05 |
| 2,916 |
Column Sketches: A Scan Accelerator for Rapid and Robust Predicate Evaluation |
2018 |
SIGMOD |
8.0325296e-05 |
| 3,152 |
Locality-aware Partitioning in Parallel Database Systems |
2015 |
SIGMOD |
7.7601909e-05 |
| 3,453 |
WANalytics: Analytics for a Geo-Distributed Data-Intensive World |
2015 |
CIDR |
7.4746036e-05 |
| 3,553 |
Access Path Selection in Main-Memory Optimized Data Systems: Should I Scan or Should I Probe? |
2017 |
SIGMOD |
7.3775219e-05 |
| 3,650 |
In-RDBMS Hardware Acceleration of Advanced Analytics |
2018 |
VLDB |
7.291319e-05 |
| 4,239 |
WANalytics: Geo-Distributed Analytics for a Data Intensive World |
2015 |
SIGMOD |
6.8751474e-05 |
| 4,419 |
Dynamically Optimizing Queries over Large Scale Data Platforms |
2014 |
SIGMOD |
6.7759069e-05 |
| 4,787 |
Design Tradeoffs of Data Access Methods |
2016 |
SIGMOD |
6.574662e-05 |
| 5,366 |
Fine-Grained Modeling and Optimization for Intelligent Resource Management in Big Data Processing |
2022 |
VLDB |
6.3169865e-05 |
| 5,639 |
Opportunistic Physical Design for Big Data Analytics |
2014 |
SIGMOD |
6.2030946e-05 |
| 6,039 |
A Performance Study of Big Data on Small Nodes |
2015 |
VLDB |
6.0592432e-05 |
| 6,164 |
Elastic Pipelining in an In-Memory Database Cluster |
2016 |
SIGMOD |
6.0304568e-05 |
| 6,247 |
Adaptive and Robust Query Execution for Lakehouses at Scale |
2024 |
VLDB |
6.0034176e-05 |
| 6,248 |
Towards General and Efficient Online Tuning for Spark |
2023 |
VLDB |
6.0022621e-05 |
| 6,536 |
Understanding Insights into the Basic Structure and Essential Issues of Table Placement Methods in Clusters |
2013 |
VLDB |
5.9059287e-05 |
| 6,544 |
Adaptive Data Skipping in Main-Memory Systems |
2016 |
SIGMOD |
5.9038156e-05 |
| 6,629 |
Liquid: Unifying Nearline and Offline Big Data Integration |
2015 |
CIDR |
5.8773042e-05 |
| 6,715 |
JetScope: Reliable and Interactive Analytics at Cloud Scale |
2015 |
VLDB |
5.8520075e-05 |
| 6,785 |
SparkR: Scaling R Programs with Spark |
2016 |
SIGMOD |
5.8328772e-05 |
| 7,037 |
Bubble Execution: Resource-aware Reliable Analytics at Cloud Scale |
2018 |
VLDB |
5.7720294e-05 |
| 7,120 |
Kodiak: Leveraging Materialized Views For Very Low-Latency Analytics Over High-Dimensional Web-Scale Data |
2016 |
VLDB |
5.7519336e-05 |
| 7,368 |
Decentralized Actor Scheduling and Reference-based Storage in Xorbits: a Native Scalable Data Science Engine |
2025 |
VLDB |
5.6897772e-05 |
| 7,583 |
Quill: Efficient, Transferable, and Rich Analytics at Scale |
2016 |
VLDB |
5.6487011e-05 |
| 7,587 |
JoinBoost: Grow Trees Over Normalized Data Using Only SQL |
2023 |
VLDB |
5.6470468e-05 |
| 7,608 |
Using VDMS to Index and Search 100M Images |
2021 |
VLDB |
5.6423505e-05 |
| 8,046 |
SparkCruise: Workload Optimization in Managed Spark Clusters at Microsoft |
2021 |
VLDB |
5.5585663e-05 |
| 8,321 |
Parallel-Correctness and Transferability for Conjunctive Queries |
2015 |
PODS |
5.5080112e-05 |
| 8,329 |
Piranha: Optimizing Short Jobs in Hadoop |
2013 |
VLDB |
5.5065836e-05 |
| 8,472 |
A Spark Optimizer for Adaptive, Fine-Grained Parameter Tuning |
2024 |
VLDB |
5.4888877e-05 |
| 8,951 |
QMapper for Smart Grid: Migrating SQL-based Application to Hive |
2015 |
SIGMOD |
5.4076395e-05 |
| 9,448 |
Cost-based Fault-tolerance for Parallel Data Processing |
2015 |
SIGMOD |
5.3309226e-05 |
| 9,678 |
Introduction to Spark 2.0 for Database Researchers |
2016 |
SIGMOD |
5.293328e-05 |
| 10,415 |
Dynamic Pruning for Recursive Joins |
2025 |
SIGMOD |
5.1725247e-05 |
| 11,695 |
Integration of Large-Scale Data Processing Systems and Traditional Parallel Database Technology |
2019 |
VLDB |
5.1725247e-05 |
| 11,839 |
Logical Aspects of Massively Parallel and Distributed Systems |
2016 |
PODS |
5.1725247e-05 |
| 11,890 |
Parallel Evaluation of Multi-Semi-Joins |
2016 |
VLDB |
5.1725247e-05 |