DBScholar

Back to papers

RHEEM: Enabling Cross-Platform Data Processing - May The Big Data Be With You! -

Summary: Rheem enables cross-platform data processing by decoupling applications from execution platforms and partitioning tasks. A cost-based optimizer selects platforms and maps subtasks, with an executor orchestrating multi-platform workflows for lower cost. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
h57b4174d6ca0e37c
Venue
VLDB
Year
2018
Pagerank
7.3269775e-05
Overall Rank
3,402 | 77.14%
DOI
10.14778/3236187.3236195

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{agrawal_vldb18,
        title = {{RHEEM: Enabling Cross-Platform Data Processing - May The Big Data Be With You! -}},
        author = {Agrawal, Divy and Chawla, Sanjay and Contreras-Rojas, Bertty and Elmagarmid, Ahmed and Idris, Yasser and Kaoudi, Zoi and Kruse, Sebastian and Lucas, Ji and Mansour, Essam and Ouzzani, Mourad and Papotti, Paolo and Quiané-Ruiz, Jorge-Arnulfo and Tang, Nan and Thirumuruganathan, Saravanan and Troudi, Anis},
        journal = {PVLDB},
        series = {{VLDB} '18},
        volume = {11},
        number = {11},
        pages = {1414--1427},
        doi = {10.14778/3236187.3236195},
        url = {https://doi.org/10.14778/3236187.3236195},
        year = {2018}
}

Incoming Citations (Sorted by Pagerank)

Showing 22 of 22 citing papers.

Rank Citing Paper Year Venue Pagerank
3,717 Logical and Physical Optimizations for SQL Query Execution over Large Language Models 2025 SIGMOD 7.0712441e-05
5,094 Towards Scalable Hybrid Stores: Constraint-Based Rewriting to the Rescue 2019 SIGMOD 6.2767625e-05
5,234 Babelfish: Efficient Execution of Polyglot Queries 2022 VLDB 6.2159562e-05
6,141 Expand your Training Limits! Generating Training Data for ML-based Data Management 2021 SIGMOD 5.8708409e-05
6,347 ESTOCADA: Towards Scalable Polystore Systems 2020 VLDB 5.8064898e-05
6,840 Skeena: Efficient and Consistent Cross-Engine Transactions 2022 SIGMOD 5.6657341e-05
7,210 Dataset Relationship Management 2019 CIDR 5.5821013e-05
7,869 Blueprinting the Cloud: Unifying and Automatically Optimizing Cloud Data Infrastructures with BRAD 2024 VLDB 5.4342182e-05
9,042 The Power of Nested Parallelism in Big Data Processing – Hitting Three Flies with One Slap – 2021 SIGMOD 5.2310219e-05
9,098 Hyperspace: The Indexing Subsystem of Azure Synapse 2021 VLDB 5.2258409e-05
9,124 Check Out the Big Brain on BRAD: Simplifying Cloud Data Processing with Learned Automated Data Meshes 2023 VLDB 5.2247088e-05
9,325 On-Demand State Separation for Cloud Data Warehousing 2022 VLDB 5.1933808e-05
9,541 Farm Your ML-based Query Optimizer's Food! - Human-Guided Training Data Generation - 2022 CIDR 5.1612414e-05
9,807 Apache Wayang in Action: Enabling Data Systems Integration via a Unified Data Analytics Framework 2025 SIGMOD 5.1233734e-05
9,928 Polyglot Data Management: State of the Art & Open Challenges 2022 VLDB 5.1079647e-05
9,929 Unified Data Analytics: State-of-the-art and Open Problems 2022 VLDB 5.1079647e-05
10,312 Fast and Scalable Data Transfer Across Data Systems 2025 SIGMOD 5.0376863e-05
10,353 Rethinking Query Optimization for Multi-Agent Systems 2027 VLDB 4.9769913e-05
10,743 APEROL: Adaptive Parallel Edge-to-cloud Runtime Optimization for Layered Workflow Execution 2026 VLDB 4.9769913e-05
11,266 Accio: Bolt-on Query Federation 2025 VLDB 4.9769913e-05
11,720 QaaD (Query-as-a-Data): Scalable Execution of Massive Number of Small Queries in Spark 2023 SIGMOD 4.9769913e-05
12,011 In the Land of Data Streams where Synopses are Missing, One Framework to Bring Them All 2021 VLDB 4.9769913e-05
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 22 of 22 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
6 Pig Latin: A Not-So-Foreign Language for Data Processing 2008 SIGMOD 0.0010515896
15 How Good Are Query Optimizers, Really? 2016 VLDB 0.00061067652
43 A Comparison of Approaches to Large-Scale Data Analysis 2009 SIGMOD 0.0004552807
186 Federated Database Systems for Managing Distributed, Heterogeneous, and Autonomous Databases 1991 VLDB 0.00025938943
416 SystemML: Declarative Machine Learning on Spark 2016 VLDB 0.00018650998
471 Robust Query Processing through Progressive Optimization 2004 SIGMOD 0.00017744392
697 NADEEF: A Commodity Data Cleaning System 2013 SIGMOD 0.00014687805
972 The Data Civilizer System 2017 CIDR 0.00012757732
1,193 Weld: A Common Runtime for High Performance Data Analytics 2017 CIDR 0.00011582619
1,752 Split Query Processing in Polybase 2013 SIGMOD 9.7246277e-05
2,197 Opening the Black Boxes in Data Flow Optimization 2012 VLDB 8.8740089e-05
2,424 BigDansing: A System for Big Data Cleansing 2015 SIGMOD 8.483813e-05
2,492 A Demonstration of the BigDAWG Polystore System 2015 VLDB 8.3861255e-05
3,258 How to Fit when No One Size Fits 2013 CIDR 7.4839611e-05
3,296 Lightning Fast and Space Efficient Inequality Joins 2015 VLDB 7.444648e-05
3,339 Optimizing Analytic Data Flows for Multiple Execution Engines 2012 SIGMOD 7.4045097e-05
3,687 The Myria Big Data Management and Analytics System and Cloud Service 2017 CIDR 7.0954736e-05
3,841 MISO: Souping Up Big Data Query Processing with a Multistore System 2014 SIGMOD 6.98877e-05
3,863 Husky: Towards a More Efficient and Expressive Distributed Computing Framework 2016 VLDB 6.961908e-05
5,703 A Demo of the Data Civilizer System 2017 SIGMOD 6.02892e-05
7,144 A Cost-based Optimizer for Gradient Descent Optimization 2017 SIGMOD 5.5979118e-05
10,165 Rheem: Enabling Multi-Platform Task Execution 2016 SIGMOD 5.0664415e-05
Previous Page 1 / 1 Next

Semantically Similar Papers