Hadoop++: Making a Yellow Elephant Run Like a Cheetah (Without It Even Noticing)
Summary: Hadoop++ transparently accelerates Hadoop by injecting indexing and join optimizations through UDFs, without modifying its framework or interface. It substantially outperforms Hadoop and HadoopDB while remaining compatible with future Hadoop changes. (summarized by gpt-5.6-luna on Jul 24 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Jens Dittrich (Saarland University)
- 2. Yagiz Kargin (International Max Planck Research School for Computer Science)
- 3. Jorge-Arnulfo Quiané-Ruiz (Saarland University)
- 4. Vinay Setty (International Max Planck Research School for Computer Science)
- 5. Alekh Jindal (International Max Planck Research School for Computer Science; Saarland University)
- 6. Jörg Schad (Saarland University)
BibTeX Citation
@article{dittrich_vldb10,
title = {{Hadoop++: Making a Yellow Elephant Run Like a Cheetah (Without It Even Noticing)}},
author = {Dittrich, Jens and Kargin, Yagiz and Quiané-Ruiz, Jorge-Arnulfo and Setty, Vinay and Jindal, Alekh and Schad, Jörg},
journal = {PVLDB},
series = {{VLDB} '10},
volume = {3},
number = {1},
pages = {518--529},
doi = {10.14778/1920841.1920908},
url = {https://doi.org/10.14778/1920841.1920908},
year = {2010}
}
Incoming Citations (Sorted by Pagerank)
Showing 39 of 39 citing papers.
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 10 of 10 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 6 | Pig Latin: A Not-So-Foreign Language for Data Processing | 2008 | SIGMOD | 0.0010686205 |
| 30 | SCOPE: Easy and Efficient Parallel Processing of Massive Data Sets | 2008 | VLDB | 0.00051174276 |
| 32 | Hive - A Warehousing Solution Over a Map-Reduce Framework | 2009 | VLDB | 0.00050111008 |
| 44 | A Comparison of Approaches to Large-Scale Data Analysis | 2009 | SIGMOD | 0.00046055057 |
| 72 | Map-Reduce-Merge: Simplified Relational Data Processing on Large Clusters | 2007 | SIGMOD | 0.00037695852 |
| 120 | HadoopDB: An Architectural Hybrid of MapReduce and DBMS Technologies for Analytical Workloads | 2009 | VLDB | 0.00031680027 |
| 155 | MAD Skills: New Analysis Practices for Big Data | 2009 | VLDB | 0.00028713176 |
| 204 | Cache Conscious Indexing for Decision-Support in Main Memory | 1999 | VLDB | 0.00025342994 |
| 642 | Building a High-Level Dataflow System on top of Map-Reduce: The Pig Experience | 2009 | VLDB | 0.00015395331 |
| 895 | Runtime Measurements in the Cloud: Observing, Analyzing, and Reducing Variance | 2010 | VLDB | 0.00013357681 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 12,131 | FP-Hadoop: Efficient Execution of Parallel Jobs Over Skewed Data | 2015 | VLDB |
| 2 | 3,749 | Integrating Hadoop and Parallel DBMS | 2010 | SIGMOD |
| 3 | 12,298 | Optimization Strategies for A/B Testing on HADOOP | 2013 | VLDB |
| 4 | 7,689 | Oracle In-Database Hadoop: When MapReduce Meets RDBMS | 2012 | SIGMOD |
| 5 | 2,159 | Efficient Processing of Data Warehousing Queries in a Split Execution Environment | 2011 | SIGMOD |
| 6 | 2,849 | Column-Oriented Storage Techniques for MapReduce | 2011 | VLDB |
| 7 | 2,265 | A Platform for Scalable One-Pass Analytics using MapReduce | 2011 | SIGMOD |
| 8 | 9,510 | Efficient Big Data Processing in Hadoop MapReduce | 2012 | VLDB |
| 9 | 120 | HadoopDB: An Architectural Hybrid of MapReduce and DBMS Technologies for Analytical Workloads | 2009 | VLDB |
| 10 | 1,436 | The Performance of MapReduce: An In-depth Study | 2010 | VLDB |