Early Accurate Results for Advanced Analytics on MapReduce
Summary: EARL extends Hadoop with iterative sampling and bootstrap-based online error estimates, enabling early approximate results for arbitrary MapReduce workflows. It adaptively expands samples to meet user-defined accuracy, reducing time/I/O and aiding fault-tolerance. (summarized by gpt-5.6-luna on Jul 24 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Nikolay Laptev (University of California Los Angeles)
- 2. Kai Zeng (University of California Los Angeles)
- 3. Carlo Zaniolo (University of California Los Angeles)
BibTeX Citation
@article{laptev_vldb12,
title = {{Early Accurate Results for Advanced Analytics on MapReduce}},
author = {Laptev, Nikolay and Zeng, Kai and Zaniolo, Carlo},
journal = {PVLDB},
series = {{VLDB} '12},
volume = {5},
number = {10},
pages = {1028--1039},
doi = {10.14778/2336664.2336674},
url = {https://doi.org/10.14778/2336664.2336674},
year = {2012}
}
Incoming Citations (Sorted by Pagerank)
Showing 12 of 12 citing papers.
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 10 of 10 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 6 | Pig Latin: A Not-So-Foreign Language for Data Processing | 2008 | SIGMOD | 0.0010686205 |
| 9 | Online Aggregation | 1997 | SIGMOD | 0.00077458002 |
| 32 | Hive - A Warehousing Solution Over a Map-Reduce Framework | 2009 | VLDB | 0.00050111008 |
| 372 | HaLoop: Efficient Iterative Data Processing on Large Clusters | 2010 | VLDB | 0.0001981521 |
| 923 | Starfish: A Self-tuning System for Big Data Analytics | 2011 | CIDR | 0.00013189886 |
| 1,009 | Online Aggregation for Large MapReduce Jobs | 2011 | VLDB | 0.00012684342 |
| 1,806 | Effective Use of Block-Level Sampling in Statistics Estimation | 2004 | SIGMOD | 9.7112151e-05 |
| 2,265 | A Platform for Scalable One-Pass Analytics using MapReduce | 2011 | SIGMOD | 8.8398946e-05 |
| 2,312 | Online Aggregation and Continuous Query support in MapReduce | 2010 | SIGMOD | 8.7642158e-05 |
| 2,633 | Relational Confidence Bounds Are Easy With The Bootstrap* | 2005 | SIGMOD | 8.3224527e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 2,312 | Online Aggregation and Continuous Query support in MapReduce | 2010 | SIGMOD |
| 2 | 1,009 | Online Aggregation for Large MapReduce Jobs | 2011 | VLDB |
| 3 | 9,510 | Efficient Big Data Processing in Hadoop MapReduce | 2012 | VLDB |
| 4 | 120 | HadoopDB: An Architectural Hybrid of MapReduce and DBMS Technologies for Analytical Workloads | 2009 | VLDB |
| 5 | 12,131 | FP-Hadoop: Efficient Execution of Parallel Jobs Over Skewed Data | 2015 | VLDB |
| 6 | 9,640 | Supporting Scalable Analytics with Latency Constraints | 2015 | VLDB |
| 7 | 660 | Hadoop++: Making a Yellow Elephant Run Like a Cheetah (Without It Even Noticing) | 2010 | VLDB |
| 8 | 5,522 | Building Wavelet Histograms on Large Data in MapReduce | 2012 | VLDB |
| 9 | 1,436 | The Performance of MapReduce: An In-depth Study | 2010 | VLDB |
| 10 | 2,265 | A Platform for Scalable One-Pass Analytics using MapReduce | 2011 | SIGMOD |