Optimistic Recovery for Iterative Dataflows in Action
Summary: Optimistic recovery for iterative dataflows eliminates intermediate state checkpoints via compensation functions to reach a consistent state after failure. Demonstrated on Apache Flink with graph algorithms, it provides fault tolerance without checkpoint overhead and near-optimal failure-free performance. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Sergey Dudoladov (Technical University of Berlin)
- 2. Chen Xu (Technical University of Berlin)
- 3. Sebastian Schelter (Technical University of Berlin)
- 4. Asterios Katsifodimos (Technical University of Berlin)
- 5. Stephan Ewen (Data Artisans GmbH)
- 6. Kostas Tzoumas (Data Artisans GmbH)
- 7. Volker Markl (Technical University of Berlin)
BibTeX Citation
@inproceedings{dudoladov_sigmod15,
title = {{Optimistic Recovery for Iterative Dataflows in Action}},
author = {Dudoladov, Sergey and Xu, Chen and Schelter, Sebastian and Katsifodimos, Asterios and Ewen, Stephan and Tzoumas, Kostas and Markl, Volker},
series = {{SIGMOD} '15},
booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
publisher = {Association for Computing Machinery},
doi = {10.1145/2723372.2735372},
url = {https://dl.acm.org/doi/10.1145/2723372.2735372},
year = {2015}
}
Incoming Citations (Sorted by Pagerank)
Showing 2 of 2 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 5,388 | Fine-Grained Modeling and Optimization for Intelligent Resource Management in Big Data Processing | 2022 | VLDB | 6.2362811e-05 |
| 8,615 | A Spark Optimizer for Adaptive, Fine-Grained Parameter Tuning | 2024 | VLDB | 5.4005602e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 7 of 7 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 3 | Pregel: A System for Large-Scale Graph Processing | 2010 | SIGMOD | 0.0012250108 |
| 20 | Distributed GraphLab: A Framework for Machine Learning and Data Mining in the Cloud | 2012 | VLDB | 0.00056944564 |
| 30 | SCOPE: Easy and Efficient Parallel Processing of Massive Data Sets | 2008 | VLDB | 0.00051174276 |
| 372 | HaLoop: Efficient Iterative Data Processing on Large Clusters | 2010 | VLDB | 0.0001981521 |
| 1,054 | Interactive Analytical Processing in Big Data Systems: A Cross-Industry Study of MapReduce Workloads | 2012 | VLDB | 0.00012390673 |
| 2,196 | Spinning Fast Iterative Data Flows | 2012 | VLDB | 8.9704984e-05 |
| 2,622 | A Latency and Fault-Tolerance Optimizer for Online Parallel Query Plans | 2011 | SIGMOD | 8.3330136e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 4,078 | Asynchronous and Fault-Tolerant Recursive Datalog Evaluation in Shared-Nothing Engines | 2015 | VLDB |
| 2 | 3,776 | Fault-tolerant Stream Processing using a Distributed, Replicated File System | 2008 | VLDB |
| 3 | 1,789 | Fault-Tolerance in the Borealis Distributed Stream Processing System | 2005 | SIGMOD |
| 4 | 983 | Integrating Scale Out and Fault Tolerance in Stream Processing using Operator State Management | 2013 | SIGMOD |
| 5 | 1,209 | Highly Available, Fault-Tolerant, Parallel Dataflows | 2004 | SIGMOD |
| 6 | 1,422 | State Management in Apache Flink: Consistent Stateful Distributed Stream Processing | 2017 | VLDB |
| 7 | 9,596 | Cost-based Fault-tolerance for Parallel Data Processing | 2015 | SIGMOD |
| 8 | 7,218 | Fast Failure Recovery in Distributed Graph Processing Systems | 2015 | VLDB |
| 9 | 2,196 | Spinning Fast Iterative Data Flows | 2012 | VLDB |
| 10 | 13,523 | Fault-Tolerance for Distributed Iterative Dataflows in Action | 2018 | VLDB |