Back to papers
Cerebro: A Data System for Optimized Deep Learning Model Selection
Summary: Cerebro is a data system for optimized deep-learning model selection, boosting throughput and reproducibility at lower cost. Model hopper parallelism, a hybrid task/data-parallel SGD, yields 3–10x speedups and memory/network savings across varied resources.
(summarized by gpt-5-nano on Feb 09 2026)
- Paper ID
- 12107
- Venue
- VLDB
- Year
- 2020
- Pagerank
- 0.00018152321
- Overall Rank
- 684 | 95.25%
- DOI
-
10.14778/3407790.3407816
Incoming Non-self Citations Over Time
Incoming Citations (Sorted by Pagerank)
Showing 26 of 26 citing papers.
| Rank |
Citing Paper |
Year |
Venue |
Pagerank |
| 1,500 |
Analyzing and Mitigating Data Stalls in DNN Training |
2021 |
VLDB |
0.00011636174 |
| 1,942 |
SliceLine: Fast, Linear-Algebra-based Slice Finding for ML Model Debugging |
2021 |
SIGMOD |
0.00010010569 |
| 2,845 |
VolcanoML: Speeding up End-to-End AutoML via Scalable Search Space Decomposition |
2021 |
VLDB |
8.0301674e-05 |
| 4,601 |
Distributed Deep Learning on Data Systems: A Comparative Analysis of Approaches |
2021 |
VLDB |
6.05274e-05 |
| 4,962 |
Doing More with Less: Characterizing Dataset Downsampling for AutoML |
2021 |
VLDB |
5.7979872e-05 |
| 6,361 |
Galvatron: Efficient Transformer Training over Multiple GPUs Using Automatic Parallelism |
2023 |
VLDB |
5.0903244e-05 |
| 6,573 |
In-Database Machine Learning with CorgiPile: Stochastic Gradient Descent without Full Data Shuffle |
2022 |
SIGMOD |
5.0009513e-05 |
| 6,885 |
Lotan: Bridging the Gap between GNNs and Scalable Graph Analytics Engines |
2023 |
VLDB |
4.8908367e-05 |
| 7,656 |
Nautilus: An Optimized System for Deep Transfer Learning over Evolving Training Datasets |
2022 |
SIGMOD |
4.6826896e-05 |
| 8,096 |
Saga: A Scalable Framework for Optimizing Data Cleaning Pipelines for Machine Learning Applications |
2023 |
SIGMOD |
4.583522e-05 |
| 8,183 |
SHiFT: An Efficient, Flexible Search Engine for Transfer Learning |
2023 |
VLDB |
4.5615358e-05 |
| 8,515 |
UPLIFT: Parallelization Strategies for Feature Transformations in Machine Learning Workloads |
2022 |
VLDB |
4.4901466e-05 |
| 8,731 |
TensorSocket: Shared Data Loading for Deep Learning Training |
2026 |
SIGMOD |
4.4520434e-05 |
| 8,864 |
Cerebro: A Layered Data Platform for Scalable Deep Learning |
2021 |
CIDR |
4.4283952e-05 |
| 9,196 |
Hyper-Tune: Towards Efficient Hyper-parameter Tuning at Scale |
2022 |
VLDB |
4.3723457e-05 |
| 9,225 |
Towards an Optimized GROUP BY Abstraction for Large-Scale Machine Learning |
2021 |
VLDB |
4.3656789e-05 |
| 9,226 |
Intermittent Human-in-the-Loop Model Selection using Cerebro: A Demonstration |
2021 |
VLDB |
4.3656789e-05 |
| 9,271 |
COMET: A Novel Memory-Efficient Deep Learning Training Framework by Using Error-Bounded Lossy Compression |
2022 |
VLDB |
4.3625977e-05 |
| 9,787 |
The Image Calculator: 10x Faster Image-AI Inference by Replacing JPEG with Self-designing Storage Format |
2024 |
SIGMOD |
4.2799988e-05 |
| 10,846 |
ML-Asset Management: Curation, Discovery, and Utilization |
2025 |
VLDB |
4.1905499e-05 |
| 11,001 |
Database Native Model Selection: Harnessing Deep Neural Networks in Database Systems |
2024 |
VLDB |
4.1905499e-05 |
| 11,341 |
Redundancy Elimination in Distributed Matrix Computation |
2022 |
SIGMOD |
4.1905499e-05 |
| 11,434 |
Ease.ML: A Lifecycle Management System for MLDev and MLOps |
2021 |
CIDR |
4.1905499e-05 |
| 11,450 |
Grouped Learning: Group-By Model Selection Workloads |
2021 |
SIGMOD |
4.1905499e-05 |
| 13,185 |
Reimagining Deep Learning Systems Through the Lens of Data Systems |
2024 |
VLDB |
- |
| 13,284 |
Errata for “Cerebro: A Data System for Optimized Deep Learning Model Selection” |
2021 |
VLDB |
- |
Outgoing Citations (Sorted by Pagerank)
Showing 14 of 14 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank |
Cited Paper |
Year |
Venue |
Pagerank |
| 139 |
The MADlib Analytics Library or MAD Skills, the SQL |
2012 |
VLDB |
0.00042320525 |
| 638 |
Towards a Unified Architecture for in-RDBMS Analytics |
2012 |
SIGMOD |
0.00018810785 |
| 1,407 |
Hybrid Parallelization Strategies for Large-Scale Machine Learning in SystemML |
2014 |
VLDB |
0.00012163413 |
| 1,534 |
Data Management in Machine Learning: Challenges, Techniques, and Systems |
2017 |
SIGMOD |
0.00011462072 |
| 1,946 |
Heterogeneity-aware Distributed Parameter Servers |
2017 |
SIGMOD |
9.9983926e-05 |
| 2,446 |
FlexPS: Flexible Parallelism Control in Parameter Server Architecture |
2018 |
VLDB |
8.7988018e-05 |
| 2,871 |
Incremental and Approximate Inference for Faster Occlusion-based Deep CNN Explanations |
2019 |
SIGMOD |
7.9800983e-05 |
| 2,892 |
VISTA: Optimized System for Declarative Feature Transfer from Deep CNNs at Scale |
2020 |
SIGMOD |
7.9570135e-05 |
| 3,372 |
CROSSBOW: Scaling Deep Learning with Small Batch Sizes on Multi-GPU Servers |
2019 |
VLDB |
7.1633912e-05 |
| 3,953 |
A Comparative Evaluation of Systems for Scalable Linear Algebra-based Analytics |
2018 |
VLDB |
6.5896733e-05 |
| 4,262 |
Mariana: Tencent Deep Learning Platform and its Applications |
2014 |
VLDB |
6.3019728e-05 |
| 6,535 |
Tuple-oriented Compression for Large-scale Mini-batch Stochastic Gradient Descent |
2019 |
SIGMOD |
5.0184189e-05 |
| 13,284 |
Errata for “Cerebro: A Data System for Optimized Deep Learning Model Selection” |
2021 |
VLDB |
- |
| 13,326 |
Demonstration of Krypton: Optimized CNN Inference for Occlusion-based Deep CNN Explanations |
2019 |
VLDB |
- |
Semantically Similar Papers
| Overall Rank |
Paper |
Year |
Venue |
Pagerank |
| 9,603 |
Saturn: An Optimized Data System for Multi-Large-Model Deep Learning Workloads |
2024 |
VLDB |
4.3136057e-05 |
| 9,174 |
MemFlow: Memory-Aware Distributed Deep Learning |
2020 |
SIGMOD |
4.3807157e-05 |
| 8,731 |
TensorSocket: Shared Data Loading for Deep Learning Training |
2026 |
SIGMOD |
4.4520434e-05 |
| 13,284 |
Errata for “Cerebro: A Data System for Optimized Deep Learning Model Selection” |
2021 |
VLDB |
- |
| 9,225 |
Towards an Optimized GROUP BY Abstraction for Large-Scale Machine Learning |
2021 |
VLDB |
4.3656789e-05 |
| 3,372 |
CROSSBOW: Scaling Deep Learning with Small Batch Sizes on Multi-GPU Servers |
2019 |
VLDB |
7.1633912e-05 |
| 9,269 |
Model-Parallel Model Selection for Deep Learning Systems |
2021 |
SIGMOD |
4.3633561e-05 |
| 9,226 |
Intermittent Human-in-the-Loop Model Selection using Cerebro: A Demonstration |
2021 |
VLDB |
4.3656789e-05 |
| 4,601 |
Distributed Deep Learning on Data Systems: A Comparative Analysis of Approaches |
2021 |
VLDB |
6.05274e-05 |
| 8,864 |
Cerebro: A Layered Data Platform for Scalable Deep Learning |
2021 |
CIDR |
4.4283952e-05 |