DBScholar

Back to papers

Biathlon: Harnessing Model Resilience for Accelerating ML Inference Pipelines

Summary: Biathlon, an ML-serving system, exploits model resilience to input perturbations by selecting per-aggregation-feature approximation levels to maximize latency reduction while guaranteeing bounded end-to-end accuracy loss. Evaluated on real pipelines, it achieves 5.3x–16.6x speedups with negligible accuracy drop. (summarized by gpt-5-mini on Feb 09 2026)

Paper ID
13674
Venue
VLDB
Year
2024
Pagerank
5.5330423e-05
Overall Rank
7,847 | 46.17%
DOI
10.14778/3675034.3675052

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{chang_vldb24,
        title = {{Biathlon: Harnessing Model Resilience for Accelerating ML Inference Pipelines}},
        author = {Chang, Chaokun and Lo, Eric and Ye, Chunxiao},
        journal = {PVLDB},
        series = {{VLDB} '24},
        volume = {17},
        number = {10},
        pages = {2631--2640},
        doi = {10.14778/3675034.3675052},
        url = {https://doi.org/10.14778/3675034.3675052},
        year = {2024}
}

Incoming Citations (Sorted by Pagerank)

Showing 3 of 3 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 27 of 27 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
9 Online Aggregation 1997 SIGMOD 0.00077458002
284 NoScope: Optimizing Neural Network Queries over Video at Scale 2017 VLDB 0.00022370521
323 DeepDB: Learn from Data, not from Queries! 2020 VLDB 0.00021264788
593 Wander Join: Online Aggregation via Random Walks 2016 SIGMOD 0.00016027871
772 VerdictDB: Universalizing Approximate Query Processing 2018 SIGMOD 0.00014147905
819 Quickr: Lazily Approximating Complex AdHoc Queries in BigData Clusters 2016 SIGMOD 0.00013815639
1,401 Knowing When You’re Wrong: Building Fast and Reliable Approximate Query Processing Systems 2014 SIGMOD 0.00010889902
1,799 DBEst: Revisiting Approximate Query Processing Engines with Machine Learning Models 2019 SIGMOD 9.7326398e-05
1,827 G-OLA: Generalized On-Line Aggregation for Interactive Analysis on Big Data 2015 SIGMOD 9.6690206e-05
1,962 Sample + Seek: Approximating Aggregates with Distribution Precision Guarantee 2016 SIGMOD 9.3978414e-05
1,995 Database Learning: Toward a Database that Becomes Smarter Every Time 2017 SIGMOD 9.3403665e-05
2,293 Extending Relational Query Processing with ML Inference 2020 CIDR 8.7949378e-05
2,316 Evaluating End-to-End Optimization for Data Analytics Applications in Weld 2018 VLDB 8.7596739e-05
2,633 Relational Confidence Bounds Are Easy With The Bootstrap* 2005 SIGMOD 8.3224527e-05
2,823 Query Processing on Tensor Computation Runtimes 2022 VLDB 8.0893814e-05
2,865 End-to-end Optimization of Machine Learning Prediction Queries 2022 SIGMOD 8.0180243e-05
3,157 Turbo-Charging Estimate Convergence in DBO 2009 VLDB 7.6911286e-05
3,438 A Demonstration of Willump: A Statistically-Aware End-to-end Optimizer for Machine Learning Inference 2020 VLDB 7.415647e-05
4,118 Rafiki: Machine Learning as an Analytics Service System 2019 VLDB 6.8908973e-05
4,436 Serving and Optimizing Machine Learning Workflows on Heterogeneous Infrastructures 2023 VLDB 6.7069035e-05
5,927 Incremental Computation of Common Windowed Holistic Aggregates 2016 VLDB 6.0384172e-05
6,105 Optimizing In-memory Database Engine for AI-powered On-line Decision Augmentation Using Persistent Memory 2021 VLDB 5.9743349e-05
6,553 Containerized Execution of UDFs: An Experimental Evaluation 2022 VLDB 5.8417517e-05
6,585 JoinBoost: Grow Trees Over Normalized Data Using Only SQL 2023 VLDB 5.8350362e-05
9,433 FEBench: A Benchmark for Real-Time Relational Data Feature Extraction 2023 VLDB 5.2696166e-05
9,927 Everest: A Top-K Deep Video Analytics System 2022 SIGMOD 5.1955087e-05
9,940 RALF: Accuracy-Aware Scheduling for Feature Store Maintenance 2024 VLDB 5.1924403e-05
Previous Page 1 / 1 Next

Semantically Similar Papers