DBScholar

Back to papers

BladeDISC: Optimizing Dynamic Shape Machine Learning Workloads via Compiler Approach

Summary: BladeDISC is a compiler-driven optimizer for dynamic-shape ML workloads, tackling fusion and codegen with unknown shapes. Key ideas: symbolic shape representation and shape-information propagation to enable shape-agnostic fusion and generic codegen. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
h160c2f26e908fb5e
Venue
SIGMOD
Year
2023
Pagerank
5.1453267e-05
Overall Rank
9,656 | 35.08%
DOI
10.1145/3617327

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{zheng_sigmod23,
        title = {{BladeDISC: Optimizing Dynamic Shape Machine Learning Workloads via Compiler Approach}},
        author = {Zheng, Zhen and Pan, Zaifeng and Wang, Dalin and Zhu, Kai and Zhao, Wenyi and Guo, Tianyou and Qiu, Xiafei and Sun, Minmin and Bai, Junjie and Zhang, Feng and Du, Xiaoyong and Zhai, Jidong and Lin, Wei},
        series = {{SIGMOD} '23},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/3617327},
        url = {https://dl.acm.org/doi/10.1145/3617327},
        year = {2023}
}

Incoming Citations (Sorted by Pagerank)

Showing 1 of 1 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 30 of 30 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
21 Efficiently Compiling Efficient Query Plans for Modern Hardware 2011 VLDB 0.00056855599
53 Eddies: Continuously Adaptive Query Processing 2000 SIGMOD 0.00040860054
481 Robust Query Processing through Progressive Optimization 2004 SIGMOD 0.00017603972
522 PyTorch Distributed: Experiences on Accelerating Data Parallel Training 2020 VLDB 0.00016912378
1,081 Hybrid Parallelization Strategies for Large-Scale Machine Learning in SystemML 2014 VLDB 0.00012123917
1,199 Balsa: Learning a Query Optimizer Without Expert Demonstrations 2022 SIGMOD 0.00011563985
1,444 An Architecture for Compiling UDF-centric Workflows 2015 VLDB 0.0001063181
1,533 Pipelined Query Processing in Coprocessor Environments 2018 SIGMOD 0.00010332035
1,541 HippogriffDB: Balancing I/O and GPU Bandwidth in Big Data Analytics 2016 VLDB 0.00010313459
1,803 Tuplex: Data Science in Python at Native Code Speed 2021 SIGMOD 9.6068397e-05
2,249 Evaluating End-to-End Optimization for Data Analytics Applications in Weld 2018 VLDB 8.7549752e-05
2,279 NeutronStar: Distributed GNN Training with Hybrid Dependency Management 2022 SIGMOD 8.7062637e-05
2,682 Accelerating Recommendation System Training by Leveraging Popular Choices 2022 VLDB 8.1423493e-05
3,101 On Optimizing Operator Fusion Plans for Large-Scale Machine Learning in SystemML 2018 VLDB 7.649219e-05
3,972 TCUDB: Accelerating Database with Tensor Processors 2022 SIGMOD 6.8881464e-05
3,985 Designing an Open Framework for Query Optimization and Compilation 2022 VLDB 6.8730085e-05
4,116 Resource Elasticity for Large-Scale Machine Learning 2015 SIGMOD 6.7961306e-05
4,645 Improving Execution Efficiency of Just-in-time Compilation based Query Processing on GPUs 2021 VLDB 6.4881339e-05
4,945 HET-GMP: A Graph-based System Approach to Scaling Large Embedding Model Training 2022 SIGMOD 6.3435891e-05
4,963 Heterogeneity-Aware Distributed Machine Learning Training via Partial Reduce 2021 SIGMOD 6.3384091e-05
5,009 Galvatron: Efficient Transformer Training over Multiple GPUs Using Automatic Parallelism 2023 VLDB 6.3158866e-05
5,633 NuPS: A Parameter Server for Machine Learning with Non-Uniform Parameter Access 2022 SIGMOD 6.0578661e-05
6,441 Grizzly: Efficient Stream Processing Through Adaptive Query Compilation 2020 SIGMOD 5.7834762e-05
7,789 Measuring and Optimizing Distributed Array Programs 2016 VLDB 5.4521754e-05
7,962 Excalibur: A Virtual Machine for Adaptive Fine-grained JIT-Compiled Query Execution based on VOILA 2023 VLDB 5.4167003e-05
8,521 FuseME: Distributed Matrix Computation Engine based on Cuboid-based Fused Operator and Plan Generation 2022 SIGMOD 5.3238144e-05
8,555 Harmony: Overcoming the Hurdles of GPU Memory Capacity to Train Massive DNN Models on Commodity Servers 2022 VLDB 5.3152366e-05
9,162 Optimizing Inference Serving on Serverless Platforms 2022 VLDB 5.2153488e-05
9,457 COMET: A Novel Memory-Efficient Deep Learning Training Framework by Using Error-Bounded Lossy Compression 2022 VLDB 5.1735029e-05
10,013 ETO: Accelerating Optimization of DNN Operators by High-Performance Tensor Program Reuse 2022 VLDB 5.0971861e-05
Previous Page 1 / 1 Next

Semantically Similar Papers