DBScholar

Back to papers

Compressed Linear Algebra for Large-Scale Machine Learning

Summary: CLA applies database-style compression to matrices and executes matrix-vector products directly on compressed data. Novel column compression, cache-conscious operators, and a sampling-based scheme yield near-uncompressed performance with substantial memory savings. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
hbb76a95d2a89022f
Venue
VLDB
Year
2016
Pagerank
0.00010071891
Overall Rank
1,614 | 89.15%
DOI
10.14778/2994509.2994510

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{elgohary_vldb16,
        title = {{Compressed Linear Algebra for Large-Scale Machine Learning}},
        author = {Elgohary, Ahmed and Boehm, Matthias and Haas, Peter J. and Reiss, Frederick R. and Reinwald, Berthold},
        journal = {PVLDB},
        series = {{VLDB} '16},
        volume = {9},
        number = {12},
        pages = {960--971},
        doi = {10.14778/2994509.2994510},
        url = {https://doi.org/10.14778/2994509.2994510},
        year = {2016}
}

Incoming Citations (Sorted by Pagerank)

Showing 31 of 31 citing papers.

Rank Citing Paper Year Venue Pagerank
1,255 Data Management in Machine Learning: Challenges, Techniques, and Systems 2017 SIGMOD 0.00011325762
1,256 Towards Linear Algebra over Normalized Data 2017 VLDB 0.00011314687
1,668 SystemDS: A Declarative Machine Learning System for the End-to-End Data Science Lifecycle 2020 CIDR 9.9371612e-05
1,691 MISTIQUE: A System to Store and Query Model Intermediates for Model Diagnosis 2018 SIGMOD 9.8570722e-05
2,199 Enabling and Optimizing Non-linear Feature Interactions in Factorized Linear Algebra 2019 SIGMOD 8.8750296e-05
2,264 An Intermediate Representation for Optimizing Machine Learning Pipelines 2019 VLDB 8.7289107e-05
2,284 Pump Up the Volume: Processing Large Data on GPUs with Fast Interconnects 2020 SIGMOD 8.6954168e-05
2,326 SliceLine: Fast, Linear-Algebra-based Slice Finding for ML Model Debugging 2021 SIGMOD 8.6309237e-05
3,101 On Optimizing Operator Fusion Plans for Large-Scale Machine Learning in SystemML 2018 VLDB 7.649219e-05
3,112 Incremental View Maintenance with Triple Lock Factorization Benefits 2018 SIGMOD 7.6357579e-05
3,240 Towards Demystifying Serverless Machine Learning Training 2021 SIGMOD 7.5002772e-05
3,330 SPOOF: Sum-Product Optimization and Operator Fusion for Large-Scale Machine Learning 2017 CIDR 7.4173693e-05
4,095 Distributed Deep Learning on Data Systems: A Comparative Analysis of Approaches 2021 VLDB 6.8095767e-05
4,330 MNC: Structure-Exploiting Sparsity Estimation for Matrix Expressions 2019 SIGMOD 6.6595681e-05
4,573 Accelerating Raw Data Analysis with the ACCORDA Software and Hardware Architecture 2019 VLDB 6.5265033e-05
4,879 Accelerating Generalized Linear Models with MLWeaving: A One-Size-Fits-All System for Any-Precision Learning 2019 VLDB 6.3715334e-05
5,877 BlinkML: Efficient Maximum Likelihood Estimation with Probabilistic Guarantees 2019 SIGMOD 5.9627218e-05
5,933 Doing More with Less: Characterizing Dataset Downsampling for AutoML 2021 VLDB 5.942188e-05
6,169 Automatic Optimization of Matrix Implementations for Distributed Machine Learning and Linear Algebra 2021 SIGMOD 5.8624857e-05
6,592 Tuple-oriented Compression for Large-scale Mini-batch Stochastic Gradient Descent 2019 SIGMOD 5.7421684e-05
6,662 UPLIFT: Parallelization Strategies for Feature Transformations in Machine Learning Workloads 2022 VLDB 5.7171651e-05
7,839 ExDRa: Exploratory Data Science on Federated Raw Data 2021 SIGMOD 5.4432099e-05
8,384 AWARE: Workload-aware, Redundancy-exploiting Linear Algebra 2023 SIGMOD 5.3421754e-05
8,477 TreeSensing: Linearly Compressing Sketches with Flexibility 2023 SIGMOD 5.3332727e-05
8,773 Improving Matrix-vector Multiplication via Lossless Grammar-Compressed Matrices 2022 VLDB 5.2804275e-05
9,748 PlinyCompute: A Platform for High-Performance, Distributed, Data-Intensive Tool Development 2018 SIGMOD 5.1349531e-05
9,822 BlockJoin: Efficient Matrix Partitioning Through Joins 2017 VLDB 5.1254832e-05
10,146 QStore: Quantization-Aware Compressed Model Storage 2026 VLDB 5.0715586e-05
10,935 OmniTable: A Unified Wide-Table System for Petabyte-Scale LLM Data Curation and Exploration 2026 VLDB 4.9793485e-05
10,945 Morphing-based Compression for Data-centric ML Pipelines 2026 VLDB 4.9793485e-05
11,107 HyperMR: Efficient Hypergraph-enhanced Matrix Storage on Compute-in-Memory Architecture 2025 SIGMOD 4.9793485e-05
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 16 of 16 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Previous Page 1 / 1 Next

Semantically Similar Papers