DBScholar

Back to papers

Compressed Linear Algebra for Large-Scale Machine Learning

Summary: CLA applies database-style compression to matrices and executes matrix-vector products directly on compressed data. Novel column compression, cache-conscious operators, and a sampling-based scheme yield near-uncompressed performance with substantial memory savings. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
11571
Venue
VLDB
Year
2016
Pagerank
0.00010132912
Overall Rank
1,644 | 88.73%
DOI
10.14778/2994509.2994510

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{elgohary_vldb16,
        title = {{Compressed Linear Algebra for Large-Scale Machine Learning}},
        author = {Elgohary, Ahmed and Boehm, Matthias and Haas, Peter J. and Reiss, Frederick R. and Reinwald, Berthold},
        journal = {PVLDB},
        series = {{VLDB} '16},
        volume = {9},
        number = {12},
        pages = {960--971},
        doi = {10.14778/2994509.2994510},
        url = {https://doi.org/10.14778/2994509.2994510},
        year = {2016}
}

Incoming Citations (Sorted by Pagerank)

Showing 30 of 30 citing papers.

Rank Citing Paper Year Venue Pagerank
1,235 Towards Linear Algebra over Normalized Data 2017 VLDB 0.00011548457
1,250 Data Management in Machine Learning: Challenges, Techniques, and Systems 2017 SIGMOD 0.00011485301
1,670 MISTIQUE: A System to Store and Query Model Intermediates for Model Diagnosis 2018 SIGMOD 0.00010045615
1,756 SystemDS: A Declarative Machine Learning System for the End-to-End Data Science Lifecycle 2020 CIDR 9.8172465e-05
2,179 Enabling and Optimizing Non-linear Feature Interactions in Factorized Linear Algebra 2019 SIGMOD 9.0146333e-05
2,239 An Intermediate Representation for Optimizing Machine Learning Pipelines 2019 VLDB 8.8875753e-05
2,273 SliceLine: Fast, Linear-Algebra-based Slice Finding for ML Model Debugging 2021 SIGMOD 8.8230899e-05
2,566 Pump Up the Volume: Processing Large Data on GPUs with Fast Interconnects 2020 SIGMOD 8.4116562e-05
3,169 Towards Demystifying Serverless Machine Learning Training 2021 SIGMOD 7.6715222e-05
3,205 On Optimizing Operator Fusion Plans for Large-Scale Machine Learning in SystemML 2018 VLDB 7.6386536e-05
3,206 Incremental View Maintenance with Triple Lock Factorization Benefits 2018 SIGMOD 7.6367549e-05
3,284 SPOOF: Sum-Product Optimization and Operator Fusion for Large-Scale Machine Learning 2017 CIDR 7.5663058e-05
4,067 Distributed Deep Learning on Data Systems: A Comparative Analysis of Approaches 2021 VLDB 6.9293511e-05
4,409 MNC: Structure-Exploiting Sparsity Estimation for Matrix Expressions 2019 SIGMOD 6.7178579e-05
4,483 Accelerating Raw Data Analysis with the ACCORDA Software and Hardware Architecture 2019 VLDB 6.6724044e-05
4,798 Accelerating Generalized Linear Models with MLWeaving: A One-Size-Fits-All System for Any-Precision Learning 2019 VLDB 6.5024772e-05
5,785 BlinkML: Efficient Maximum Likelihood Estimation with Probabilistic Guarantees 2019 SIGMOD 6.0892672e-05
5,839 Doing More with Less: Characterizing Dataset Downsampling for AutoML 2021 VLDB 6.0699037e-05
6,046 Automatic Optimization of Matrix Implementations for Distributed Machine Learning and Linear Algebra 2021 SIGMOD 5.9956597e-05
6,485 Tuple-oriented Compression for Large-scale Mini-batch Stochastic Gradient Descent 2019 SIGMOD 5.8657457e-05
6,538 UPLIFT: Parallelization Strategies for Feature Transformations in Machine Learning Workloads 2022 VLDB 5.8477764e-05
7,687 ExDRa: Exploratory Data Science on Federated Raw Data 2021 SIGMOD 5.5671645e-05
8,310 TreeSensing: Linearly Compressing Sketches with Flexibility 2023 SIGMOD 5.4556836e-05
8,617 Improving Matrix-vector Multiplication via Lossless Grammar-Compressed Matrices 2022 VLDB 5.4002779e-05
8,794 AWARE: Workload-aware, Redundancy-exploiting Linear Algebra 2023 SIGMOD 5.370464e-05
9,572 PlinyCompute: A Platform for High-Performance, Distributed, Data-Intensive Tool Development 2018 SIGMOD 5.2528121e-05
9,675 BlockJoin: Efficient Matrix Partitioning Through Joins 2017 VLDB 5.2380072e-05
10,584 QStore: Quantization-Aware Compressed Model Storage 2026 VLDB 5.093636e-05
10,589 Morphing-based Compression for Data-centric ML Pipelines 2026 VLDB 5.093636e-05
10,666 HyperMR: Efficient Hypergraph-enhanced Matrix Storage on Compute-in-Memory Architecture 2025 SIGMOD 5.093636e-05
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 16 of 16 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Previous Page 1 / 1 Next

Semantically Similar Papers