DBScholar

Back to papers

Tuple-oriented Compression for Large-scale Mini-batch Stochastic Gradient Descent

Summary: Proposes tuple-oriented compression (TOC) for mini-batch SGD, preserving tuple boundaries while using LZW-inspired coding. Enables compressed-domain matrix operations on TOC, delivering up to 51x compression and 10.2x speedups for MGD workloads with no decompression overhead. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
5663
Venue
SIGMOD
Year
2019
Pagerank
5.8657457e-05
Overall Rank
6,485 | 55.51%
DOI
10.1145/3299869.3300070

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{li_sigmod19,
        title = {{Tuple-oriented Compression for Large-scale Mini-batch Stochastic Gradient Descent}},
        author = {Li, Fengan and Chen, Lingjiao and Zeng, Yijing and Kumar, Arun and Wu, Xi and Naughton, Jeffrey F. and Patel, Jignesh M.},
        series = {{SIGMOD} '19},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/3299869.3300070},
        url = {https://dl.acm.org/doi/10.1145/3299869.3300070},
        year = {2019}
}

Incoming Citations (Sorted by Pagerank)

Showing 8 of 8 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 18 of 18 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
60 Integrating Compression and Execution in Column-Oriented Database Systems 2006 SIGMOD 0.0003955489
106 The MADlib Analytics Library or MAD Skills, the SQL 2012 VLDB 0.00033539462
165 DB2 with BLU Acceleration: So Much More than Just a Column Store 2013 VLDB 0.00027693424
518 Towards a Unified Architecture for in-RDBMS Analytics 2012 SIGMOD 0.00017167492
532 MLbase: A Distributed Machine-learning System 2013 CIDR 0.00017072641
536 Learning Linear Regression Models over Factorized Joins 2016 SIGMOD 0.0001693369
715 Learning Generalized Linear Models Over Normalized Data 2015 SIGMOD 0.00014655327
764 To Join or Not to Join? Thinking Twice about Joins before Feature Selection 2016 SIGMOD 0.00014226652
870 BitWeaving: Fast Scans for Main Memory Data Processing 2013 SIGMOD 0.0001350293
993 Simulation of Database-Valued Markov Chains Using SimSQL 2013 SIGMOD 0.00012789598
1,079 Hybrid Parallelization Strategies for Large-Scale Machine Learning in SystemML 2014 VLDB 0.00012258469
1,235 Towards Linear Algebra over Normalized Data 2017 VLDB 0.00011548457
1,644 Compressed Linear Algebra for Large-Scale Machine Learning 2016 VLDB 0.00010132912
2,224 An Experimental Study of Bitmap Compression vs. Inverted List Compression 2017 SIGMOD 8.9183396e-05
4,185 Bolt-on Differential Privacy for Scalable Stochastic Gradient Descent-based Analytics 2017 SIGMOD 6.8462927e-05
4,799 Scalable Asynchronous Gradient Descent Optimization for Out-of-Core Models 2017 VLDB 6.5024714e-05
6,918 Leveraging Compression in the Tableau Data Engine 2014 SIGMOD 5.7392935e-05
7,000 A Cost-based Optimizer for Gradient Descent Optimization 2017 SIGMOD 5.7287645e-05
Previous Page 1 / 1 Next

Semantically Similar Papers