DBScholar

Back to papers

Tuple-oriented Compression for Large-scale Mini-batch Stochastic Gradient Descent

Summary: Proposes tuple-oriented compression (TOC) for mini-batch SGD, preserving tuple boundaries while using LZW-inspired coding. Enables compressed-domain matrix operations on TOC, delivering up to 51x compression and 10.2x speedups for MGD workloads with no decompression overhead. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
hc04c395b26d7c401
Venue
SIGMOD
Year
2019
Pagerank
5.7421684e-05
Overall Rank
6,592 | 55.68%
DOI
10.1145/3299869.3300070

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{li_sigmod19,
        title = {{Tuple-oriented Compression for Large-scale Mini-batch Stochastic Gradient Descent}},
        author = {Li, Fengan and Chen, Lingjiao and Zeng, Yijing and Kumar, Arun and Wu, Xi and Naughton, Jeffrey F. and Patel, Jignesh M.},
        series = {{SIGMOD} '19},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/3299869.3300070},
        url = {https://dl.acm.org/doi/10.1145/3299869.3300070},
        year = {2019}
}

Incoming Citations (Sorted by Pagerank)

Showing 8 of 8 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 18 of 18 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
61 Integrating Compression and Execution in Column-Oriented Database Systems 2006 SIGMOD 0.000392237
105 The MADlib Analytics Library or MAD Skills, the SQL 2012 VLDB 0.00033638251
163 DB2 with BLU Acceleration: So Much More than Just a Column Store 2013 VLDB 0.0002749118
503 Towards a Unified Architecture for in-RDBMS Analytics 2012 SIGMOD 0.00017202276
521 Learning Linear Regression Models over Factorized Joins 2016 SIGMOD 0.00016929744
537 MLbase: A Distributed Machine-learning System 2013 CIDR 0.00016768109
730 Learning Generalized Linear Models Over Normalized Data 2015 SIGMOD 0.00014406936
777 To Join or Not to Join? Thinking Twice about Joins before Feature Selection 2016 SIGMOD 0.00014054709
873 BitWeaving: Fast Scans for Main Memory Data Processing 2013 SIGMOD 0.00013338838
1,009 Simulation of Database-Valued Markov Chains Using SimSQL 2013 SIGMOD 0.00012539827
1,081 Hybrid Parallelization Strategies for Large-Scale Machine Learning in SystemML 2014 VLDB 0.00012123917
1,256 Towards Linear Algebra over Normalized Data 2017 VLDB 0.00011314687
1,614 Compressed Linear Algebra for Large-Scale Machine Learning 2016 VLDB 0.00010071891
2,201 An Experimental Study of Bitmap Compression vs. Inverted List Compression 2017 SIGMOD 8.8726671e-05
3,711 Bolt-on Differential Privacy for Scalable Stochastic Gradient Descent-based Analytics 2017 SIGMOD 7.0783969e-05
4,901 Scalable Asynchronous Gradient Descent Optimization for Out-of-Core Models 2017 VLDB 6.3647386e-05
6,984 Leveraging Compression in the Tableau Data Engine 2014 SIGMOD 5.6285605e-05
7,142 A Cost-based Optimizer for Gradient Descent Optimization 2017 SIGMOD 5.600563e-05
Previous Page 1 / 1 Next

Semantically Similar Papers