DBScholar

Back to papers

CAFE: Towards Compact, Adaptive, and Fast Embedding for Large-scale Recommendation Models

Summary: CAFE enables compact, adaptive embedding for large-scale DLRMs; HotSketch identifies hot features and assigns them dedicated embeddings, while non-hot features share via multi-level hash. Theoretical accuracy/convergence analysis; 3.92% and 3.68% AUC gains on Criteo Kaggle and CriteoTB at 10k× compression. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
6922
Venue
SIGMOD
Year
2024
Pagerank
5.2528121e-05
Overall Rank
9,556 | 34.44%
DOI
10.1145/3639306

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{zhang_sigmod24,
        title = {{CAFE: Towards Compact, Adaptive, and Fast Embedding for Large-scale Recommendation Models}},
        author = {Zhang, Hailin and Liu, Zirui and Chen, Boxuan and Zhao, Yikai and Zhao, Tong and Yang, Tong and Cui, Bin},
        series = {{SIGMOD} '24},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/3639306},
        url = {https://dl.acm.org/doi/10.1145/3639306},
        year = {2024}
}

Incoming Citations (Sorted by Pagerank)

Showing 2 of 2 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 18 of 18 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
489 Distributed Representations of Tuples for Entity Resolution 2018 VLDB 0.0001761456
865 Natural language to SQL: Where are we today? 2020 VLDB 0.00013521464
1,294 Augmented Sketch: Faster and More Accurate Stream Processing 2016 SIGMOD 0.00011291308
2,355 QueryFormer: A Tree Transformer Model for Query Plan Representation 2022 VLDB 8.7022189e-05
2,485 HET: Scaling out Huge Embedding Model Training via Cache-enabled Distributed Framework 2022 VLDB 8.5145736e-05
2,620 Fauce: Fast and Accurate Deep Ensembles with Uncertainty for Cardinality Estimation 2021 VLDB 8.3363963e-05
2,878 Data Sketches for Disaggregated Subset Sum and Frequent Item Estimation 2018 SIGMOD 8.0058242e-05
3,516 LOGER: A Learned Optimizer towards Generating Efficient and Robust Query Execution Plans 2023 VLDB 7.3524442e-05
3,896 Scaling Attributed Network Embedding to Massive Graphs 2021 VLDB 7.0406415e-05
4,912 HET-GMP: A Graph-based System Approach to Scaling Large Embedding Model Training 2022 SIGMOD 6.4481656e-05
5,255 At-the-time and Back-in-time Persistent Sketches 2021 SIGMOD 6.2982495e-05
5,305 Parallel Training of Knowledge Graph Embedding Models: A Comparison of Techniques 2022 VLDB 6.2732352e-05
6,633 Agile and Accurate CTR Prediction Model Training for Massive-Scale Online Advertising Systems 2021 SIGMOD 5.8185371e-05
7,005 Effective and Efficient Retrieval of Structured Entities 2020 VLDB 5.7269998e-05
7,069 SKT: A One-Pass Multi-Sketch Data Analytics Accelerator 2021 VLDB 5.7117595e-05
7,268 Cardinality Estimation of Approximate Substring Queries using Deep Learning 2022 VLDB 5.6597123e-05
8,310 TreeSensing: Linearly Compressing Sketches with Flexibility 2023 SIGMOD 5.4556836e-05
9,520 Experimental Analysis of Large-scale Learnable Vector Storage Compression 2024 VLDB 5.2561283e-05
Previous Page 1 / 1 Next

Semantically Similar Papers