HongTu: Scalable Full-Graph GNN Training on Multiple GPUs
Summary: HongTu scales full-graph GNN training on multi-GPU platforms with CPU-memory vertex storage, GPU offload, and recomputation caching. It minimizes host-GPU traffic with deduplicated communication and cost-guided reorganization, and delivers speedups over DistGNN. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Qiange Wang
- 2. Yao Chen
- 3. Weng-Fai Wong
- 4. Bingsheng He
Incoming Citations (Sorted by Pagerank)
Showing 8 of 8 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 7,609 | Systems for Scalable Graph Analytics and Machine Learning: Trends and Methods | 2025 | VLDB | 4.6921981e-05 |
| 10,027 | NeutronHeter: Optimizing Distributed Graph Neural Network Training for Heterogeneous Clusters | 2026 | SIGMOD | 4.1905499e-05 |
| 10,066 | DepCache: A KV Cache Management Framework for GraphRAG with Dependency Attention | 2026 | SIGMOD | 4.1905499e-05 |
| 10,310 | NeutronCloud: Resource-Aware Distributed GNN Training in Fluctuating Cloud Environments | 2026 | VLDB | 4.1905499e-05 |
| 10,548 | Graph Neural Network Training Systems: A Performance Comparison of Full-Graph and Mini-Batch. | 2025 | VLDB | 4.1905499e-05 |
| 10,579 | NeutronTask: Scalable and Efficient Multi-GPU GNN Training with Task Parallelism | 2025 | VLDB | 4.1905499e-05 |
| 10,742 | Faster Convergence in Mini-batch Graph Neural Networks Training with Pseudo Full Neighborhood Compensation | 2025 | VLDB | 4.1905499e-05 |
| 11,029 | Improving Graph Compression for Efficient Resource-Constrained Graph Analytics | 2024 | VLDB | 4.1905499e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 8 of 8 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 271 | AliGraph: A Comprehensive Graph Neural Network Platform | 2019 | VLDB | 0.00029565193 |
| 1,102 | Large Graph Convolutional Network Training with GPU-Oriented Data Communication Architecture | 2021 | VLDB | 0.00014011556 |
| 1,162 | Sancus: Staleness-Aware Communication-Avoiding Full-Graph Decentralized Training in Large-Scale Graph Neural Networks | 2022 | VLDB | 0.00013573136 |
| 2,072 | HippogriffDB: Balancing I/O and GPU Bandwidth in Big Data Analytics | 2016 | VLDB | 9.6300019e-05 |
| 2,292 | Pipelined Query Processing in Coprocessor Environments | 2018 | SIGMOD | 9.0884645e-05 |
| 3,028 | NeutronStar: Distributed GNN Training with Hybrid Dependency Management | 2022 | SIGMOD | 7.6833093e-05 |
| 4,354 | LargeEA: Aligning Entities for Large-scale Knowledge Graphs | 2022 | VLDB | 6.2534656e-05 |
| 5,711 | EMOGI: Efficient Memory-access for Out-of-memory Graph-traversal in GPUs | 2021 | VLDB | 5.3603407e-05 |
Previous
Page 1 / 1
Next