iBFS: Concurrent Breadth-First Search on GPUs
Summary: iBFS is a GPU-based framework for concurrent BFS from multiple sources, with a single joint-traversal kernel, outdegree-based GroupBy to maximize frontier sharing, and bitwise per-vertex checks across BFS instances. Evaluations show up to 30x single-GPU speedup and near-linear scaling to 112 GPUs, achieving peak TEPS in the tens of trillions. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Hang Liu
- 2. H. Howie Huang
- 3. Yang Hu
Incoming Citations (Sorted by Pagerank)
Showing 10 of 10 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 1,715 | CECI: Compact Embedding Cluster Index for Scalable Subgraph Matching | 2019 | SIGMOD | 0.00010776518 |
| 4,525 | GPU-based Graph Traversal on Compressed Graphs | 2019 | SIGMOD | 6.1087614e-05 |
| 4,578 | Accelerating Dynamic Graph Analytics on GPUs | 2018 | VLDB | 6.0651154e-05 |
| 4,670 | Realtime Top-k Personalized PageRank over Large Graphs on GPUs | 2020 | VLDB | 6.0027844e-05 |
| 5,693 | Parallel Personalized PageRank on Dynamic Graphs | 2018 | VLDB | 5.3683002e-05 |
| 5,966 | Cache-Efficient Fork-Processing Patterns on Large Graphs | 2021 | SIGMOD | 5.2471834e-05 |
| 7,157 | GPU-Accelerated Graph Label Propagation for Real-Time Fraud Detection | 2021 | SIGMOD | 4.8097601e-05 |
| 7,226 | Self-adaptive Graph Traversal on GPUs | 2021 | SIGMOD | 4.7910164e-05 |
| 9,796 | uBlade: Efficient Batch Processing for Uncertain Graph Queries | 2024 | SIGMOD | 4.2777144e-05 |
| 10,146 | CANDOR-Bench: Benchmarking In-Memory Continuous ANNS under Dynamic Open-World Streams [Experiments & Analysis] | 2026 | SIGMOD | 4.1905499e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 6 of 6 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 280 | 3-HOP: A High-Compression Indexing Scheme for Reachability Query | 2009 | SIGMOD | 0.00029092277 |
| 727 | GRAIL: Scalable Reachability Index for Large Graphs | 2010 | VLDB | 0.00017462279 |
| 1,665 | The More the Merrier: Efficient Multi-Source Graph Traversal | 2015 | VLDB | 0.00010958779 |
| 2,762 | K-Reach: Who is in Your Small World | 2012 | VLDB | 8.1614818e-05 |
| 5,494 | Neighborhood-Privacy Protected Shortest Distance Computing in Cloud | 2011 | SIGMOD | 5.4770961e-05 |
| 7,192 | Parallel Graph Processing on Graphics Processors Made Easy | 2013 | VLDB | 4.8000919e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| Overall Rank | Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 4,756 | Efficient GPU-Accelerated Subgraph Matching | 2023 | SIGMOD | 5.9364786e-05 |
| 1,973 | Speeding Up Set Intersections in Graph Algorithms using SIMD Instructions | 2018 | SIGMOD | 9.8834701e-05 |
| 10,079 | Fast Optimal Group Steiner Tree Search using GPUs | 2026 | SIGMOD | 4.1905499e-05 |
| 3,488 | GPU-Accelerated Subgraph Enumeration on Partitioned Graphs | 2020 | SIGMOD | 7.0460627e-05 |
| 5,811 | CGgraph: An Ultra-fast Graph Processing System on Modern Commodity CPU-GPU Co-processor | 2024 | VLDB | 5.3168243e-05 |
| 2,527 | Efficient Algorithms for Maximal k-Biplex Enumeration | 2022 | SIGMOD | 8.5982912e-05 |
| 1,138 | Traversing Large Graphs on GPUs with Unified Memory | 2020 | VLDB | 0.00013715114 |
| 4,525 | GPU-based Graph Traversal on Compressed Graphs | 2019 | SIGMOD | 6.1087614e-05 |
| 5,483 | Efficient Load-Balanced Butterfly Counting on GPU | 2022 | VLDB | 5.4829054e-05 |
| 1,665 | The More the Merrier: Efficient Multi-Source Graph Traversal | 2015 | VLDB | 0.00010958779 |