Accelerating Triangle Counting on GPU
Summary: Lightweight graph preprocessing boosts GPU triangle counting without changing code. Proposes analytic models for workload imbalance and pattern diversity; uses approx edge directions and vertex reordering to balance load and boost GPU parallelism. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
Incoming Citations (Sorted by Pagerank)
Showing 8 of 8 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 3,028 | NeutronStar: Distributed GNN Training with Hybrid Dependency Management | 2022 | SIGMOD | 7.6833093e-05 |
| 5,483 | Efficient Load-Balanced Butterfly Counting on GPU | 2022 | VLDB | 5.4829054e-05 |
| 10,079 | Fast Optimal Group Steiner Tree Search using GPUs | 2026 | SIGMOD | 4.1905499e-05 |
| 10,085 | GraphTwin: Cache-Centric Bit-Level Graph Representation for Fast and Exact Graph Queries | 2026 | SIGMOD | 4.1905499e-05 |
| 10,495 | Finding Logic Bugs in Graph-processing Systems via Graph-cutting | 2025 | SIGMOD | 4.1905499e-05 |
| 10,604 | Truss Decomposition in Hypergraphs | 2025 | VLDB | 4.1905499e-05 |
| 10,867 | Towards Sufficient GPU-accelerated Dynamic Graph Management: Survey and Experiment | 2025 | VLDB | 4.1905499e-05 |
| 10,875 | Efficient Computation of Hyper-triangles on Hypergraphs | 2025 | VLDB | 4.1905499e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 6 of 6 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 106 | Truss Decomposition in Massive Networks | 2012 | VLDB | 0.00048373761 |
| 1,150 | K-Core Decomposition of Large Networks on a Single PC | 2016 | VLDB | 0.00013647447 |
| 1,973 | Speeding Up Set Intersections in Graph Algorithms using SIMD Instructions | 2018 | SIGMOD | 9.8834701e-05 |
| 2,046 | Efficient Parallel Lists Intersection and Index Compression Algorithms using Graphics Processing Units | 2011 | VLDB | 9.6922338e-05 |
| 3,976 | Accelerating Truss Decomposition on Heterogeneous Processors | 2020 | VLDB | 6.5694893e-05 |
| 4,249 | Fast Sparse Matrix-Vector Multiplication on GPUs: Implications for Graph Mining | 2011 | VLDB | 6.3173602e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| Overall Rank | Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 4,249 | Fast Sparse Matrix-Vector Multiplication on GPUs: Implications for Graph Mining | 2011 | VLDB | 6.3173602e-05 |
| 9,231 | Efficiently Counting Triangles in Large Temporal Graphs | 2025 | SIGMOD | 4.3648789e-05 |
| 5,043 | Better Algorithms for Counting Triangles in Data Streams | 2016 | PODS | 5.7350154e-05 |
| 588 | Massive Graph Triangulation | 2013 | SIGMOD | 0.00019588834 |
| 3,488 | GPU-Accelerated Subgraph Enumeration on Partitioned Graphs | 2020 | SIGMOD | 7.0460627e-05 |
| 1,348 | Counting and Sampling Triangles from a Graph Stream | 2013 | VLDB | 0.00012461666 |
| 3,067 | Sliding Window-based Approximate Triangle Counting over Streaming Graphs with Duplicate Edges | 2021 | SIGMOD | 7.6247945e-05 |
| 5,483 | Efficient Load-Balanced Butterfly Counting on GPU | 2022 | VLDB | 5.4829054e-05 |
| 4,525 | GPU-based Graph Traversal on Compressed Graphs | 2019 | SIGMOD | 6.1087614e-05 |
| 3,976 | Accelerating Truss Decomposition on Heterogeneous Processors | 2020 | VLDB | 6.5694893e-05 |