arXiv Machine Learning

Not All Neighbors Matter: Understanding the Impact of Graph Sparsification on GNN Pipelines

arXiv:2603. 06952v2 Announce Type: replace Abstract: As graphs scale to billions of nodes and edges, graph Machine Learning workloads are constrained by the cost of multi-hop traversals over exponentially growing neighborhoods.

arXiv Machine Learning
5d ago

Scaffold: Support Graph Theory Based Sparsification for Graph Neural Networks

Scaffold is a new unsupervised graph sparsification framework for graph neural networks that uses support graph theory preconditioners to jointly control dilation and congestion, thereby preserving short communication paths while avoiding bottlenecks. It achieves superior aggregate ranking across 19 homophilic and heterophilic benchmarks, recovering or closely approaching full‑graph GNN performance with only 10%–50% of the original edges. The method reduces memory usage to less than half and cuts end‑to‑end training time, including sparsification overhead.

By Siddhartha Shankar Das, Sai Karthik Navuluru, S M Ferdous, Ryan A. Rossi, Baris Coskunuzer, Lakshman Tamil, Edoardo Serra, Alex Pothen, Robert Rallo, Mahantesh M Halappanavar
arXiv Machine Learning
Aug 27

DeltaGNN: Graph Neural Network with Information Flow Control

DeltaGNN introduces an information flow control mechanism that uses a new connectivity measure, the information flow score, to mitigate over‑smoothing and over‑squashing in Graph Neural Networks. This approach enables linear computational and memory overhead while effectively capturing both short‑range and long‑range node interactions. Experiments on ten diverse real‑world datasets demonstrate superior performance with limited computational complexity.

By Kevin Mancini, Islem Rekik
arXiv Machine Learning
Aug 19

Asynchronous Message Passing for Addressing Oversquashing in Graph Neural Networks

The paper introduces an asynchronous message‑passing framework for Graph Neural Networks to mitigate oversquashing, a problem where distant nodes cannot effectively communicate due to structural bottlenecks. Unlike conventional synchronous updates, the method updates a centrality‑guided batch of nodes at each layer, allowing information to propagate sequentially and reducing the need for increased channel capacity. Experiments on six standard and two long‑range graph classification benchmarks show notable performance gains, including 5 % improvement on REDDIT‑BINARY and 4 % on Peptides‑struct.

By Kushal Bose, Swagatam Das