arXiv Machine Learning

Network Information Enhances Unreliable News Domain Detection

arXiv:2608. 02399v1 Announce Type: cross Abstract: Content-based detection of unreliable news is increasingly difficult, as low-reliability sources mimic credible journalism and generative AI makes fabricated content harder to flag.

arXiv AI
Sep 1

HeTGB: A Comprehensive Benchmark for Heterophilic Text-Attributed Graphs

HeTGB is a new benchmark for heterophilic text‑attributed graphs, consisting of five real‑world datasets where nodes have rich textual descriptions. It allows systematic evaluation of graph neural networks, pre‑trained language models, and co‑training methods on node classification. The benchmark highlights the utility of text attributes, the challenges of heterophilic TAGs, and the limitations of current models.

By Shujie Li, Yuxia Wu, Yuan Fang, Chuan Shi
arXiv AI
Aug 28

Rethinking Message Passing as Retrieval for Text-Attributed Graph Learning

The paper reinterprets graph neural networks (GNNs) as retrieval-augmented models, where each layer uses an MLP on a node representation and a permutation‑invariant summary of retrieved graph context instead of traditional message passing. It introduces RTA, a lightweight MLP‑based framework that replaces structural message passing with label‑aware retrieval and propagation, and provides theoretical links to softmax‑attention message passing and robustness to mis‑retrieved outliers. Experiments on text‑attributed graph benchmarks demonstrate that RTA matches or surpasses strong GNN and graph LLM baselines while improving efficiency and robustness.

By Jintang Li, Yuhong Chen, Ruofan Wu, Binli Luo, Jiayi Ji, Hui Li, Rongrong Ji
arXiv Machine Learning
Sep 10

Not Just Oversmoothing: Detecting the Echo Chamber Effect in Graph Neural Networks

The paper introduces the Echo Chamber Effect, a failure mode in Graph Neural Networks where intra-community representations collapse while inter-community separation remains, differing from traditional oversmoothing. It proposes the Echo Chamber Index (ECI) to detect this effect by stratifying pairwise distances by community membership. Building on this analysis, the authors present Community-Aware Split Propagation (CASP), a lightweight plugin that decouples intra- and inter-community aggregation and learns their balance from label structure, improving performance across various GNN backbones in both homophilic and heterophilic settings.

By Asela Hevapathige, Ahad N. Zehmakan, Asiri Wijesinghe, Saman Halgamuge
arXiv Machine Learning
Aug 24

TH-GNN: Heterogeneous Temporal Graph Neural Networks for LLM-Agent Shilling Attack Detection

TH-GNN is a heterogeneous temporal graph neural network designed to detect shilling attacks generated by large language model (LLM) agents. It combines a two‑layer Heterogeneous Graph Transformer with per‑type and per‑relation attention, learnable sinusoidal temporal encodings, cross‑modal attention that fuses user embeddings with frozen RoBERTa representations of reviews and item descriptions, and a GRU that models log inter‑arrival times. Across five attack families and four benchmark datasets, TH‑GNN achieves a grand‑mean F1 score of 0.870, surpassing the best text‑only baseline on Agent4SR attacks by 10.9 percentage points and 11.5 percentage points at the lowest injection rate.

By Shivam Swarup, Divya Prakash Shrivastava, Rakesh Thakur