arXiv:2609.05955v1 Announce Type: new
Abstract: Tabular foundation models have become powerful graph learners. Systems such as G2T-FM and GraphPFN encode each node as a feature row and make predictio...
By Mingqi Yang, Zidong Guo, Jihui Yang, Wenming Zuo
VisKG‑LM proposes compiling retrieved knowledge graph subgraphs into static visual memories rather than re‑encoding them during each inference step. The method serializes each subgraph as Relation‑Labeled Paths, renders them as images that preserve the graph’s branching structure, and caches these images for reuse. At inference, a language model processes the question and candidate text first, then consults the cached visual memory only at its final layer, yielding improved performance on CommonsenseQA, OpenBookQA, and MedQA‑USMLE compared to both text‑only baselines and a large vision‑language model.
By Yixin Peng, Er Jin, Shiwei Luo, Diego Collarana, Stefan Decker
arXiv:2608. 12391v1 Announce Type: cross Abstract: Graph reasoning provides a promising testbed for evaluating the reasoning ability of large language models (LLMs), as graph instances can be programmatically generated, structurally controlled, and naturally scaled to long-input settings.
By Fali Wang, Ali Al-Lawati, Iliyas Bektas, Jinxuan Fang, Alek Melenski, Tianxiang Zhao, Yao Ma, Suhang Wang
The paper introduces EffiRAG, a graph-based retrieval‑augmented generation system that reduces the cost of building and querying a graph by using it only to locate relevant passages and generating answers from the original text. On the UltraDomain benchmark, EffiRAG outperforms LightRAG‑hybrid in 93 of 120 questions while cutting total system cost by 57 % (from USD 0.952 to USD 0.408). The study shows that graph‑based RAG can be both more accurate and cheaper, especially as the corpus grows, and recommends evaluating such systems on both answer quality and cost.
By Yuzhong Zhang, Haoyang Ma, Chao Peng, Lionel Briand, Boxi Yu, Jialun Cao
arXiv:2608. 07458v1 Announce Type: cross Abstract: Recent optimization studies on Retrieval-Augmented Generation (RAG) have exploited chunk-level KV cache reuse to avoid processing long retrieved contexts for higher efficiency, while significant information redundancy and noise still remain in the coarse-grained chunks.
By Gyuwan Kim, Cheoneum Park, Tao Yang
LADDER is a new framework that combines diffusion language modeling with Graph Retrieval-Augmented Generation (GraphRAG) to enable efficient multi‑hop reasoning. It introduces an event‑driven self‑clocking retrieval mechanism that triggers graph queries only when new entities appear, and an incomplete‑query graph propagation module that aggregates multi‑hop evidence during parallel decoding. Experiments on three multi‑hop QA benchmarks show that LADDER improves exact match from 39.6% to 45.2% while reducing latency by 4.1×.
By Senlei Zhang, Linhao Luo, Qian-Wen Zhang, Siyu An, Junnan Dong, Shuhao Zhang, Xing Sun
arXiv:2511. 07457v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in modeling sequential textual data and generalizing across diverse tasks.
By Jiarui Feng, Donghong Cai, Yixin Chen, Muhan Zhang
The paper introduces WFM, a Wiki Foundation Model designed to support complex agentic reasoning by combining dense document contexts with markdown-based topological linkages. It formalizes a Wiki Graph schema that preserves explicit topologies while embedding continuous semantics, and employs a query‑conditioned attentive aggregation for efficient message passing. The authors also propose an NCCL‑based protocol to reduce distributed system overhead, achieving a 10.5× training speedup and strong performance on long‑term memory and multi‑hop reasoning benchmarks.
By Junnan Dong, Linhao Luo, Senlei Zhang, Gong Chen, Taian Guo, Yifei Yu, Rong Tao, Tao Guo, Qian-Wen Zhang, Siyu An, Ruizhi Qiao, Xing Sun
The paper investigates a parametric approach to knowledge graph memory by compiling each entity into a LoRA adapter, enabling zero‑cost query-time retrieval via weight injection. On the MetaQA dataset, these adapters encode context‑free factual knowledge, improving exact‑match scores by up to +0.243 over a base model and achieving an oracle gap of +0.283. However, the stored knowledge is not recoverable through similarity or embedding‑based methods, indicating that knowledge is stored locally and does not transfer across semantically neighboring entities.
By Martino M. L. Pulici, Cuong Xuan Chu, Evgeny Kharlamov, Volker Tresp
arXiv:2607. 01241v1 Announce Type: cross Abstract: Existing prompt compression methods treat text as flat token sequences, failing to capture the distributed nature of important information, which is often spread across multiple locations and connected through both local syntactic dependencies and global semantic relations.
By Yaxin Gao, Yao Lu, Jinhong Deng, Jiaqi Nie, Zhe Tang, Jian Zhang, Zhaowei Zhu, Shanqing Yu, Qi Xuan, Joey Tianyi Zhou
arXiv:2601. 08187v3 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated promising capabilities in Text-Attributed Graph (TAG) understanding.
By Zijun Di, Bin Lu, Huquan Kang, Luoyi Fu, Jiaxin Ding, Xiaoying Gan, Lei Zhou, Xinbing Wang
arXiv:2606. 11898v1 Announce Type: cross Abstract: Research on Text-Attributed Graphs (TAGs) has gained significant attention recently due to its broad applications across various real-world data scenarios, such as citation networks, e-commerce platforms, social media, and web pages.
By Hengyi Feng, Zeang Sheng, Meiyi Qiang, Meiyi Qiang, Wentao Zhang