arXiv:2606. 12841v1 Announce Type: cross Abstract: Masked diffusion language models (MDLMs) such as LLaDA now rival autoregressive (AR) LLMs, but every existing knowledge-editing and unlearning method (ROME, MEMIT, etc.
By Zhengtao Yao, Liuyang Song, Hongbo Zhang, Chenhao Wei, Haoyan Xu, Guang Yang, Siheng Wang
arXiv:2609.00184v1 Announce Type: cross
Abstract: Large language models (LLMs) rely on static pretraining corpora, causing their knowledge to become outdated over time. Existing approaches for evalua...
By Jonathan Zheng, Zirui Shao, Alan Ritter, Wei Xu
arXiv:2601.07148v4 Announce Type: replace-cross
Abstract: Tool use, such as web search, has become a standard capability even in freely available large language models (LLMs). However, existing bench...
By Zhengxiang Wang, Zeyu Dong
arXiv:2608. 15507v1 Announce Type: cross Abstract: A consistent concept of the current time is important for temporal reasoning, yet how language models represent the current time is not well understood.
By Suze van Adrichem, Aditi Bhaskar, Diyi Yang, Christopher Potts, Jing Huang
arXiv:2510. 27544v2 Announce Type: replace Abstract: Temporal reasoning involves understanding how systems evolve over time through input-driven state transitions.
By Nikolaus Holzer, William Fishell, Baishakhi Ray, Mark Santolucito
Large language models (LLMs) rely on static pretraining corpora, causing their knowledge to become outdated over time. Existing approaches for evaluating knowledge edits either suffer from rapid conta...