arXiv Computation and Language By Haewon Park, Yohan Jo

Knowledge Editing for Masked Diffusion Language Models

Read the original on arXiv Computation and Language →

The paper investigates whether the locate‑then‑edit approach for knowledge editing, previously applied only to autoregressive language models, can be transferred to masked diffusion models (MDMs). It finds that the optimal edit location—an early‑to‑mid‑layer MLP at the last subject token—remains the same for both model types, but that MDMs suffer a sharper decline in performance when editing longer, multi‑token facts. By incorporating intermediate partially‑unmasked states into the edit optimization, the authors restore multi‑token editing performance in MDMs.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Computation and Language.

arXiv AI
3d ago

How Should Diffusion Language Models Edit Code?

arXiv:2609.38257v1 Announce Type: cross Abstract: Code editing requires a model to decide where to make changes, generate the new content, and preserve everything else. We study how masked diffusion...

By Xijia Tao, Ziru Liu, Shansan Gong, Jiacheng Ye, Kecheng Chen, Zirui Wu, Lin Zheng, Xinyu Fu, Rui Liu, Lingpeng Kong
arXiv Computation and Language
3d ago

DEdit: Iterative Draft Editing for Speculative Decoding

arXiv:2609.38510v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive LLMs by having a lightweight drafter propose tokens that the target model verifies in parallel. Diffusi...

By Longxuan Yu, Bingsen Chen, Peng Shi, Dongkyu Lee, Yi Xiang, Hideo Kobayashi, Sheng Zhang, Shuaichen Chang, Xing Niu, Zhuoyan Xu, Greg Ver Steeg, Jiarong Jiang