arXiv Machine Learning By Xiaopu Wang, Zelin He, Chengyuan Liu, Runze Li

Beyond Heuristic Tuning: Power-Calibrated LLM Watermarking

Read the original on arXiv Machine Learning →

arXiv:2607. 05694v1 Announce Type: cross Abstract: Logit-based watermarking is a widely used mechanism for identifying LLM generated content, yet its effectiveness is governed by a fundamental trade-off between detectability and semantic distortion.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
Sep 24

Optimizing watermarks for large language models

The paper "Optimizing watermarks for large language models" discusses the growing importance of watermarks for generative LLMs amid concerns about misuse. It presents a systematic multi‑objective optimization framework to balance watermark identifiability with the impact on text quality. The authors identify Pareto‑optimal solutions for a broad class of robust, efficient watermarks that outperform the current default watermark.

By Bram Wouters
arXiv AI
Sep 18

Watermarking Diffusion Language Models

The paper introduces the first watermark designed specifically for diffusion language models (DLMs), which generate tokens in arbitrary order unlike traditional autoregressive models. It overcomes the challenge of missing prior tokens by applying the watermark in expectation over the context and promoting tokens that strengthen the watermark when used as context. Experiments show a >99% true positive rate with minimal quality loss and comparable robustness to existing autoregressive watermarks.

By Thibaud Gloaguen, Robin Staab, Nikola Jovanovi\'c, Martin Vechev