arXiv AI

Security and Detectability Analysis of Unicode Text Watermarking Methods against Large Language Models

arXiv:2512. 13325v2 Announce Type: replace-cross Abstract: Securing digital text is becoming increasingly relevant due to the widespread use of large language models.

arXiv Machine Learning
Jul 2

Watermarking for Proprietary Dataset Protection

arXiv:2607. 00325v1 Announce Type: new Abstract: A growing body of literature suggests that training data membership inference problems are fundamentally hard tasks in modern language modeling settings.

By John Kirchenbauer, Brian R. Bartoldson, Bhavya Kailkhura, Tom Goldstein
arXiv Machine Learning
Jul 8

Multi-Channel Spread-Spectrum Code Watermarking

arXiv:2607. 06009v1 Announce Type: cross Abstract: Attributing code to the large language model that produced it is essential for provenance, licensing, and misuse accountability, yet no deployed watermark meets this need.

By Soohyeon Choi, Debin Gao, Yue Duan