arXiv AI By Zhenlong Liu, Hao Zeng, Weiran Huang, Hongxin Wei

Provable Training Data Identification for Large Language Models

Read the original on arXiv AI →

arXiv:2510. 09717v3 Announce Type: replace-cross Abstract: Identifying training data of large-scale models is critical for copyright litigation, privacy auditing, and ensuring fair evaluation.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.