arXiv AI By Haritz Puerto, Haonan Li, Xudong Han, Timothy Baldwin, Iryna Gurevych

From Leaky Thoughts to Private Reasoning: Controlling What LRMs Say to Themselves

Read the original on arXiv AI →

arXiv:2602. 24210v3 Announce Type: replace-cross Abstract: Large reasoning models (LRMs) produce reasoning traces (RTs) that often contain sensitive information.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.