arXiv AI By Baturay Saglam, Dionysis Kalogerias

Test-Time Detoxification without Training or Learning Anything

Read the original on arXiv AI →

arXiv:2602. 02498v2 Announce Type: replace-cross Abstract: Large language models can produce toxic or inappropriate text even for benign inputs, creating risks when deployed at scale.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.