arXiv AI By Sean Brynj\'olfsson, Shashvat Jayakrishnan, Esha Sali, Diptanshu Purwar, Madhav Aggarwal

RedactionBench

Read the original on arXiv AI →

arXiv:2606. 18782v1 Announce Type: cross Abstract: Large Language Models are increasingly applied to sensitive domains that require redaction of personally identifiable information (PII).

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Aug 20

Redakto - The Incognito Tab for LLMs

Redakto is a new tool designed to anonymize text before it is processed by large language models (LLMs). It offers state‑of‑the‑art redaction of personally identifiable information (PII) and pseudonymization, accessible via a web interface, REST APIs, and model context protocol hooks. The authors evaluate its performance on legal and medical datasets, showing that anonymized texts retain utility comparable to the originals, enabling LLM tasks without significant loss of effectiveness.

By Saurav Kumar Saha, Tom R\"ohr, Felix Bie{\ss}mann
arXiv Computation and Language
Sep 11

On the Impact of Anonymization on the Performance of Large Language Models

The paper systematically studies how anonymizing input data affects large language models (LLMs). Five prominent LLMs were evaluated on eleven benchmarks, comparing performance on original versus pseudonymized inputs. Results show that anonymization generally degrades performance, with larger drops for more capable models and task-dependent effects; reversible anonymization preserves entity uniqueness better than irreversible redaction, and prompting about anonymization offers no benefit.

By Tobias Deu{\ss}er, Max Hahnb\"uck, Lorenz Sparrenberg, Tobias Uelwer, Christian Bauckhage, Rafet Sifa
arXiv AI
Sep 25

ASIRF: An Agentic Framework for Context-Dependent Sensitive Information Redaction

ASIRF (Agentic Sensitive Information Redaction Framework) is a system that retrieves domain‑specific definitions of sensitive information from a flexible knowledge base at inference time, eliminating the need for retraining when adapting to new domains. It offers two architectures—a three‑call multi‑agent pipeline and a single‑agent variant—and has been evaluated on ten small open‑weight models across eight datasets, including out‑of‑distribution fictional domains. In 68 of 80 model‑domain combinations (85 %), ASIRF’s recall surpasses that of the OpenAI Privacy Filter, with most shortfalls limited to the filter’s training‑distribution domains.

By Sudha Priyadarshini, Mohamed Chahine Ghanem