arXiv AI By Enrico Cassano, Riccardo Renzulli, Rayyan Ahmed, Marco Grangetto, Stephan Alaniz

Look But Don't Touch with Sparse Autoencoders for Unlearning in Diffusion Models

Read the original on arXiv AI →

arXiv:2606. 31699v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have recently been proposed as interpretable tools for concept-level manipulation, under the assumption that isolated features can serve as controllable intervention points.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.