arXiv AI By Naveen George, Naoki Murata, Yuhta Takida, Konda Reddy Mopuri, Yuki Mitsufuji

TILDE: TILt-based Distributional Erasure for Concept Unlearning

Read the original on arXiv AI →

arXiv:2607. 06432v1 Announce Type: cross Abstract: Concept unlearning in text-to-image diffusion models is critical for safe and practical deployment: with rising privacy concerns, copyright disputes, trademark constraints, and safety regulations, deployed systems must be able to suppress unwanted concepts after training.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Computer Vision
6d ago

Weeding Out Bad Seeds: Initial-Noise-Robust Unlearning for Text-to-Image Diffusion Models

arXiv:2609.37537v1 Announce Type: new Abstract: Machine unlearning has emerged as a critical post-hoc safety measure to erase sensitive concepts from Text-to-Image (T2I) models without prohibitive re...

By Arian Komaei Koma, Seyed Amir Kasaei, Aida Aryafar, Matin Ghiasi, Ali Aghayari, Amirhossein Souri, Mohammad Mosayyebi, AmirMahdi Sadeghzadeh, Mohammad Hossein Rohban
arXiv Computer Vision
4d ago

RASteer: Retain-Aware Activation Steering for Concept Erasure in Diffusion Models

RASteer is a training‑free method for concept erasure in text‑to‑image diffusion models that retains other concepts. It constructs a retain subspace from concepts to preserve, then removes components aligned with this subspace from the erasure direction using Retain‑Orthogonal Steering (ROS). Overlap‑Adaptive Calibration (OAC) further adjusts the removal of shared components at each layer and denoising step, balancing target erasure with concept preservation. Experiments show RASteer matches or outperforms existing activation steering and weight editing baselines on unsafe‑content, instance, and artistic‑style erasure tasks across multiple backbones and benchmarks.

By Yongliang Wu, Haori Lu, Yulun Wu, Jinqi Luo, Xingyu Zhu, Yaoyao Liu
arXiv Machine Learning
6d ago

Reference-Guided Machine Unlearning

Reference-Guided Machine Unlearning (ReGUn) is a vision unlearning framework that prioritizes distributional indistinguishability over degradation-based heuristics. It uses disjoint held-out data to create a class-conditioned reference distribution for distillation, guiding forget samples toward non-member behavior without explicitly degrading predictions. Experiments across various architectures and datasets show that ReGUn achieves a competitive forgetting–utility trade-off and closely matches retrain-like membership inference behavior.

By Jonas Mirlach, Sonia Laguna, Julia E. Vogt