arXiv Computer Vision By {\L}ukasz Rudnik, Agnieszka Polowczyk, Alicja Polowczyk, Przemys{\l}aw Spurek

FOMO: Forget the Concept, Don't Miss Out on the Scene in Selective Video Unlearning

Read the original on arXiv Computer Vision →

FOMO is a training‑based selective video unlearning method that prioritizes preserving the original scene while removing targeted concepts. It localizes concept‑related representations for modification and employs a preservation mechanism that maintains non‑target scene information without auxiliary data. The approach extends to motion unlearning, enabling removal of concepts defined by temporal behavior, and achieves a strong balance between concept removal and scene preservation.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Computer Vision.

arXiv AI
Sep 18

CleanVideo: Adaptive Concept Erasure for Text-to-Video Diffusion Models

CleanVideo introduces a selective erasure framework for text-to-video diffusion models, addressing the challenge of removing undesired visual concepts from videos. The method uses a low-dimensional subspace intervention guided by a tri-modal gating mechanism that jointly considers spatiotemporal visual features, timestep signals, and textual semantics to decide where, when, and whether to intervene. Experiments on three video diffusion models demonstrate that CleanVideo effectively erases target concepts while preserving visual fidelity, temporal coherence, and outperforming existing baselines in both frame-level and video-level evaluations, even under concept-recovery attacks.

By Junchi Liao, Hongji Li, Wenrui Zhou, Lijie Hu