arxivcs.CVcs.LG2026-07-15
Inference-Time Concept Suppression and Video-Centric Evaluation for Text-to-Video Models
Text-to-video (T2V) generators can synthesize realistic and temporally coherent videos, but controllably removing a target concept from a generator remains difficult. Unlike text-to-image concept erasure, T2V unlearning must suppress a target concept that may persist across frame…