arxivcs.CVcs.AI2026-06-30
Look But Don't Touch with Sparse Autoencoders for Unlearning in Diffusion Models
Enrico Cassano, Riccardo Renzulli, Rayyan Ahmed, Marco Grangetto, Stephan Alaniz
Sparse autoencoders (SAEs) have recently been proposed as interpretable tools for concept-level manipulation, under the assumption that isolated features can serve as controllable intervention points. In this work, we systematically evaluate this assumption in the context of obje…