cs.CVOct 5, 2026

Compositional Concept Erasure in Text-to-Image Diffusion Models via Hierarchically Grounded Semantic Surgery

Authors: Chen Dai, Ganyu Zou, Nathan Self, Kevin Piper, Ramachandra Rao Seethiraju, Karthik Shyamsunder, Chang-Tien Lu, Naren Ramakrishnan

Organizations: Department of Computer Science Virginia Tech Alexandria, Virginia, USA · Verisign, Inc. Reston, Virginia, USA

Abstract

Removing copyrighted, unsafe, or user-specified concepts from a deployed text-to-image diffusion model is now a practical requirement. Weight-editing methods can suppress fixed targets, but they require per-target retraining and modify the model checkpoint. Training-free methods, on the other hand, are deployment-friendly, but they suffer from text-side routing failures on compositional prompts. In such prompts, the erase target may be invoked through a related class rather than its lexical name, and its modifiers may migrate onto preserved objects. This paper proposes Hierarchically Grounded Semantic Surgery (HGSS), a training-free framework for compositional concept erasure. The framework lifts both the routing signal and the edit operator used by text-side erasure. First, hierarchical span grounding resolves erase-target spans through lexical, taxonomic, and semantic evidence, while guarding against broad-hypernym and compound-head false positives. Second, dynamic attribute binding refines the text conditioning during early denoising via a counterfactual reference and a preserve-aware cross-attention objective, keeping surviving attribute-noun bindings intact. HGSS selectively removes the erase target without updating model weights or adding learned parameters. On SEE, HGSS cuts hierarchical evasion from 29.54 to 10.02 and roughly halves pairwise attribute leakage, achieving the best Neighbor E and AttrP scores among the reported erasure methods. On UnlearnCanvas, HGSS slightly improves the six-metric average over the matched Semantic Surgery baseline, reaching state-of-the-art.

Figures & tables

Explore similar work

CardsList
  1. Erasing Without Collateral Damage: Precise Concept Removal in Diffusion Models

    Jul 6, 2026Parth Upman, Nishita Jain, Shreyank N GowdaConcept ErasureText-To-Image Diffusion Models

  2. Continual Concept Erasure in Diffusion Models by Suppressing Cross-Edit Interference

    Oct 1, 2026Yongliang Wu, Haori Lu, Jinqi Luo +3Concept ErasureText-To-Image Diffusion Models

  3. GRACE: Adaptive Concept Erasure with Geometry-Guided Retention in Diffusion Models

    Sep 14, 2026Qinghui Gong, Yihuai Liang, Yuanlun Xie +3Concept ErasureText-To-Image Diffusion Models