cs.CVSep 29, 2026

Targeted Visual Counterfactual Explanations for Contrastive Vision-Language Model

Authors: Van Bach Nguyen, Jörg Schlötterer, Christin Seifer

Organizations: Marburg University, Germany

Abstract

Current explanation methods for contrastive vision-language models such as CLIP mainly identify important regions without showing how to change the input in order to get a target prediction. We introduce Mask-guided Adaptive Counterfactual Explanations (MACE), a targeted visual counterfactual method designed specifically for CLIP zero-shot classification. MACE constructs an editable region from either source attribution or source-target attribution differences and expands the mask only when needed to reach a specified target class. A latent diffusion inpainting model then modifies the selected region, while a frozen CLIP model provides modification guidance and anchors the remaining image content to the original input. We evaluate MACE on ImageNet, Food-101, Oxford Pets, and CUB-200. The source-mask variant achieves the highest target top-1 success rate across all four datasets, while the difference-mask variant produces the smallest pixel level and perceptual changes and the best realism scores. Both variants improve proximity and realism over a Stable Diffusion-only baseline using the same generative backbone. These results show that adaptive mask-guided editing produces effective CLIP counterfactuals. They further reveal a tradeoff between counterfactual validity and source-image preservation.

Figures & tables

Explore similar work

CardsList
  1. FiRe: Fixed-Noise Refinement for Visual Counterfactual Explanations

    Aug 9, 2026Yan Zeng, Changlu Guo, Oskar Kristoffersen +3Counterfactual ExplanationCounterfactuals

  2. EDCT-Bench: Uncovering Faithfulness Gaps in VLMs via Explanation-Driven Counterfactual Testing

    Sep 16, 2026Sihao Ding, Santosh Vasa, Aditi Ramadwar +1Visual EvidenceExplainable AI Methods

  3. Did Models Learn Sufficiently? Attribution-Guided Training via Subset-Selected Counterfactual Augmentation

    Nov 15, 2025Yannan Chen, Ruoyu Chen, Wei Wang +6Counterfactual LearningAttention-Guided Training