cs.CVSep 28, 2026

DF-CBM: Region-Aware Concept Bottleneck Models for Deepfake Detection

Authors: Georgios Tsoumplekas, Vazgken Vanian, Alexandros Doumanoglou, Panos K. Papadopoulos, Yannis Spyridis, Dimitrios Zarpalas, Vasileios Argyriou

Organizations: Department of Networks and Digital Media, Kingston University London, UK · Centre for Research and Technology Hellas (CERTH), Thessaloniki, Greece · Department of Computer Science, Kingston University London, UK

Abstract

Deepfake detection methods have become increasingly effective yet most provide limited insight into the evidence behind their predictions. However, in forensic settings users also need to know which manipulation cues support the decision and where they appear. Existing explainability methods only partially address this need since localization-based approaches lack semantic descriptions while language-based explanation methods are only weakly grounded in visual evidence. In this work, we propose DF-CBM, a region-aware concept bottleneck model for explainable deepfake detection. DF-CBM builds a compact vocabulary of manipulation-related concepts from textual artifact annotations and links each concept to plausible facial and boundary regions. It then predicts these concepts from visual features using a concept-specific masked attention mechanism guided by parsed facial masks and the final real/fake decision is made from the predicted concept bottleneck. Our experiments show that DF-CBM outperforms concept-based baselines in concept prediction and deepfake classification while remaining competitive with state-of-the-art black-box detectors. Finally, qualitative results and intervention analyses demonstrate that DF-CBM provides spatially grounded concept evidence and enables counterfactual explanations of how individual manipulation concepts influence the final prediction. Our code is available at: https://github.com/GeorgeTsoumplekas/DF-CBM.

Figures & tables

Explore similar work

CardsList
  1. Explainable Deepfake Detection Challenge

    Jul 23, 2026Abhijeet Narang, Kartik Kuckreja, Shreya Ghosh +4Deepfake DetectionAi-Generated Video Detection

  2. Look Before You Judge: Training-Free Region Mining for Grounded and Explainable Deepfake Detection

    Sep 28, 2026Chia-Ling Chen, Yu-Ting Ta, Jian-Yu Jiang-Lin +8Deepfake DetectionUnsupervised Detection

  3. Why Fake ? Unveiling the Semantic Vocabulary of Deepfake Detectors

    Jul 8, 2026Vazgken Vanian, Alexandros Doumanoglou, Dimitris ZarpalasDeepfake DetectionXai