cs.CRSep 29, 2026

ModalFidelity: Routing Modalities for Deepfake Detection on a Budget

Authors: Oguzhan Baser, Kaan Kale, Sriram Vishwanath, Sandeep Chinchali

Organizations: The University of Texas at Austin, Austin, TX, USA · Georgia Institute of Technology, Atlanta, GA, USA

Abstract

Deepfakes no longer need to fake a whole video. Generators that read the transcript now alter only the few seconds in which a video's meaning turns, so a forgery hides in a small, unknown fraction of the video. Yet detectors still read every one-second window of both the audio and image streams, spending nearly all of their compute where nothing was altered. We observe that deciding where to look is far cheaper than looking. We present ModalFidelity, a lightweight router that previews each window and decides, before any forensic detector runs, which stream is worth reading, under a hard compute budget it can never exceed. On AV-Deepfake1M, reading at most a fifth of the windows, it is more accurate than gating after the detectors at 15.9x less compute, and retains over 96% of the accuracy of an oracle that knows where every forgery lies.

Figures & tables

Explore similar work

CardsList
  1. Look Before You Judge: Training-Free Region Mining for Grounded and Explainable Deepfake Detection

    Sep 28, 2026Chia-Ling Chen, Yu-Ting Ta, Jian-Yu Jiang-Lin +8Deepfake DetectionUnsupervised Detection

  2. DFD-Lab: A Modular Audio-Visual Deepfake Detection Pipeline

    Sep 20, 2026Jan Rybarczyk, Mateusz Roszkowski, Jacek KomorowskiDeepfake DetectionUnsupervised Detection