cs.CVSep 27, 2026

IVT-Guard: All-in-One Reasoning Model for AI-Generated Content Detection

Authors: Hongwei Niu, Yunpeng Luo, Hanjun Li, Ziyin Zhou, Jianghang Lin, Ke Yan, Shouhong Ding, Shengchuan Zhang, +1 more

Organizations: Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University, Xiamen 361005, P.R. China · Tencent YouTu Lab

Abstract

The rapid proliferation of highly realistic AI-Generated Content (AIGC) necessitates robust and interpretable detection mechanisms. However, existing detectors are predominantly confined to single modalities and provide binary outputs without reasoning. While Multimodal Large Language Models (MLLMs) present a promising solution, their development is constrained by the scarcity of multimodal reasoning data and the reasoning-detection optimization dilemma, where explicit reasoning supervision can compromise detection accuracy. To this end, we introduce IVT-Set, a comprehensive dataset comprising over 152K diverse image, video, and text samples equipped with multi-granularity Chain-of-Thought (CoT) reasoning trajectories. Based on it, we propose IVT-Guard, a pioneering framework for unified and interpretable AIGC detection across image, video, and text modalities. Furthermore, to overcome the aforementioned optimization dilemma, we design a novel three-stage training paradigm: Artifact-Aware Pre-training, Artifact-to-Evidence Supervised Fine-Tuning via artifact-aware injection, and Evidence-Verdict Consistency Group Relative Policy Optimization. Extensive experiments demonstrate that IVT-Guard achieves state-of-the-art detection performance across in-domain, out-of-domain, and cross-dataset settings while delivering faithful reasoning. Code and data will be released.

Figures & tables

Appendix figures & tables25 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. EvoGuard: An Extensible Agentic RL-based Framework for Practical and Evolving AI-Generated Image Detection

    Mar 18, 2026Chenyang Zhu, Maorong Wang, Jun Liu +2Ai-Generated Image DetectionMultimodal Large Language Models

  2. Reasoning-Aware AIGC Detection via Alignment and Reinforcement

    Apr 21, 2026Zhao Wang, Max Xiong, Jianxun Lian +1Ai-Generated Content DetectionGenerative Artificial Intelligence

  3. Video as Natural Augmentation: Towards Unified AI-Generated Image and Video Detection

    May 21, 2026Zhengcen Li, Chenyang Jiang, Liangxu Su +4Ai-Generated Video DetectionCross-Modal