cs.CVOct 1, 2026

Cog-VADU: A Training-Free Cognitive Reasoning Framework for Video Anomaly Detection and Understanding

Authors: Mohd Ubaid Wani, Sara Atito, Josef Kittler, Muhammad Awais

Organizations: Centre for Vision, Speech and Signal Processing (CVSSP) University of Surrey, UK · Surrey Institute for People-Centred AI (PAI) University of Surrey, UK

Abstract

Video Anomaly Detection (VAD) aims to temporally localize abnormal events in videos. Most existing approaches rely on dataset-specific training and curated annotations, limiting generalization in open-set scenarios. Recent zero-shot methods based on Large Vision- Language Models (LVLMs) alleviate this dependency but often lack temporal continuity and structured reasoning. We propose Cog-VADU, a fully training-free framework that reformulates VAD as a sequential cognitive reasoning task. Cog-VADU introduces Chain-of- Anomaly Detection Thought Prompting (CoADTP), which unrolls an LVLM into a recurrent reasoning chain across video segments. By propagating structured rationales over time, the model maintains implicit temporal memory, enabling robust discrimination between com- plex anomalies and high-motion normal activities. To improve reliability, we further design a cross-modal re-ranking stage that aligns textual rationales with visual embeddings, enforcing semantic consistency and temporal coherence for refined and stable predictions. Extensive experiments on multiple public VAD benchmarks demonstrate that Cog-VADU achieves competitive zero-shot performance. Moreover, cross-model evaluations show that CoADTP consistently enhances reasoning-based anomaly detection in a model-agnostic manner, pro- viding interpretable and generalizable anomaly understanding for real-world applications.

Figures & tables

Appendix figures & tables15 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. CoReVAD: A Contextual Reasoning Framework for Training-Free Video Anomaly Detection

    May 22, 2026Hyeongmuk Lim, Youngbum HurTraining-Free Video Anomaly DetectionFine-Grained Video Understanding

  2. Glance, Scrutinize, and Think: Advancing Video Anomaly Detection from Training-Free to Agentic Reasoning

    Aug 7, 2026Shibo Gao, Peipei Yang, Xu-Yao Zhang +1Training-Free Video Anomaly DetectionFine-Grained Video Understanding

  3. Context-structured Video Anomaly Detection with Large Vision-Language Models

    Jul 21, 2026Dongjun Kim, Changjae Oh, Andrea Cavallaro +1Training-Free Video Anomaly DetectionFine-Grained Video Understanding