cs.CVOct 8, 2025

Continual Action Quality Assessment via Adaptive Manifold-Aligned Graph Regularization

Authors: Kanglei Zhou, Qingyi Pan, Xingxing Zhang, Hubert P. H. Shum, Frederick W. B. Li, Xiaohui Liang, Liyuan Wang

Organizations: Department of Psychological and Cognitive Sciences, Tsinghua University, Beijing 100084, China · Department of Statistics and Data Science, Tsinghua University, Beijing 100084, China · Department of Computer Science and Technology, Institute for AI, BNRist Center, Tsinghua-Bosch Joint ML Center, THBI Lab, Tsinghua University · Department of Computer Science, Durham University, DH1 3LE Durham, U.K. · State Key Laboratory of Virtual Reality Technology and Systems, Beihang University, Beijing 100191, China · Zhongguancun Laboratory, Beijing 100190, China

Abstract

Action Quality Assessment (AQA) quantifies human actions in videos, supporting applications in sports scoring, rehabilitation, and skill evaluation. A major challenge lies in the non-stationary nature of quality distributions in real-world scenarios, which limits the generalization of conventional methods. We introduce Continual AQA (CAQA), which equips AQA with Continual Learning (CL) capabilities to handle evolving distributions while mitigating catastrophic forgetting. Although parameter-efficient fine-tuning of pretrained models has shown promise in continual learning, our empirical study shows that the evaluated adapter-based PEFT setting provides less effective downstream adaptation than FPFT for fine-grained AQA. Our empirical and theoretical analyses reveal two insights: (i) sufficiently expressive backbone adaptation is important for bridging the upstream--downstream representation gap; yet (ii) uncontrolled FPFT may induce overfitting and feature manifold shift, thereby aggravating forgetting. To address this, we propose Adaptive Manifold-Aligned Graph Regularization (MAGR++), which couples backbone fine-tuning that stabilizes shallow layers while adapting deeper ones with a two-step feature rectification pipeline: a manifold projector to translate deviated historical features into the current representation space, and a graph regularizer to align local and global distributions. We construct four CAQA benchmarks from three datasets with tailored evaluation protocols and strong baselines, enabling systematic cross-dataset comparison. Extensive experiments show that MAGR++ achieves state-of-the-art performance, with average correlation gains of 3.6% offline and 12.2% online over the strongest baseline, confirming its robustness and effectiveness.

Figures & tables

Appendix figures & tables7 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. MoAKE: Toward Unified All-in-One Action Quality Assessment via Mixture of Action Knowledge Experts

    Jul 22, 2026Huangbiao Xu, Huanqi Wu, Xiao Ke +3Open-Vocabulary Action RecognitionMixtures

  2. HyLoVQA: Dynamic Hypernetwork-Generated Low-Rank Adaptation for Continual Visual Question Answering

    May 21, 2026Yiran Wang, Chenyi Xiong, Ziyue Qin +3Knowledge-Based Visual Question AnsweringVision-Language Model Adaptation

  3. AIM: Asymmetric Information Masking for Visual Question Answering Continual Learning

    Apr 16, 2026Peifeng Zhang, Zice Qiu, Donghua Yu +4Knowledge-Based Visual Question AnsweringMultiple-Choice Visual Question Answering