cs.LGOct 5, 2026

Source-Learned Reliance for Selective Test-Time Adaptation of Multimodal Time Series

Authors: Payal Mohapatra, Yueyuan Sui, Haodong Yang, Benjamin Lundell, Stephen Xia, Qi Zhu

Organizations: Arm Inc., USA · Northwestern University, USA

Abstract

Multimodal wearable systems must remain reliable when sensor streams become noisy or unavailable. Existing multimodal test-time adaptation (TTA) methods often assess reliability online, but cross-modal agreement can be misleading when sensors measure different physical processes, and evaluating alternative modality configurations adds inference cost. We propose CARAT, which decouples model reliance from runtime corruption detection to guide omission or attenuation, amortizing reliance estimation through source training. An asymmetric modality-dropout curriculum prepares a missingness-resilient backbone for omission and derives a frozen, backbone-specific reliance proxy from windowed input-projection gradient norms. At deployment, a lightweight one-class detector flags suspect streams, and the proxy guides a joint choice between replacing the suspect set with the backbone's trained missingness symbol and attenuating its representations before fusion, without candidate-subset evaluation. Across four wearable datasets, five corruption types, three backbones, and eight TTA baselines, CARAT achieves the highest overall macro-F1 and best mean rank (2.42), exceeding EATA, the strongest baseline, by 1.58 F1 points across 12 equally weighted dataset-backbone settings. Across five profiled configurations, CARAT uses 9.49% fewer GFLOPs and updates 47.82% fewer parameters than EATA. A pattern also emerges across sensing regimes: multimodal TTA methods such as PTA are competitive on IMU-dominated homogeneous datasets, whereas unimodal TTA methods like TENT and EATA match or exceed it on heterogeneous datasets. These results position CARAT as a practical default to wearable TTA, offering competitive robustness with modest computational requirements and benefits that vary across backbones and dataset regimes.

Figures & tables

Appendix figures & tables39 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Temporal Structure Matters for Efficient Test-Time Adaptation in Wearable Human Activity Recognition

    May 6, 2026Zishu Zhou, Zaipeng Xie, Xuanyao JieHuman Activity RecognitionTest-Time Adaptation

  2. Multi-modal Test-time Adaptation via Adaptive Probabilistic Gaussian Calibration

    Apr 21, 2026Jinglin Xu, Yi Li, Chuxiong Sun +3ModalitiesConditional Distribution

  3. VCR: Learning Valid Contextual Representation for Incomplete Wearable Signals

    May 13, 2026Yuxuan Weng, Wenhan Luo, Qijia ShaoMissing ModalitiesWearable Physiology