cs.ROOct 8, 2026

CAPABLE: Capability-Aware Policy Adaptation via Behavioral Latent Encoding

Authors: Mohammad Khoshnazar, Mohammad Dehghani Tezerjani, Deyuan Qu, Zhiyuan Gao, Yanxiang Zhan, Jeroen Schafer, Andrew Melnik, Qing Yang, +1 more

Organizations: University of Bremen, Bremen, Germany. · University of North Texas, Denton, TX, USA. · Toyota Motor North America, InfoTech Labs, Mountain View, CA, USA.

Abstract

Vision-language-action (VLA) policies assume the embodiment on which they were trained and can fail when a joint fault changes how commanded actions are physically executed. Existing fault-recovery methods often require task-specific retraining, fault labels, explicit diagnosis, or privileged embodiment information. We introduce CAPABLE, a unified capability-aware adaptation framework for frozen VLAs that integrates self-supervised capability inference with residual reinforcement learning. CAPABLE infers capability, how much of the commanded motion each joint actually realizes and how that motion contributes to end-effector behavior, online from command-response history and kinematics using a temporal encoder shared across joints, Jacobian grounding, cross-joint attention, and self-supervised physical prediction. The resulting representation conditions a residual policy that adds bounded corrections to the VLA arm action without fault labels or faulty-joint identifiers. Across 28 LIBERO tasks, CAPABLE raises success on an actuator excluded from fault training from 24.8% to 59.3%, outperforming a parameter-matched global-history baseline by 17.4 points while preserving healthy performance. Leave-one-actuator-out experiments across six joints show that this transfer is not specific to one actuator, and additional evaluations characterize transfer to unseen fault families and demonstrate recovery on a physical Franka Panda. https://capable-vla.github.io/

Figures & tables

Explore similar work

CardsList
  1. Uncovering Vulnerability of Vision-Language-Action Models under Joint-Level Physical Faults

    Jun 9, 2026Minsoo Jo, Taeju Kwon, Junha Chun +2Robotic ControlVision-Language-Action Models

  2. ReCoVLA: VLM-Guided Reward Compilation for Failure Recovery in Vision-Language-Action Policies

    Jun 8, 2026Haodi Hu, Chung-Ta Huang, Jing Liu +4Robot Failure RecoveryVision-Language-Action Models

  3. RePO-VLA: Recovery-Driven Policy Optimization for Vision-Language-Action Models

    May 10, 2026Weijia Liufu, Xiaoyu Guo, Ruiyi Chen +16VLM RobustnessRobot Failure Recovery