cs.ROSep 5, 2026

Where Success Breaks: Failure-Boundary Learning for Robust Vision-Language-Action Models

Authors: Yanzhe Chen, Zhijun Cao, Mike Zheng Shou

Organizations: Showlab, National University of Singapore

Abstract

Vision-language-action (VLA) models adapted through supervised fine-tuning (SFT) inherit a structural asymmetry: expert demonstrations teach the policy where success behavior lies, but provide no signal about where it ceases to be reliable. We argue that robust VLA adaptation should therefore be viewed not as further demonstration fitting, but as Failure-Boundary Learning -- the problem of Discovering, Localizing, and Shaping the boundary between recoverable deviations and task failure. To instantiate this view, we propose DLS: built on a real-grounded behavioral prior from few real demonstrations and simulated co-training, DLS discovers failure boundaries at scale through on-policy digital twin rollouts. Rather than reducing each rollout to a binary label, semantic progress localization uses privileged simulator states to assign progress-aware signals that capture where the failure boundary is crossed, not merely whether. These signals drive directional boundary shaping in the flow dynamics -- reinforcing success-producing denoising directions and suppressing failure-producing ones, without action likelihoods or auxiliary critics. Across real-robot manipulation tasks, DLS improves robustness over SFT and online RL baselines, especially under randomized initial states and unseen visual conditions.

Figures & tables

Appendix figures & tables9 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Failing Forward: Adaptive Failure-Informed Learning for Vision-Language-Action Models

    May 8, 2026Meng Zheng, Samhita Marri, Anwesa Choudhuri +6Action GenerationNegative Results

  2. FailPatch: Failure Residual Patching for Vision-Language-Action Models

    Sep 28, 2026Peng Yu, Jiacheng Wang, Ziheng Zhang +6Photometric SupervisionRobot State

  3. RePO-VLA: Recovery-Driven Policy Optimization for Vision-Language-Action Models

    May 10, 2026Weijia Liufu, Xiaoyu Guo, Ruiyi Chen +16Frictive Policy OptimizationImitation