Robustness of VLA Models

VLA: Vision-Language-Action

Latest papers 142

All topics
CardsList
  1. VersaCamVLA: Camera-Configurable VLA Policies for Robotic Manipulation

    Oct 8, 2026Boyao Han, Chen Shi, Jingjing Qian +2Vision-Language-Action ModelsRobotic Manipulation

  2. CAPABLE: Capability-Aware Policy Adaptation via Behavioral Latent Encoding

    Oct 8, 2026Mohammad Khoshnazar, Mohammad Dehghani Tezerjani, Deyuan Qu +6Robot Policy AdaptationRobot Failure Recovery

  3. WARP-VLA: Wrist-Camera Adaptation for View-Robust Policy Execution in Vision-Language-Action Models

    Oct 8, 2026Junmyeong Lee, Dongmin Shin, Min-Gyu Park +3Robustness of VLA ModelsVision-Language-Action Models

  4. When Listening Becomes Easier: Scrubbing Visual Cues for Shortcut-Free VLAs

    Oct 7, 2026Jasper Gerigk, Kenzo Aspuru-Takata, Chin-Hsuan Wu +3Vision-Language-Action ModelsRobustness of VLA Models

  5. Do Vision-Language-Action Models Understand Instructions? A Mechanistic Interpretability Study on Language Grounding

    Oct 7, 2026Theodor Wulff, Angelo CangelosiVision-Language-Action ModelsRobustness of VLA Models

  6. TMT: Runtime Backdoor Detection for Vision-Language-Action Policies on Unseen Tasks

    Oct 7, 2026Zirun Zhou, Jingfeng Zhang, HaoChuan Xu +4Backdoor DetectionRobustness of VLA Models

  7. OGAM: Connecting Systematic Testing to Runtime Assurance through Object-Grounded Attention Monitoring for VLA Policies

    Oct 5, 2026Haki Darwish, Xiangyu Yin, Changwen Li +3Instruction Generalization in VLA ModelsRobustness of VLA Models

  8. Beyond In-Distribution Preservation: Recovering Generalization in Quantized VLAs via Vulnerability-Oriented Tuning

    Oct 5, 2026Shen Ruan, Wenchang Gao, Jin Wang +4VLM QuantizationOOD Generalization

  9. Continuous Conditioning of VLAs with Augmenting EMG and Visual Task Descriptors

    Oct 1, 2026Edward W. Staley, Connor O. Pyles, Rahul Hingorani +5Robotic ControlVision-Language-Action Models

  10. Is Success All You Need? Investigating the Impact of Input Perturbations on VLA Behaviour in Tabletop Manipulation Tasks

    Oct 1, 2026Sophie Higham, Riccardo Andrea Izzo, Matteo Matteucci +1Vision-Language-Action ModelsRobot Manipulation Benchmarks

  11. Multi-Link Safety Filtering for VLA Policies Around Moving Hazards

    Sep 30, 2026Yatharth Agarwal, Vijay RaghunathanRobot SafetySafety Filtering

  12. When Instructions Retrieve Trajectories: Diagnosing and Mitigating Generalization Failures in VLA Models

    Sep 30, 2026Hung-Jen Chen, Yu-Hsun Hou, Yan-Hong Chen +4Robotic ManipulationVision-Language-Action Models

  13. Blackout vs. Freeze: Analyzing Physical Failure Modes of VLAs under Camera Faults

    Sep 30, 2026Heejae Suh, Jongwook Han, Zahra Gholami +1Robotic ControlVision-Language-Action Models

  14. Beyond Prediction: Steering VLM Agents with Retrospective World Modeling

    Sep 30, 2026Yongjiang Liu, Jie Zhang, Haoyue Zhang +3World Model LearningWorld Models

  15. Refusals That Bend: Measuring and Predicting Task Malleability in Embodied VLM Planners

    Sep 30, 2026Leo Y. Lin, Mikhail Kuznetsov, Muslum Ozgur Ozmen +1Robot SafetyLLM Safety Evaluation

  16. RawVLA: Embodied Neural Image Signal Processor For Robotic Manipulation

    Sep 29, 2026Shuhong Liu, Heng Zhou, Lingfeng Qian +7Robotic ManipulationVision-Language-Action Models

  17. Disentangling Spurious Correlations in Vision-Language-Action Models via Predicting Domain-Invariant Latent Lookahead

    Sep 29, 2026Junghyun Kim, Ngseo Kim, ChungWoo Lee +7Domain GeneralizationVision-Language-Action Models

  18. CoRe-VLA: Preserving Cross-View Coordination in VLAs under Camera Shifts

    Sep 29, 2026Tianhang Pan, Xuanhao Wang, Yiwen Pang +4Point Cloud Surface ReconstructionVision-Language-Action Models

  19. Where Predictive Supervision Goes Shapes What VLA Policies Learn

    Sep 29, 2026Hanseul Kim, Jewon Yeom, Youngjoon Jeong +2Visuomotor Policy LearningVision-Language-Action Models

  20. Learning to Act under Visual Interruptions with Vision-Language-Action Models

    Sep 28, 2026Mingle Jiang, Rui Xu, Yunke Wang +1Vision-Language-Action ModelsVision-Based Robot Control

  21. ActionUNet: Improving Robustness of VLA Models with Efficient Multi-scale Fine-tuning

    Sep 28, 2026Di Zhu, Ziheng Yan, Fang WanRobotic ControlEfficient VLA Models

  22. RoboIRGBench: Benchmarking Implicit Referential Grounding in Vision-Language-Action Models

    Sep 28, 2026Aernaer Akelijiang, Jiannan Li, Zhineng Chen +2Language-Conditioned Robot ManipulationVision-Language-Action Models

  23. TLC-DiT: Task-Aligned Local Visual Conditioning for Robust Multitask Robot Manipulation

    Sep 28, 2026Xianbo Cai, Hideyuki Ichiwara, Zihang Wang +2Language-Conditioned Robot ManipulationMulti-Task Robotic Manipulation