cs.ROSep 30, 2026

Multi-Link Safety Filtering for VLA Policies Around Moving Hazards

Authors: Yatharth Agarwal, Vijay Raghunathan

Organizations: School of Electrical and Computer Engineering Purdue University, West Lafayette, IN, USA

Abstract

A vision-language-action (VLA) policy can finish a manipulation task while knocking over objects unrelated to it, so task success alone does not show that the policy is safe to deploy in clutter. We study how to keep a pretrained VLA policy clear of such hazards at run time without retraining it, which requires guarding more of the arm than the end effector, following the hazard as it moves, and sharing onboard compute with the policy. Our training-free shield covers the gripper, wrist, and forearm with five ellipsoids and filters every commanded motion through one barrier program against a keep-out ellipsoid fitted from RGB-D perception at reset. Sparse optical flow then carries that ellipsoid's center along with the hazard, with no repeated detection or refitting. Over six simulated hazard-motion conditions, the shield lowers collision from 65.62%65.62\% to 27.27%27.27\% and raises safe-success, task completion without collision, from 29.35%29.35\% to 50.43%50.43\%. Ablations show that guarding the arm links protects beyond end-effector shielding, and that tracking recovers most of the protection lost when the hazard estimate is frozen at reset. On heterogeneous edge hardware, the five-ellipsoid barrier runs on the CPU in 2.22.2~ms at the 99th percentile, and trimming the vision--language prefix and taking fewer flow-matching steps shortens each π0.5π_{0.5} policy call on the integrated GPU from 343343 to 177.3177.3~ms. On a physical SO-101 arm across four tasks, the arm touched the hazard in 3 of 16 shielded episodes versus 11 of 16 unshielded ones. Project page: https://yathag.github.io/multilink-safety-filter/

Figures & tables

Explore similar work

CardsList
  1. Your Model Already Knows: Attention-Guided Safety Filter for Vision-Language-Action Models

    Jun 8, 2026Seongbin Park, Fan Zhang, Baharan Mirzasoleiman +2Safety FiltersCollision Avoidance

  2. SafeLoop: Risk-Aware Rollback for Vision-Language-Action Manipulation

    Sep 22, 2026Zeyu Lou, Tianran Zhang, Xinquan Yue +2Vision-Language-Action FrameworkRobotic Manipulation

  3. SafeVLA-Bench: A Benchmark for the Success-Safety Gap in Vision-Language-Action Models

    May 30, 2026Jialiang Fan, Weizhe Xu, Zijun Wang +3