cs.ROSep 29, 2026

Reactive Real-Time Flow Policies via Asynchronous Distribution Alignment

Authors: Moritz Zoellner, Reece O'Mahoney, Ioannis Havoutis, Rohan Paleja

Organizations: Department of Computer Science, Purdue University · Oxford Robotics Institute, University of Oxford

Abstract

Generalist robot policies such as vision-language-action models (VLAs) have achieved remarkable generalization, but their inference delays can conflict with the demands of real-time control. Asynchronous execution avoids pauses between action chunks by predicting the next sequence of actions while the robot carries out the previous one. In this paper, we study whether asynchronous execution produces the same action distribution as the original VLA. We find that, for non-Markovian demonstrations, asynchronous execution can produce a fundamentally different action distribution, which can limit the policy's reactivity. In our method, we seek to restore this reactivity by aligning the asynchronously produced action distribution with that of the original VLA through two complementary mechanisms. First, Recursive Flow-Field Distillation trains the asynchronous policy using the VLA's action-generation flow. We characterize the learned distribution theoretically and show experimentally that our asynchronous policy can generate nearly the full range of actions the original VLA would produce, while existing asynchronous methods recover only a fraction of that range. Second, Propose-Resolve prepares multiple action sequences asynchronously and uses the latest observation to select among them based on a lightweight approximation of their likelihood under the VLA's action distribution. Our resulting method matches the original VLA's success on LIBERO and retains about 80% of its success on RoboMimic, about 30 percentage points more than existing asynchronous methods.

Figures & tables

Appendix figures & tables5 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. DEFLECT: Temporal Counterfactual Preference Learning for Delay-Robust Asynchronous VLAs

    May 19, 2026Yixiang Zhu, Yonghao Chen, Zijie Yang +2Counterfactual LearningInference Latency

  2. FutureRTC: Real-Time Robot Execution with Anticipatory-Conditioned Action Chunking

    Jul 27, 2026Hai Jiang, Yixian Zou, Binbin Liang +3Action PredictionRobot Systems