cs.ROSep 24, 2026

Training-free Behavior Cloning

Authors: Maximilian Adang, Timothy Chen, Lars Osterberg, Aiden Swann, Mac Schwager

Organizations: Stanford University, Stanford, CA 94404, USA.

Abstract

Neural behavior cloning compresses demonstrations into large models, making individual actions difficult to trace and policy updates costly. Retrieval policies retain access to demonstrations but struggle with mismatch between recorded and live behavior. We introduce Behavior Predictive Control (BPC), which synthesizes policies without end-to-end policy training by combining an action-aware retrieval metric, a Hankel-based action-continuation prior, and a closed-form one-step residual correction. Inspired by behavioral systems theory, BPC predicts future actions by blending stored observation-action data that best reconstructs the recent runtime observation--action history. Across simulated benchmarks and real-robot deployments, BPC is competitive with learned policies such as π0.5π_{0.5} (surpassing it in some cases), while reducing policy fitting from hours to seconds on consumer GPUs and supporting closed-loop control upwards of 75 Hz on a Jetson Orin Nano. The retrieved demonstration windows and their coefficients also provide an intrinsic estimate of task progress. Retaining demonstrations within the deployed policy makes its predictions traceable to supporting trajectories and enables behavior revision through the demonstration bank.

Figures & tables

Explore similar work

CardsList
  1. When Does Predictive Inverse Dynamics Outperform Behavior Cloning?

    Jan 29, 2026Lukas Schäfer, Pallavi Choudhury, Abdelhak Lemkhenter +10Inverse DynamicsBehavior Cloning

  2. Scalable Behavior Cloning with Open Data, Training, and Evaluation

    Jun 25, 2026Arthur Allshire, Himanshu Gaurav Singh, Ritvik Singh +15Behavior CloningTeleoperation

  3. Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering

    Jun 28, 2026Hao Wang, Jiuzhou Lei, Dayou Li +5Behavior CloningRobot Policies