cs.LGSep 28, 2026

Action Chunking Proximal Policy Optimization with Feedback Correction

Authors: Sanghyun Hahn, Jonghyun Choi

Organizations: Cornell University · Seoul National University

Abstract

Action chunking provides temporal abstraction in reinforcement learning by selecting short action sequences instead of individual actions, but many existing approaches face two limitations in high-dimensional robotic control. First, many rely on value functions over action chunks, which can be difficult to learn as action dimensionality and chunk length grow. Second, executing chunks open-loop removes within-chunk feedback, limiting reactivity in contact-rich tasks. We present Action Chunking PPO (ACPPO), a PPO extension that uses a chunked actor while retaining a standard state-value critic, thereby avoiding chunked Q-functions. We further propose ACPPO-Corr, which augments the chunk planner with a stepwise feedback corrector that adjusts planned actions online within each chunk. Across 25 simulated robotics tasks from IsaacGym and Bi-DexHands, spanning locomotion, arm manipulation, and dexterous hand-object interaction, ACPPO-Corr achieves the strongest aggregate performance among evaluated methods and performs best on both decision-frequency-sensitive and decision-frequency-neutral task subsets. Ablations show that moderate chunk lengths work best and that corrector regularization is important for balancing chunk-level planning with local feedback. These results suggest that action chunking can be effective in online PPO when chunk-level planning is paired with closed-loop correction. The code is available at: https://github.com/hshhahn/ACPPO.

Figures & tables

Appendix figures & tables8 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Why Does Action Chunking Improve Behavioral Cloning Performance in Robotic Control?

    Aug 3, 2026Filippo Lazzati, Kyle Stachowicz, William Chen +3Action ChunksRobot Policies

  2. Adaptive Q-Chunking for Offline-to-Online Reinforcement Learning

    May 7, 2026Nandiraju Gireesh, Yuanliang Ju, He WangAction ChunksChunk

  3. PACE: Phase-Aware Chunk Execution for Robot Policies with Action Chunking

    May 30, 2026Junnan Nie, Jiayi Li, Chenghao Liu +5Action ChunksRobot Policies