cs.CVOct 5, 2026

Scalable Minimal-Change Learning for Controllable Image Editing

Authors: Shuo Chen, Fengming Huang, Yu Yao, Mingming Gong, Tongliang Liu

Organizations: Sydney AI Centre, The University of Sydney

Abstract

Image editing should change only the attributes specified by an instruction while preserving everything else, yet current methods often make unintended changes. We treat this minimal-change principle as an optimization objective for instruction-based editing. Latent L1 regularization is a poor proxy for output locality in modern nonlinear generators and often requires supervision unavailable at scale. We instead optimize edit outcomes with reinforcement learning. An agentic vision-language reward model audits each source image, instruction, and edited image for two failure types: unimplemented requested changes and unintended changes. A group-level rubric merges and verifies these issues to provide consistent rewards across candidate edits without per-instruction human annotations. On FLUX.1 Kontext-dev, ARRO raises average EditScore from 5.21 to 5.88 across MinEval, MagicBrush, AnyBench, and Emu-Edit. On 600 evaluation examples, it reduces off-target pixel change by 8.4% relative to the base editor. Reward and SFT controls, blinded human evaluations, and transfer to OmniGen2 provide complementary evidence. Code: https://github.com/Showwwwwwwww/ARRO

Figures & tables

Appendix figures & tables13 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Leveraging Verifier-Based Reinforcement Learning in Image Editing

    Apr 30, 2026Hanzhong Guo, Jie Wu, Jie Liu +6Image EditingContrastive Language-Image Pre-Training Model

  2. Refinement Is Inherently Editable: Training-Free Prompt-to-Prompt Image Editing with Generative Refinement Network

    Sep 17, 2026Yulong Chen, Ziqian Zhang, Haoyu Zhang +4Diffusion-Based Image EditingImage Editing

  3. RewardHarness: Self-Evolving Agentic Post-Training

    May 9, 2026Yuxuan Zhang, Penghui Du, Bo Li +11Progress Reward ModelingMulti-Image Editing Benchmarks