cs.AIOct 6, 2026

RippleCP: Measuring Counterfactual Checkpoint Advantage in Coding Agents

Authors: Mayur Akewar, Ravi Ranjan

Organizations: Florida International University, Miami, USA

Abstract

Agent checkpoint systems decide what state is recovery-relevant, how to snapshot it, and whether rollback is admissible. None decides which of the safe boundaries they expose are worth materializing. We formulate this as counterfactual checkpoint advantage, the reduction in future recovery cost obtained by checkpointing a candidate rather than skipping it, and measure it by driving a CP branch and a SKIP branch to the same logical failure and recovering both under matched model, tool, verifier, and stopping conditions. On a frozen pilot of 12 SWE-bench Verified tasks and 106 real recovery branches, checkpointing saves 49.4 s per task, and that figure resolves into two regimes two orders of magnitude apart. The first checkpoint returns 100.0 s on 156.5 s of protected work, a conversion of 0.64; a second one step later returns-1.1 s on 41.8 s, a conversion of -0.03. Recovery is a re-derivation rather than a replay, so preserved work is a poor guide to saved work, and the classical elapsed-work rule misprices the second checkpoint by its full nominal cost. We identify where placement can pay, and set the bar a placement policy must clear.

Figures & tables

Appendix figures & tables11 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. CodeRescue: Budget-Calibrated Recovery Routing for Coding Agents

    Jul 21, 2026Qijia He, Jiayi Cheng, Chenqian Le +8AI Coding AgentsCost-Aware Inference

  2. Crab: A Semantics-Aware Checkpoint/Restore Runtime for Agent Sandboxes

    Apr 30, 2026Tianyuan Wu, Chaokun Chang, Lunxi Cao +2Agent ReliabilityLLM Agent Reliability