Gradient Staleness

Momentum

7 papers in the last four weeks, up 17% on the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 46

All topics
CardsList
  1. When Context Changes: Understanding Update Failures in LLMs

    Sep 30, 2026Junyu Guo, Yuchen Fang, Shangding Gu +3Large Language Models FailGradient Staleness

  2. Where Does Staleness Accumulate? Pool Aware Effective Staleness Control for Asynchronous RL in LLM Post-Training

    Sep 29, 2026Chenliang Li, Neiwen Ling, Zijun Wei +1Gradient Staleness

  3. Grow the Harness, Not the Context: From Strategy-Free Scaffolds to Reusable Specialist Agents

    Sep 22, 2026Laizhen Li, Jiarui Li, Juanjuan Zhao +4Agent HarnessScaffolds

  4. Beyond Appearance Shifts: Task-Semantic Action Calibration for VLA Models

    Sep 20, 2026Shuaijun Liu, Feiyang You, Chengyu Wu +5Gradient Staleness

  5. BRACE: Anchored Bellman-Residual Correction for Stale Critics in Asynchronous RL

    Sep 9, 2026Guanqun Zhao, Zijun Xie, Binbin Zheng +3Gradient StalenessLarge Language Model Training

  6. From Version Conflicts to Decision Conflicts: Selective Revalidation for Long-Running AI Agents

    Sep 7, 2026Yongjian Lyu, Yang Ren, Ruofei Lai +1VersionTransaction Evidence

  7. Fresh Memory, Stale Plans: Derivation Currency for Distributed LLM-Agent Memory

    Sep 3, 2026Evan Chen, Shiqiang Wang, Christopher G. BrintonGradient StalenessPeak Memory

  8. REVISE: Validity-Guided Recovery for Online Revisions in Agent Workflows

    Sep 1, 2026Ruoling Qi, Xuaner Wu, Penghang Liu +2RollbackAgentic Workflows

  9. Invalidation Contracts for Cross-Episode Agent Memory

    Aug 31, 2026Michael Wu, Arquimedes CanedoGradient StalenessContracts

  10. Context Is Not Authority: Structured Runtime Governance for Financial Market Agents

    Aug 10, 2026Rui Tang, Qiangqiang Liu, Yichi Zhang +3AuthorityBanking

  11. TEPA: Revoking Stale Memories for Conflict-Robust Language Agents

    Aug 7, 2026Yan Zhou, Yue Ouyang, Kaiyang Zheng +1Long-Term Agent MemoryGradient Staleness

  12. Stateful Governance for Concurrent Agentic Systems

    Aug 3, 2026Yuxiang Peng, Xiaodi WuGovernanceProduction Agentic Systems

  13. SyncPlan: Long-Horizon LLM Coordination with Explicit Synchronization and Adaptive Correction

    Aug 3, 2026Shen You, Xiaoming Zhu, Weining Weng +17CoordinationGradient Staleness

  14. When Memory Updates but Behavior Does Not: Repairing Implicit Stale Dependencies in Personalized Agent Responses

    Aug 3, 2026Haofei Sun, Lin HeGradient StalenessLong-Term Agent Memory

  15. Looping Is Not Reliability: State-Bound Evidence and Typed Revision Contracts for Agentic Code Repair

    Jul 27, 2026Xueping Gao, Jianwei Yang, Qiang YangAutomated Program RepairRepair

  16. Stale but Stable: Staleness-Adaptive Trust Regions for Stabilizing Asynchronous Reinforcement Learning

    Jul 21, 2026Junyao Yang, Yucheng Shi, Zongxia Li +6Trust RegionGradient Staleness

  17. AgentCheck: A Reproduce-Intervene-Mitigate Workbench for LLM Agents over MCP

    Jul 13, 2026Aritra Mazumder, Nusrat jahan LiaLarge Language Model AgentsRetrying

  18. One-Step Gradient Delay is Not a Barrier for Large-Scale Asynchronous Pipeline Parallel LLM Pretraining

    Jun 29, 2026Philip Zmushko, Egor Petrov, Nursultan Abdullaev +2Asynchronous ExecutionLarge Language Model Training

  19. LNN-Fly: Continuous-Time UAV Navigation for Robust Obstacle Avoidance under Timing Mismatch

    Jun 27, 2026Yulin Huang, Shaojie Chen, Di Feng +3Unmanned Aerial VehiclesObstacle Avoidance

  20. AsyncOPD: How Stale Can On-Policy Distillation Be?

    Jun 23, 2026Wonjun Kang, Kevin Galim, Seunghyuk Oh +9Efficient On-Policy DistillationAsynchronous Execution

  21. FedSteer: Taming Extreme Gradient Staleness in Federated Learning with Corrective Projections and Caching

    Jun 8, 2026Haoran Zhang, Cainã Figueiredo Pereira, Marie Siew +3Federated LearningGradient

  22. Context Rot in AI-Assisted Software Development: Repurposing Documentation Consistency for AI Configuration Artifacts

    Jun 8, 2026Christoph Treude, Sebastian BaltesCode QualityDocumentation

  23. A Kinetic Theory of Encounter-Based Information Propagation in Multi-Robot Systems

    Jun 1, 2026Alkesh K. Srivastava, Philip DamesMulti-Robot SystemsStatistical Physics

  24. Masking Stale Observations Helps Search Agents -- Until It Doesn't: A Regime Map and Its Mechanism

    May 29, 2026Haoxiang Zhang, Qixin Xu, Zhuofeng Li +4Search AgentsSingle-Agent Baselines

  25. f\boldsymbol{f}-OPD: Stabilizing Long-Horizon On-Policy Distillation with Freshness-Aware Control

    May 18, 2026Xianwei Chen, Shimin Zhang, Jibin WuRp-OpsdGradient Staleness

  26. When Retrieval Hurts Code Completion: A Diagnostic Study of Stale Repository Context

    May 14, 2026Haojun Weng, Qianqian Yang, Hao Fu +2Code CompletionGradient Staleness

  27. IGT-OMD: Implicit Gradient Transport for Decision-Focused Learning under Delayed Feedback

    May 12, 2026Benjamin Amoh, Geoffrey G. Parker, Wesley MarreroBilevel OptimizationOptimal Regret

  28. Scratchpad Patching: Decoupling Compute from Patch Size in Byte-Level Language Models

    May 10, 2026Lin Zheng, Vasilisa Bashlovkina, Timothy Dozat +3Gradient Staleness

  29. The Geometry of Forgetting: Temporal Knowledge Drift as an Independent Axis in LLM Representations

    May 9, 2026Rania Elbadry, Ahmed Heakl, Fan Zhang +4ForgettingGradient Staleness

  30. FlashEvolve: Accelerating Agent Self-Evolution with Asynchronous Stage Orchestration

    May 8, 2026Zhengding Hu, Mingge Lu, Zhen Wang +8Self-Evolving AgentsAlphaevolve

  31. STALE: Can LLM Agents Know When Their Memories Are No Longer Valid?

    May 7, 2026Hanxiang Chao, Yihan Bai, Rui Sheng +2Agentic MemoryGradient Staleness

  32. Bringing Order to Asynchronous SGD: Towards Optimality under Data-Dependent Delays with Momentum

    May 3, 2026Tehila Dahan, Roie Reshef, Sharon Goldstein +1Stochastic Gradient DescentAsynchronous Execution

  33. Adapt or Forget: Provable Tradeoffs Between Adam and SGD in Nonstationary Optimization

    Date pendingSharan Sahu, Abir Sarkar, Cameron J. Hogan +1AdamStationarity