cs.AIOct 5, 2026

Understanding and Mitigating Inference-Time Overreliance Using Agentic Memory

Authors: Luoxi Tang, Yuqiao Meng, Nilesh Auradkar, Muchao Ye, Dazheng Zhang, Zhaohan Xi

Abstract

Agentic memory allows LLM agents to reuse past experience, yet retrieved memories can also distort inference even when they are benign, correctly stored, and appropriately retrieved. We study this failure mode, which we call memory over-reliance. Across benchmarks and memory architectures, we find that memory is useful when past experience transfers to the current task, but can become misleading when only part of the evidence transfers. Failures are strongest under partial query-memory overlap, a pattern further confirmed by controlled experiments thatvary the amount of overlapping evidence. Motivated by this finding, we propose MEMTRIM, a plug-and-play framework that indexes memory evidence at write time and controls its reuse at read time. MEMTRIM removes repeated or conflicting evidence while preserving useful memory-specific information, requires no retraining, and applies to both embedding-based and structured memory systems.Experiments show that MEMTRIM reduces memory overreliance while preserving the benefits of useful memory across models and memory settings.

Explore similar work

CardsList
  1. TRUSTMEM: Learning Trustworthy Memory Consolidation for LLM Agents with Long-Term Memory

    Jun 23, 2026Tianyu Yang, Sudipta Paul, Vijay Srinivasan +2LLM Agent MemoryAgent Memory Management

  2. MemTrace: Probing What Final Accuracy Misses in Long-Term Memory

    Jun 15, 2026Xianxuan Long, Zhikai Chen, Shenglai Zeng +3Temporal Reasoning in Language ModelsLLM Agent Memory

  3. Memory as a Controlled Process: Learned Adaptive Memory Management for LLM Agents

    Jul 15, 2026Eric Hanchen Jiang, Zhi Zhang, Yuchen Wu +11Persistent Memory for Language ModelsAgent Memory Management