cs.CLOct 5, 2026

MemPilot: Orchestrating On-Demand Multimodal Memory Curation for LLM Agents

Authors: Haozhen Zhang, Haodong Yue, Quanyu Long, Jianzhu Bao, Qingyuan Liu, Tao Feng, Bohan Liu, Weida Liang, +1 more

Organizations: Nanyang Technological University · Tsinghua University · University of Illinois Urbana-Champaign

Abstract

Memory has become integral to the LLM agent ecosystem, supporting information retention and reuse across interactions. However, most existing agent memory systems construct memory in a query-agnostic manner, which can incur unnecessary preprocessing cost and discard details that later prove essential. Recent studies have begun shifting memory processing toward runtime adaptation, but typically specialize in particular operations or fixed processing schemes, leaving flexible control over performance, cost, and latency largely underexplored. To address this challenge, we present \textbf{MemPilot}, a flexible framework that orchestrates on-demand memory curation under different performance--cost--latency preferences. Specifically, we optimize a multi-step LLM policy via reinforcement learning to iteratively choose between retrieving from query-agnostic memory and delegating query-specific curation of raw multimodal history to heterogeneous LLMs and VLMs. The policy jointly controls evidence amount, curation instructions, model selection, and visual access, enabling fine-grained allocation of runtime computation. To optimize this policy under competing objectives, we adapt objective-wise advantage decoupling by separately estimating each objective's advantage before aggregation. Moreover, we introduce prefix-based marginal utility estimation for fine-grained credit assignment across multi-step rollouts. Experiments on five multimodal agent-memory benchmarks demonstrate favorable performance--cost--latency trade-offs across optimization preferences, with preference sweeps yielding broader frontiers than existing trade-off-aware baselines.

Figures & tables

Appendix figures & tables6 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Memory as a Controlled Process: Learned Adaptive Memory Management for LLM Agents

    Jul 15, 2026Eric Hanchen Jiang, Zhi Zhang, Yuchen Wu +11Large Language Model AgentsGraph-Structured Memories

  2. ElasticMem: Latent Memory as a Learnable Resource for LLM Agents

    May 29, 2026Tao Feng, Chongrui Ye, Tianyang Luo +5Large Language Model AgentsNativemem

  3. MemLens: A Value-Aware Memory Management System with Interactive Analytics for LLM-based Agents

    Jul 28, 2026Shuyue Wei, Chang Liu, Zimu Zhou +2Large Language Model MemoryPeak Memory