LLM Agent Planning

LLM: Large Language Model

Momentum

22 papers in the last four weeks, up 175% on the four weeks before. 0.2% of all new papers.

Jul 13Week of Sep 28

Latest papers 144

All topics
CardsList
  1. AgentFloor: How Far Up the tool use Ladder Can Small Open-Weight Models Go?

    May 1, 2026Ranit Karmakar, Jayita ChatterjeeComputer-Use Agent BenchmarksLLM Agent Evaluation

  2. Bridging Values and Behavior: A Hierarchical Framework for Proactive Embodied Agents

    Apr 30, 2026Chunhui Zhang, Yuxuan Wang, Aoyang Qin +4Hierarchical PlanningAgent Evaluation

  3. AgenticCache: Cache-Driven Asynchronous Planning for Embodied AI Agents

    Apr 27, 2026Hojoon Kim, Yuheng Wu, Thierry TambeLLM PlanningLLM Inference Acceleration

  4. LLM-Guided Agentic Floor Plan Parsing for Accessible Indoor Navigation of Blind and Low-Vision People

    Apr 27, 2026Aydin Ayanzadeh, Tim OatesLLM Agent Planning

  5. From Coarse to Fine: Self-Adaptive Hierarchical Planning for LLM Agents

    Apr 25, 2026Haoran Tan, Zeyu Zhang, Chen Ma +3Hierarchical PlanningLong-Horizon Agent Tasks

  6. WebUncertainty: Dual-Level Uncertainty Driven Planning and Reasoning For Autonomous Web Agent

    Apr 20, 2026Lingfeng Zhang, Yongan Sun, Jinpeng Hu +5Web Search AgentsMonte Carlo Tree Search

  7. Provable Coordination for LLM Agents via Message Sequence Charts

    Apr 19, 2026Benedikt Bollig, Matthias Függer, Thomas NowakMulti-Agent LLM SystemsMulti-Agent Coordination

  8. ProtoCycle: Reflective Tool-Augmented Planning for Text-Guided Protein Design

    Apr 18, 2026Yutang Ge, Guojiang Zhao, Sihang Li +7LLM Self-RefinementLLM Tool Use

  9. PersonalHomeBench: Evaluating Agents in Personalized Smart Homes

    Apr 18, 2026Manasa Bharadwaj, Yolanda Liu, InJung Yang +5AI Agent EvaluationMultimodal Agents

  10. SocialGrid: A Benchmark for Planning and Social Reasoning in Embodied Multi-Agent Systems

    Apr 17, 2026Hikaru Shindo, Hanzhao Lin, Lukas Helff +2LLM Agent EvaluationAI Agent Benchmarks

  11. ProAct: A Benchmark and Multimodal Framework for Structure-Aware Proactive Response

    Feb 3, 2026Xiaomeng Zhu, Fengming Zhu, Weijie Zhou +8AI Agent BenchmarksProactive Assistance

  12. Toward Efficient Agents: Memory, Tool learning, and Planning

    Jan 20, 2026Xiaofang Yang, Lijun Li, Heng Zhou +12LLM Agent MemoryAI Agent Benchmarks

  13. AstroAgentBench: Evaluating Agentic Planning on Space Mission Planning Tasks

    Jan 16, 2026Weiyi Wang, Xinchi Chen, Jingjing Gong +2LLM Agent EvaluationAI Agent Benchmarks

  14. Imagine-then-Plan: Agent Learning from Adaptive Lookahead with World Models

    Jan 13, 2026Youwei Liu, Jian Wang, Hanlin Wang +2World Model-Based PlanningModel-Based Planning

  15. CostBench: Evaluating Multi-Turn Cost-Optimal Planning and Adaptation in Dynamic Environments for LLM Tool-Use Agents

    Nov 4, 2025Jiayu Liu, Cheng Qian, Zhaochen Su +4LLM PlanningLLM Agent Evaluation

  16. CXRAgent: Director-Orchestrated Multi-Stage Reasoning for Chest X-Ray Interpretation

    Oct 24, 2025Jinhui Lou, Yan Yang, Zhou Yu +4Chest X-Ray ClassificationLLM Agent Orchestration

  17. Agent+P: Guiding UI Agents via Symbolic Planning

    Oct 7, 2025Shang Ma, Xusheng Xiao, Yanfang YeGUI AgentsMobile GUI Agents

  18. DeepTravel: An End-to-End Agentic Reinforcement Learning Framework for Autonomous Travel Planning Agents

    Sep 26, 2025Yansong Ning, Rui Liu, Jun Wang +6Agentic RLLLM Agent Training

  19. COCORELI: Enforcing Execution Preconditions for Reliable Collaborative Instruction Following

    Aug 29, 2025Swarnadeep Bhar, Omar Naim, Eleni Metheniti +4Instruction FollowingLLM Agent Planning

  20. City Editing: Hierarchical Agentic Execution for Dependency-Aware Urban Geospatial Modification

    Date pendingRui Liu, Steven Jige Quan, Zhong-Ren Peng +6Hierarchical PlanningGeospatial Reasoning