Tool-Using Agents

Momentum

31 papers in the last four weeks, up 210% on the four weeks before. 0.3% of all new papers.

Jul 13Week of Sep 28

Latest papers 199

All topics
CardsList
  1. An Empirical Study of Harness Design for Coding Agents

    Sep 17, 2026Run-Ze Fan, Zihao Zhang, Simin Ma +6Coding AgentsAgent Harness Optimization

  2. Ask the Tool, Don't Guess: Agent Tool Calls Hold Their Progress, and the Serving System Should Read It

    Sep 16, 2026Yipeng Liu, Yingqiang Zhang, Feifei Li +1LLM Inference EfficiencyKV Caching

  3. TuiML: Machine Learning for AI Agents

    Sep 16, 2026Nilesh Verma, Nick Lim, Albert Bifet +1Tool-Using AgentsML Reproducibility

  4. Semantic-TVM: Structure-Preserving Trustworthy Virtual Memory for Memory-Augmented and Tool-Using Agents

    Sep 14, 2026Yu Li, Qikun Cai, Tao Huang +1Agent MemoryPrivacy-Preserving Language Models

  5. ParaRecover: A Process-Level Benchmark for Error Localization and Recovery in Parallel Tool-Use Agents

    Sep 14, 2026Bowen Guan, Zhentao Yin, Yanming ShenLLM Agent EvaluationAI Agent Benchmarks

  6. Earth-Agent-Pro: Towards Real-World Full-Chain Earth Observation with Agents

    Sep 14, 2026Zhutao Lv, Chenhao Dang, Yi Feng +5Remote SensingAI Agent Benchmarks

  7. Fabrication After Tool Failure: Tool-Augmented Agents Assert Values Their Tools Did Not Return

    Sep 13, 2026Arham Sethi, Arsen Kenzhebayev, Saanvi Paturi +3Tool-Augmented Language Model AgentsDeception in Language Models

  8. The Menu Is an Execution Prior: State-Path Tool Menus for Online Agents

    Sep 10, 2026Bo Yan, Weikai Lin, Song WangTool-Using AgentsLLM Tool Use

  9. From Version Conflicts to Decision Conflicts: Selective Revalidation for Long-Running AI Agents

    Sep 7, 2026Yongjian Lyu, Yang Ren, Ruofei Lai +1Tool-Using AgentsRuntime Enforcement for AI Agents

  10. Speculative Macro Commit for Faster Tool-Using Agents

    Sep 3, 2026Zeyu Liu, Souvik Kundu, Peter A. BeerelTool-Augmented Language Model AgentsLLM Agent Workflow Optimization

  11. Source-Dependent Deference in Medical Imaging Agents Under Falsified Findings: A Pilot Audit

    Aug 30, 2026Ridam Roy, Md Shahriar Rashid, Md. Rajib MiaTool-Use EvaluationMedical VQA

  12. Robustness Analysis of Agentic AI to Inconsistent and Incomplete Tool Responses

    Aug 24, 2026Jiachen Xu, Torben Bach Pedersen, Zhongming Yao +2Agent Failure AnalysisAI Agent Reliability

  13. One Success Isn't Reliability: Thinkingbox, a Sandbox and Benchmark for Agents in Stateful Business Workflows

    Aug 20, 2026Zhuochun Li, Youngmin Ko, Ali Keramati +11Business Process AutomationAI Agent Benchmarks

  14. PIPES: Securing Agent Perception with Provenance and Priors

    Aug 13, 2026Sanjay Kariyappa, Severin Klingler, G. Edward SuhData ProvenanceAI Agent Security

  15. FlowScout: From Execution Feedback to Reliable Tool-Using Agent Workflows

    Aug 10, 2026Shuo Hao, You Lu, Bihuan Chen +1Agentic WorkflowsLLM Agent Workflow Optimization

  16. DOCSCHISEL: Adaptive Tool Documentation Optimization Framework for LLM Agents

    Aug 10, 2026You Lu, Kun Zhang, Bihuan Chen +1Tool-Augmented Language Model AgentsTool-Using Agents

  17. OmnilingualGAIA2: Evaluating the Multilingual Gap in Frontier AI Agents

    Aug 9, 2026Andrea Caciolai, Pere-Lluís Huguet Cabot, Chierh Cheng +11AI Agent EvaluationAI Agent Benchmarks

  18. The Scaffolding Matters More Than the Interface: A Controlled Comparison of MCP and CLI Tool Use Across Seven Agent Scaffoldings, Five Language Models, and One Software Task

    Aug 9, 2026Marc Alier Forment, María José Casañ Guerrero, Francisco José García-Peñalvo +1AI Coding AgentsLLM Agent Evaluation

  19. WebRider: Persona-Conditioned Intent Controllers for Live-Web Assistance

    Aug 7, 2026Zhi Li, Tao Zhou, Yeqing Li +2LLM Agent EvaluationTool-Using Agents

  20. When Do Prompt-Side Agent Playbooks Transfer? Accuracy, Cost, and Runtime Shift in Agent Deployment

    Aug 6, 2026Weihong Lin, Lin Sun, Xiangzheng ZhangLLM Agent EvaluationTool-Using Agents

  21. ETA: A New Agentic Paradigm for Embodied Tasks

    Aug 4, 2026Yitong Chen, Zezheng Huai, Sixian Li +7Tool-Using AgentsRobot Task Planning

  22. SAT-Edge-Agent: Hardware-in-the-Loop Edge-Agent Orchestration for Onboard Satellite Intelligence

    Aug 4, 2026Longji He, Jeto XuEdge ComputingLLM Agent Orchestration

  23. WeClawArena: An Auditable Sandbox and Benchmark for Cross-User Agents Collaboration and Security in Human-Centered Agent Networks

    Aug 4, 2026Prince Zizhuang Wang, Aojie Yuan, Haiyue Zhang +3Multi-Agent CollaborationAI Agent Security Benchmarks

  24. Towards Robust Tool Use in Agents via Experience-Driven Adaptive Guidance

    Aug 4, 2026Can Wang, Haoran Chen, Li Yu +4Tool-Using AgentsLLM Tool Use

  25. ScrambleToolBench: Agents Search Exhaustively Even When Their Own Map Points to the Next Step

    Aug 3, 2026Vernon Toh, Navonil Majumder, Zhengyuan Liu +2LLM Agent EvaluationAI Agent Benchmarks

  26. Homebot: A Personal AI Agent for Conversational Home Assistance and Automation

    Aug 3, 2026Shengyuan Ye, Yixin Zhang, Han Liang +3Tool-Augmented Language Model AgentsTool-Using Agents