Tool Use

Momentum

10 papers in the last four weeks, up 100% on the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 63

All topics
CardsList
  1. Toward Trustworthy Large Language Model Agents in Healthcare

    Jul 6, 2026Hadi Hasan, Safaa Salman, Adam Tai Abou Dargham +2HealthcareLLM Agent Safety

  2. Why Multi-Step Tool-Use Reinforcement Learning Collapses and How Supervisory Signals Fix It

    Jun 24, 2026Yupu Hao, Zhuoran Jin, Huanxuan Liao +2Supervised Fine-TuningLLM Reliability

  3. Beyond Function Calling: Benchmarking Tool-Using Agents under Tool-Environment Unreliability

    Jun 24, 2026Yang Tian, Zhengpeng Shi, Yu Zhou +1AI Agent ReliabilityAI Agent Benchmarks

  4. PACT: Privileged Trace Co-Training for Multi-Turn Tool-Use Agents

    Jun 15, 2026Zhenbang Du, Jun Luo, Zhiwei Zheng +8Learning from DemonstrationLLM Agent Training

  5. HyperTool: Beyond Step-Wise Tool Calls for Tool-Augmented Agents

    Jun 11, 2026Yaxin Du, Yifan Zhou, Yujie Ge +7Tool-Augmented Language Model AgentsLLM Agent Workflow Optimization

  6. Beyond APIs: Probing the Limits of MLLMs in Physical Tool Use

    Jun 9, 2026Zhixin Ma, Yutong Zhou, Yongqi Li +2Tool Use in VLMsTool-Use Planning

  7. SCOPE: Real-Time Natural Language Camera Agent at the Edge

    Jun 1, 2026Nikolaj Hindsbo, Sina Ehsani, Pragyana MishraEfficient VLM InferenceAI Control

  8. Ghost Tool Calls: Issue-Time Privacy for Speculative Agent Tools

    Jun 1, 2026Bardia Mohammadi, Lars Klein, Akhil Arora +1Data LeakageLLM Agent Security

  9. Learning When Not to Act: Mitigating Tool Abuse in Agentic Reinforcement Learning

    Jun 1, 2026Liuji Chen, Dianxing Tang, Xing Shi +4Agentic RLRL for Language Model Reasoning

  10. DeltaMCP: Incremental Regeneration via Spec-Aware Transformation for MCP servers

    May 27, 2026Aditya Pujara, Xiaogang Zhu, Hsiang-Ting ChenCode GenerationLLM Tool Use

  11. Efficient Agentic Reinforcement Learning with On-Policy Intrinsic Knowledge Boundary Enhancement

    May 26, 2026Dingwei Chen, Zefang Zong, Zhipeng Ma +5Agentic RLLLM Agent Training

  12. Mind the Tool Failures: Achieving Synergistic Tool Gains for Medical Agents

    May 26, 2026Yunhui Gan, Tan Pan, Kaiyu Guo +5Tool UseRL for Tool Use

  13. Enabling Extensible Embodied Capabilities with Tools

    May 26, 2026Xueyang Zhou, Zijia Wang, Qianjiang Li +7Robotic Tool UseEmbodied AI

  14. Diversity Over Frequency: Rethinking Tool Use in Visual Chain-of-Thought Agents

    May 25, 2026Dong-Hee Kim, Reuben Tan, Donghyun KimTool Use in VLMsAgentic Reasoning

  15. AgroTools: A Benchmark for Tool-Augmented Multimodal Agents in Agriculture

    May 21, 2026Zi Ye, Yibin Wen, Xiaoya Fan +10AI Agent BenchmarksTool-Using Vision-Language Agents

  16. ParaVT: Taming the Tool Prior Paradox for Parallel Tool Use in Agentic Video Reinforcement Learning

    May 19, 2026Zuhao Yang, Kaichen Zhang, Sudong Wang +7Tool Use in VLMsAgentic RL

  17. Are Tools Always Beneficial? Learning to Invoke Tools Adaptively for Dual-Mode Multimodal LLM Reasoning

    May 19, 2026Qinghe Ma, Zhen Zhao, Yiming Wu +3Multimodal Large Language ModelsMultimodal Reasoning

  18. TOBench: A Task-Oriented Omni-Modal Benchmark for Real-World Tool-Using Agents

    May 16, 2026Zhiqiang Liu, Wenhui Dong, Yilang Tan +3Tool-Use EvaluationAI Agent Benchmarks

  19. Concurrency without Model Changes: Future-based Asynchronous Function Calling for LLMs

    May 14, 2026Guangyu Feng, Huanzhi Mao, Prabal Dutta +1LLM Agent OrchestrationLLM Tool Use

  20. It's not the Language Model, it's the Tool: Deterministic Mediation for Scientific Workflows

    May 13, 2026Marios Adamidis, Danae Katrisioti, Yannis Tzitzikas +1LLM Tool UseScientific Reproducibility

  21. ReTool-Video: Recursive Tool-Using Video Agents with Meta-Augmented Tool Grounding

    May 13, 2026Xiao Liu, Nayu Liu, Junnan Zhu +6Video QATool-Using Vision-Language Agents

  22. Trajectory Supervision for Continual Tool-Use Learning in LLMs

    May 10, 2026Vishnu Vardhan Reddy, Sagnik Chatterjee, Soumik BhattaContinual Learning for LLMsLLM Tool Use

  23. AutomationBench

    Apr 21, 2026Daniel Shepard, Robin SalimansLLM Agent EvaluationBusiness Process Automation

  24. RaTA-Tool: Retrieval-based Tool Selection with Multimodal Large Language Models

    Apr 16, 2026Gabriele Mattioli, Evelyn Turri, Sara Sarto +4Multimodal Foundation ModelsTool Selection