Tool Use

Momentum

10 papers in the last four weeks, up 100% on the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 63

All topics
CardsList
  1. SchemaFill: Efficient LLM Tool Calling via Slot-Parallel Speculative Decoding

    Oct 5, 2026Zhi-Kai Chen, Song-Yan Li, De-Chuan Zhan +1Speculative DecodingLLM Inference Acceleration

  2. Turnslide: Scalable Multi-Turn Data Synthesis by Walking a Finite-State Machine

    Oct 5, 2026Aaron Fainman, Gabriela Kadlecová, Maciej Gryka +7LLM Fine-TuningSynthetic Data Generation

  3. ParaAgent: Reinforcing Parallel Acting in Open-World Tool Environments

    Sep 27, 2026Shengbin Yue, Hongru Wang, Siyuan Wang +3LLM Agent TrainingTool Use

  4. A Wrong Turn Does Not Ruin the Journey: Deviation-Guided Skill Self-Evolution for LLM Agents

    Sep 24, 2026Yichun Feng, Jiawei Wang, Haozhe SunLLM Agent Self-ImprovementLLM Agent Skill Learning

  5. Toolcompass: Guiding Tool Trialing, Not Suppressing It

    Sep 22, 2026Junlin Fang, Chong Zhang, Do Nguyen-Thanh +3Representation LearningTool-Augmented Language Model Agents

  6. MATCH: Model-Aware Tool Learning with Curriculum Scheduling and Hierarchically Gated Rewards

    Sep 17, 2026Shihao Liu, Hao Yin, Lijun Liu +3Reward ModelingCurriculum RL

  7. Ask the Tool, Don't Guess: Agent Tool Calls Hold Their Progress, and the Serving System Should Read It

    Sep 16, 2026Yipeng Liu, Yingqiang Zhang, Feifei Li +1LLM Inference EfficiencyKV Caching

  8. CoBRA: Learning Tool-Use Boundaries via Counterfactual Margins

    Sep 1, 2026Wenhao Zou, Xianglong Liu, Wendong Bi +3Tool UseRL for Tool Use

  9. One Policy Is Enough: Single-Agent Reinforcement Learning Outperforms Tree Search for Chemistry Tool Learning

    Aug 31, 2026Armin Dariani, Sifan Wu, Bang Liu +1LLM Tool UseTool Use

  10. SPIEval: Evaluating Large Language Models as Mobile Assistants over Scattered Personal Information

    Aug 11, 2026Junjie Ye, Zhuohui Sheng, Shaofan Liu +12LLM EvaluationAI Agent Benchmarks

  11. VTO: Visual Tool Orchestration for Video Anomaly Detection

    Aug 8, 2026Rui Wang, Yeteng Wu, Xianling Zhang +1Reinforcement LearningVideo Anomaly Detection

  12. Which Decisions Low-Bit Quantization Breaks, and How to Predict Them

    Aug 6, 2026Zekun Wu, Swati Dhiman, Adriano KoshiyamaPost-Training QuantizationLow-Bit Quantization

  13. A Master-Slave Robot Manipulator for Needle-Based Teleoperation in MRI Chamber

    Aug 6, 2026Omar Curiel, Jing-Yuan Huang, Po-Chih Chen +6Robot TeleoperationRobotics

  14. ToolLIFT: Lifting Tool-Specific Trajectories into Function-Level Graphs for Generalizable Tool Planning

    Aug 4, 2026Xiuhui You, Jiayi Luo, Zichao Shen +2Tool-Use PlanningTool Use

  15. Towards Robust Tool Use in Agents via Experience-Driven Adaptive Guidance

    Aug 4, 2026Can Wang, Haoran Chen, Li Yu +4Tool-Using AgentsLLM Tool Use

  16. Screenshots or Tools? Eliciting Tool Use and Managing Multimodal Context in Hybrid GUI-MCP Computer-Use Agents

    Aug 4, 2026Siqi Fan, Minghao Li, Xiaoqian Ma +6Computer-Use AgentsRL for LLM Agents

  17. ScrambleToolBench: Agents Search Exhaustively Even When Their Own Map Points to the Next Step

    Aug 3, 2026Vernon Toh, Navonil Majumder, Zhengyuan Liu +2LLM Agent EvaluationAI Agent Benchmarks

  18. Verified Tool Calls Improve LLM Agent Reliability Under Non-Atomic Failures

    Jul 31, 2026Isham Kalappurackal Mansoor, Abhishek Phadke, Pratip RanaAI Agent ReliabilityLLM Tool Use

  19. E-Bench: Benchmarking Multi-Step Tool-Use Agents in Real-World Product Scenarios

    Jul 26, 2026Weihuang Zheng, Tianyuan Zou, Eileen Ye +5Computer-Use Agent BenchmarksTool-Use Evaluation

  20. FuncBridge: Towards Functional Tool-Use Generalization via Keypoint Trajectory Reasoning

    Jul 7, 2026Chuhao Zhou, Liquan Wang, Shuxin Cao +5Robot Skill LearningGeneralization in Robotic Manipulation

  21. Toward Trustworthy Large Language Model Agents in Healthcare

    Jul 6, 2026Hadi Hasan, Safaa Salman, Adam Tai Abou Dargham +2HealthcareLLM Agent Safety

  22. Why Multi-Step Tool-Use Reinforcement Learning Collapses and How Supervisory Signals Fix It

    Jun 24, 2026Yupu Hao, Zhuoran Jin, Huanxuan Liao +2Supervised Fine-TuningLLM Reliability

  23. Beyond Function Calling: Benchmarking Tool-Using Agents under Tool-Environment Unreliability

    Jun 24, 2026Yang Tian, Zhengpeng Shi, Yu Zhou +1AI Agent ReliabilityAI Agent Benchmarks

  24. PACT: Privileged Trace Co-Training for Multi-Turn Tool-Use Agents

    Jun 15, 2026Zhenbang Du, Jun Luo, Zhiwei Zheng +8Learning from DemonstrationLLM Agent Training

  25. HyperTool: Beyond Step-Wise Tool Calls for Tool-Augmented Agents

    Jun 11, 2026Yaxin Du, Yifan Zhou, Yujie Ge +7Tool-Augmented Language Model AgentsLLM Agent Workflow Optimization

  26. Beyond APIs: Probing the Limits of MLLMs in Physical Tool Use

    Jun 9, 2026Zhixin Ma, Yutong Zhou, Yongqi Li +2Tool Use in VLMsTool-Use Planning

  27. SCOPE: Real-Time Natural Language Camera Agent at the Edge

    Jun 1, 2026Nikolaj Hindsbo, Sina Ehsani, Pragyana MishraEfficient VLM InferenceAI Control

  28. Ghost Tool Calls: Issue-Time Privacy for Speculative Agent Tools

    Jun 1, 2026Bardia Mohammadi, Lars Klein, Akhil Arora +1Data LeakageLLM Agent Security

  29. Learning When Not to Act: Mitigating Tool Abuse in Agentic Reinforcement Learning

    Jun 1, 2026Liuji Chen, Dianxing Tang, Xing Shi +4Agentic RLRL for Language Model Reasoning

  30. DeltaMCP: Incremental Regeneration via Spec-Aware Transformation for MCP servers

    May 27, 2026Aditya Pujara, Xiaogang Zhu, Hsiang-Ting ChenCode GenerationLLM Tool Use

  31. Efficient Agentic Reinforcement Learning with On-Policy Intrinsic Knowledge Boundary Enhancement

    May 26, 2026Dingwei Chen, Zefang Zong, Zhipeng Ma +5Agentic RLLLM Agent Training

  32. Mind the Tool Failures: Achieving Synergistic Tool Gains for Medical Agents

    May 26, 2026Yunhui Gan, Tan Pan, Kaiyu Guo +5Tool UseRL for Tool Use

  33. Enabling Extensible Embodied Capabilities with Tools

    May 26, 2026Xueyang Zhou, Zijia Wang, Qianjiang Li +7Robotic Tool UseEmbodied AI

  34. Diversity Over Frequency: Rethinking Tool Use in Visual Chain-of-Thought Agents

    May 25, 2026Dong-Hee Kim, Reuben Tan, Donghyun KimTool Use in VLMsAgentic Reasoning

  35. AgroTools: A Benchmark for Tool-Augmented Multimodal Agents in Agriculture

    May 21, 2026Zi Ye, Yibin Wen, Xiaoya Fan +10AI Agent BenchmarksTool-Using Vision-Language Agents

  36. ParaVT: Taming the Tool Prior Paradox for Parallel Tool Use in Agentic Video Reinforcement Learning

    May 19, 2026Zuhao Yang, Kaichen Zhang, Sudong Wang +7Tool Use in VLMsAgentic RL

  37. Are Tools Always Beneficial? Learning to Invoke Tools Adaptively for Dual-Mode Multimodal LLM Reasoning

    May 19, 2026Qinghe Ma, Zhen Zhao, Yiming Wu +3Multimodal Large Language ModelsMultimodal Reasoning

  38. TOBench: A Task-Oriented Omni-Modal Benchmark for Real-World Tool-Using Agents

    May 16, 2026Zhiqiang Liu, Wenhui Dong, Yilang Tan +3Tool-Use EvaluationAI Agent Benchmarks

  39. Concurrency without Model Changes: Future-based Asynchronous Function Calling for LLMs

    May 14, 2026Guangyu Feng, Huanzhi Mao, Prabal Dutta +1LLM Agent OrchestrationLLM Tool Use

  40. It's not the Language Model, it's the Tool: Deterministic Mediation for Scientific Workflows

    May 13, 2026Marios Adamidis, Danae Katrisioti, Yannis Tzitzikas +1LLM Tool UseScientific Reproducibility

  41. ReTool-Video: Recursive Tool-Using Video Agents with Meta-Augmented Tool Grounding

    May 13, 2026Xiao Liu, Nayu Liu, Junnan Zhu +6Video QATool-Using Vision-Language Agents

  42. Trajectory Supervision for Continual Tool-Use Learning in LLMs

    May 10, 2026Vishnu Vardhan Reddy, Sagnik Chatterjee, Soumik BhattaContinual Learning for LLMsLLM Tool Use

  43. AutomationBench

    Apr 21, 2026Daniel Shepard, Robin SalimansLLM Agent EvaluationBusiness Process Automation

  44. RaTA-Tool: Retrieval-based Tool Selection with Multimodal Large Language Models

    Apr 16, 2026Gabriele Mattioli, Evelyn Turri, Sara Sarto +4Multimodal Foundation ModelsTool Selection