LLM Agents

LLM: Large Language Model

Latest papers 433

All topics
CardsList
  1. LACUNA: Safe Agents as Recursive Program Holes

    May 27, 2026Yaoyu Zhao, Yichen Xu, Oliver Bračevac +3AI Agent ReliabilityLLM Agents

  2. Audio-Mind: An Auditable Agentic Framework for Audio Understanding

    May 27, 2026Yucheng Wang, Jing Peng, Hanqi Li +6Audio QAAudio Reasoning

  3. Beyond One Path: Evaluating and Enhancing Divergent Thinking in Interactive LLM Agents

    May 27, 2026Jihyeong Park, Ingeol Baek, Jeonghyun Park +1LLM Agent EvaluationAI Agent Benchmarks

  4. Skill-as-Pseudocode: Refactoring Skill Libraries to Pseudocode for LLM Agents

    May 27, 2026Xinze Li, Yuhang Zang, Yixin Cao +1Tool-Augmented Language Model AgentsLLM Agent Skill Retrieval

  5. Retrieval, Reward, and Training Protocols: What Matters in Training Search Agents?

    May 27, 2026Yibo Zhao, Zichen Ding, Jiayi Wu +2LLM AgentsInformation Retrieval

  6. A Query Engine for the Agents

    May 27, 2026Kenny DanielLLM Agents

  7. Maat: The Agentic Legal Research Assistant for Competition Protection

    May 26, 2026Basant Mounir, Farida Madkour, Amira Abdelaziz +1Legal IRAgentic Search

  8. Learning to Act under Noise: Enhancing Agent Robustness via Noisy Environments

    May 26, 2026Yuxin Chen, Xiaodong Cai, Junfeng Fang +9LLM AgentsLLM Agent Training

  9. Automating Formal Verification with Agent-Guided Tree Search

    May 26, 2026Leo YaoCode GenerationLLM-Based Program Synthesis

  10. AGORA: Adapter-Grounded Observation-Action Retention for Inference-Free Prompt Compression in LLM Agents

    May 26, 2026Haoran Zhang, Zhaohua SunLLM CompressionLLM Agents

  11. CRPO: Character-centric Group Relative Policy Optimization for Role-aware Reasoning in Role-playing Agents

    May 25, 2026Yihong Tang, Kehai Chen, Liang Yue +2RL for Language ModelsPersona Consistency

  12. Retrieval as Reasoning: Self-Evolving Agent-Native Retrieval via LLM-Wiki

    May 25, 2026Haoliang Ming, Feifei Li, Xiaoqing Wu +1Multi-Hop QARetrieval-Augmented Generation

  13. CODESKILL: Learning Self-Evolving Skills for Coding Agents

    May 25, 2026Yanzhou Li, Yiran Zhang, Xiaoyu Zhang +2Coding AgentsLLM Agent Skill Learning

  14. Test-Time Deep Thinking to Explore Implicit Rules

    May 24, 2026Wentong Chen, Xin Cong, Zhong Zhang +8RL for Language Model ReasoningLLM Agents

  15. How Many Tools Should an LLM Agent See? A Chance-Corrected Answer

    May 23, 2026Vyzantinos Repantis, Ameya Gawde, Harshvardhan Singh +1Tool-Use EvaluationTool-Augmented Language Model Agents

  16. DeltaBox: Scaling Stateful AI Agents with Millisecond-Level Sandbox Checkpoint/Rollback

    May 21, 2026Yunpeng Dong, Jingkai He, Shiqi Liu +7LLM Agents

  17. Contractual Skills: A GovernSpec Design Framework for Enterprise AI Agents

    May 21, 2026Ting LiuLLM AgentsAI Agent Governance

  18. Ratchet: A Minimal Hygiene Recipe for Self-Evolving LLM Agents

    May 21, 2026Xing Zhang, Yanwei Cui, Guanghui Wang +4Continual Learning for LLM AgentsLLM Agent Self-Improvement

  19. Diagnosis Is Not Prescription: Linguistic Co-Adaptation Explains Patching Hazards in LLM Pipelines

    May 21, 2026Yoon Jeonghun, Kim DongchanLLM AgentsCausal Interventions in Language Models

  20. APEX: Autonomous Policy Exploration for Self-Evolving LLM Agents

    May 20, 2026Yibo Li, Jiashuo Yang, Zhi Zheng +5LLM AgentsLong-Horizon LLM Agents

  21. VBFDD-Agent for Electric Vehicle Battery Fault Detection and Diagnosis: Descriptive Text Modeling of Battery Digital Signals

    May 20, 2026Joey Chan, Zhen Chen, Ershun PanPredictive MaintenanceLLM Agents

  22. A Methodology for Selecting and Composing Runtime Architecture Patterns for Production LLM Agents

    May 19, 2026Vasundra SrinivasanLLM Agent OrchestrationLLM Agents