LLM Agents

LLM: Large Language Model

Latest papers 433

All topics
CardsList
  1. Towards Efficient and Evidence-grounded Mobility Prediction with LLM-Driven Agent

    Jun 3, 2026Linyao Chen, Qinlao Zhao, Zechen Li +7LLM AgentsHuman Behavior Prediction

  2. SePO: Self-Evolving Prompt Agent for System Prompt Optimization

    Jun 3, 2026Wangcheng Tao, Han Wu, Weng-Fai WongLLM Agent Self-ImprovementAutomatic Prompt Optimization

  3. Towards Multi-Agent-Simulation-Based Community Note Evaluation

    Jun 3, 2026Changxi Wen, Shuning Zhang, Bohao Chu +5Automated Fact-CheckingLLM Agents

  4. Agent libOS: A Runtime Substrate for Capability-Controlled Self-Evolving LLM Agents

    Jun 2, 2026Yingqi ZhangAI Agent SecurityLLM Agents

  5. SkillDAG: Self-Evolving Typed Skill Graphs for LLM Skill Selection at Scale

    Jun 2, 2026Tong Bai, Zhenglin Wan, Pengfei Zhou +3LLM Agent Skill RetrievalDynamic Graph Learning

  6. Adaptive Latent Agentic Reasoning

    Jun 1, 2026Dongwon Jung, Peng Shi, Yi Zhang +2Agentic ReasoningLLM Agents

  7. ExpWeaver: LLM Agents Learn from Experience via Latent RAG

    May 31, 2026Tao Feng, Tianyang Luo, Jingjun Xu +5Retrieval-Augmented GenerationAgentic RAG

  8. SkillPager: Query-Adaptive Intra-Skill Navigation via Semantic Node Retrieval

    May 30, 2026Zicai Cui, Zihan Guo, Weiwen Liu +1LLM Agent Skill RetrievalLLM Agents

  9. I-WebGenBench : Evaluating Interactivity in LLM-Generated Scientific Web Applications

    May 30, 2026Dasen Dai, Biao Wu, Meng Fang +2AI for ScienceWeb Application Generation

  10. BAGEN: Are LLM Agents Budget-Aware?

    May 29, 2026Yuxiang Lin, Zihan Wang, Mengyang Liu +9LLM Agent EvaluationAI Agent Benchmarks

  11. Skill Availability and Presentation Granularity in Large-Language-Model Agents: A Controlled SkillsBench Study

    May 29, 2026Xiaonan Xu, Wenjing WuLLM Agent EvaluationLLM Agent Skill Learning

  12. COMPASS: Cognitive MCTS-Guided Process Alignment for Safe Search Agents

    May 29, 2026Wenkai Shen, Pengyang Zhou, Jiahe Xu +5LLM Safety AlignmentLLM Agents

  13. Harnessing Agent Skills: Architectural Patterns and a Reference Architecture for Skill-Mediated LLM Agents

    May 29, 2026Boming Xia, Liming Zhu, Zhenchang Xing +3LLM Agent OrchestrationLLM Agent Skill Retrieval

  14. MAVEN: Improving Generalization in Agentic Tool Calling

    May 29, 2026Omkar Ghugarkar, Vishvesh Bhat, Muhammad Ahmed Mohsin +1Logical ReasoningLLM Agent Evaluation

  15. Skill is Not One-Size-Fits-All: Model-Aware Skill Alignment for LLM Agents

    May 29, 2026Jianxiang Yu, Jiapeng Zhu, Bochen Lin +3LLM Agent Skill LearningLLM Agent Skill Retrieval

  16. An Organization-Scoped LLM Agent Runtime Architecture for Regulated Cybersecurity Operations

    May 28, 2026George Fatouros, Georgios Makridis, George Kousiouris +2CybersecurityLLM Agent Security

  17. Specialize Roles, Mix Deployments: Pushing the Cost-Accuracy Frontier of LLM Agent Teams

    May 28, 2026Yinsicheng Jiang, Liang Cheng, Yeqi Huang +6Multi-Agent LLM SystemsLLM Agent Evaluation

  18. Modularizing Educational LLM-Agency for Fostering Responsible Learning Assistance

    May 28, 2026Julius Gabelmann, Felix Jahn, Kevin Baum +4AI in EducationGenerative AI in Education

  19. Do Proactive Agents Need an LLM to Decide When to Act?

    May 28, 2026Xiaoze Liu, Ruowang Zhang, Amir H. Abdi +5Temporal GNNsLLM Agents

  20. PokerSkill: LLMs Can Play Expert-Level Poker without Training or Solvers

    May 28, 2026Boning Li, Baoxiang Wang, Longbo HuangGame-Playing AgentsImperfect-Information Games

  21. SkillsInjector: Dynamic Skill Context Construction for LLM Agents

    May 28, 2026Yanchao Li, Wanhao Liu, Ben Gao +5Tool-Augmented Language Model AgentsLLM Agent Skill Learning

  22. GRASP: Gated Regression-Aware Skill Proposer for Self-Improving LLM Agents

    May 28, 2026Johannes Moll, Jean-Philippe Corbeil, Jiazhen Pan +4LLM Agent Self-ImprovementLLM Agent Skill Learning

  23. VitalAgent: A Tool-Augmented Agent for Reactive and Proactive Physiological Monitoring over Wearable Health Data

    May 28, 2026Di Zhu, Yu Yvonne Wu, Hong Jia +3Tool-Augmented Language Model AgentsLLM Agents

  24. BenchTrace: A Benchmark for Testing Reflection Ability and Controlled Evolution in LLM Agents

    May 28, 2026Jiahao Huang, Fei Cheng, Junfeng Jiang +2LLM Agent EvaluationAI Agent Benchmarks