LLM Agents

LLM: Large Language Model

Latest papers 433

All topics
CardsList
  1. Avatar: Toward Autonomous End-to-End Orchestration of Scientific Workflows using LLMs

    Sep 9, 2026Suman Raj, Hai Duc Nguyen, Haochen Pan +3LLM Agent OrchestrationLLM Agent Workflow Optimization

  2. Belief-State Engine: Augmenting LLMs for Principled Planning Under Partial Observability

    Sep 9, 2026Arnab Chattopadhayay, Debdipta HalderBelief-Space PlanningLLM Agents

  3. LexAgentHallu: A Hierarchical Benchmark for Profiling Hallucinations in Legal Agents

    Sep 9, 2026Yujin Zhou, Mingxuan Zheng, Chuxue Cao +4Hallucination in Language ModelsLLM Agent Evaluation

  4. LLM Agents as Computational Typologists

    Sep 7, 2026Changbing Yang, Christopher Hammerly, Freda Shi +1LLM AgentsLinguistic Typology

  5. SkillAlign: Aligning Skill Interfaces for LLM-based Agents

    Sep 7, 2026Shuo Ren, Xiaomian Kang, Jiajun ZhangLLM Agent EvaluationLLM Agent Skill Learning

  6. ττ\tau^\tau-Bench: An Environment for End-To-End, Realistic Agent Construction

    Sep 7, 2026Quan Shi, Keshav Dhandhania, Karthik Narasimhan +1AI Coding AgentsLLM Agent Evaluation

  7. SLIDEFORGE: An LLM Agent for Controllable Editing of Slides as Structured Artifacts

    Sep 2, 2026Haozhen Zheng, Fulin Wang, Tianhu Xiong +6LLM Agents

  8. CORAL: An LLM-Native Harness for Production Recommender Systems

    Sep 2, 2026Muhammad Rafay Azhar, Yuhang Zhou, Gilbert Jiang +7Recommender SystemsLLM Agents

  9. Competitive Market Behavior of LLMs

    Sep 2, 2026Pawel Struski, Jakub Swistak, Inez Okulska +1LLM Agents

  10. Coverage, Not Targeting: A Structural Regime in Multi-Turn Agent Credit Assignment

    Sep 2, 2026Chenyu Zhou, Qiliang Jiang, Shuning Wu +1Credit AssignmentLLM Agents

  11. APEx: Distillation of Agent Procedural Experience for Adaptive Deep Research Question Answering

    Sep 2, 2026Jie Ding, Rui Sun, Xinyuan Zhang +2LLM AgentsAgent Memory

  12. AI agents reshape consensus formation in human groups

    Sep 2, 2026Lin Chen, Ziyi Liu, Xia Hu +1Human-AI InteractionLLM Agents

  13. When Guardrails Look Effective: Construct Validity Failures in LLM Agent Commerce Evaluation

    Sep 1, 2026Peiying Zhu, Sidi ChangLLM GuardrailsLLM Agent Evaluation

  14. Polished but Unresolved: Identifying Late-Stage Pressure States in Long-Horizon Tool-Use Agents

    Sep 1, 2026Haoyang Chen, Yi Liu, Jianzhi Shao +3LLM AgentsLong-Horizon LLM Agents

  15. S3C-LLM: Skill-Code Guided Agentic Language Models for Spectrum-to-Structure Elucidation

    Aug 31, 2026Xuanle Zhao, Xinyuan Cai, Xiang Cheng +1LLM Agent Skill LearningLLM Agents

  16. ATLAS: Dual-Horizon Diagnostic Evaluation for Industrial Tool-Use Agents

    Aug 31, 2026Wei Chen, Peilun Zhou, Zhaoyu Hu +8Tool-Use EvaluationLLM Agent Evaluation

  17. Agents in the Large: Perception-Centered Architecture for Persistent Agents

    Aug 31, 2026Shihan Dou, Haoxiang Jia, Shichun Liu +14Cognitive Architectures for AI AgentsLifelong Learning Agents

  18. FaVOR: LLM-Based Agentic Framework for Factor Mining via Empirical Validation

    Aug 31, 2026Hyeonjin Kim, Minseok Kim, Seunghyeon Jung +3Quantitative FinanceInterpretable ML

  19. CAST: Critique-Aware Supervision for Training Reliable Long-Horizon Tool-Calling Agents

    Aug 31, 2026Amir Saeidi, Zehua Zhang, Rishitosh Singh +6LLM AgentsLLM Agent Training

  20. Don't Solve, Just Compare: Tiny Advisors for Runtime Intervention in LLM Agents

    Aug 21, 2026Yanze Jiang, Mingxuan Li, Yuhao Wang +2LLM AgentsLLM Agent Reliability

  21. Optimal Skill Selection for LLM Agents with Provable Bicriteria Guarantees

    Aug 20, 2026Yu Chen, Ruishuo Chen, Xun Wang +2Submodular OptimizationLLM Agent Skill Learning