LLM Agents

LLM: Large Language Model

Latest papers 433

All topics
CardsList
  1. NeSyFS: A Neuro-symbolic Fast-Slow Thinking Framework for LLM Agent under Partial Observability

    Jul 31, 2026Duo Xu, Faramarz FekriBelief-Space PlanningLLM Agents

  2. Agentic Coding in the Wild: Characterizing GitHub Copilot Traces at Production Scale

    Jul 30, 2026Banruo Liu, Haoran Qiu, Íñigo Goiri +3LLM InferenceAI Coding Agents

  3. Paying for Honesty Without Knowing the Truth: Reputation-Penalty Design for LLM Marketplace Agents

    Jul 30, 2026Mingdai Yang, Shicheng Fan, Kejing Yu +5Agent ReliabilityDeception in Language Models

  4. Think Short, Defer Smart, Act, and Repeat: Calibrated Reasoning and Uncertainty-Aware Deferral for Edge LLM Agents

    Jul 29, 2026Amirmohammad Farzaneh, Osvaldo SimeoneLLM Inference EfficiencyAI Agent Reliability

  5. PowerAtlas: Towards Electricity-Computing Co-Scheduling for Power Systems

    Jul 29, 2026Kaiwen Jiang, Siya Xu, Ziyue Zhu +3LLM AgentsPower Systems

  6. CaM-Wolf: Causal-Aware Multimodal Agents for Social Deduction Games

    Jul 29, 2026Zheng Zhang, Nanjie Yao, Jiarui He +3Game-Playing AgentsSocial Deduction Games

  7. PatientAgentBench: A Benchmark Framework for Evaluating Patient-Facing Health AI Agents

    Jul 28, 2026Korosh Vatanparvar, Ashutosh Joshi, Maria Xenochristou +11AI Agent EvaluationAI Agent Safety

  8. Matryoshka Agent: Unfolding Sub-Agents for Long-Horizon Machine Learning Engineering

    Jul 27, 2026Rushi Qiang, Changhao Li, Haotian Sun +3LLM Agent OrchestrationLLM Agents

  9. Authoring Agent Skills: A Software-Engineering Approach

    Jul 27, 2026Giuseppe DestefanisLLM Agents

  10. A Self-Calibrating Agentic AI Framework for Autonomous Edge Resource Allocation

    Jul 24, 2026Fin Gentzen, Marla Grunewald, Iulisloi Zacarias +2Edge ComputingLanguage Model Calibration

  11. When Language Models Meet NeuroGraphs: Exploring Enhanced Agentic LLM Framework Towards Brain Network Analysis

    Jul 24, 2026Jiaxing Li, Rui Dong, Muyao Tang +1LLM InterpretabilityLLM Agents

  12. Agentic Evaluation of Copyright Law Compliance

    Jul 23, 2026Zheng Hui, Doni Bloomfield, Noam KoltAI Agent EvaluationAI Agent Benchmarks

  13. HiMe: Real-Time Self-Hosted Personal Agent Platform for Health Insights with Wearable Devices

    Jul 23, 2026Wei Liu, Siya Qi, Linhai Zhang +2LLM AgentsWearable Health Monitoring

  14. EvoDRC: A Self-Evolving Agentic Framework for Automated DRC Violation Repair

    Jul 22, 2026Bing-Yue Wu, Chia-Tung Ho, Haoyu Yang +2Electronic Design AutomationLLM Agents

  15. When Shippers Become Algorithms: Candidate Exposure, Information Design, and the Concentration of LLM-Mediated Freight Markets

    Jul 22, 2026Takahiro Ezaki, Naoto Imura, Katsuhiro NishinariLLM AgentsInformation Design

  16. A Framework of User Experience Principles for Human-AI Agent Interaction in the Workplace

    Jul 22, 2026Kathrin Paimann, Elizangela Valarini, Sebastian JuhlHuman-AI InteractionHuman-Centered AI

  17. Agents in the Wild: Where Research Meets Deployment

    Jul 21, 2026Grace Hui Yang, Pranav N. Venkit, Hooman Sedghamiz +3AI Agent ReliabilityMulti-Agent LLM Systems

  18. AgentDebugX: An Open-Source Toolkit for Failure Observability, Attribution, and Recovery in LLM Agents

    Jul 21, 2026Kunlun Zhu, Xuyan Ye, Zhiguang Han +9Agent Failure AnalysisAI Agent Reliability

  19. Operational Hallucination and Safety Drift in AI Agents

    Jul 20, 2026Shasha Yu, Fiona Carroll, Barry L. BentleyLLM Agent EvaluationAI Agent Safety

  20. LLMs and Agentic AI Systems for Smart Grids: A Tutorial on Architectures and Applications

    Jul 20, 2026Daniela Rojas, Abdulwahab Albassam, Aidan G. Leung +11LLM Agent OrchestrationLLM Agent Evaluation