LLM Agent Security

LLM: Large Language Model

Latest papers 234

All topics
CardsList
  1. Swarm-Driven Multi-Agent Reasoning for Smart City Security

    Jul 3, 2026Saeid Jamshidi, Kawser Wazed Nafi, Carol Fung +1Multi-Agent LLM SystemsCybersecurity

  2. Securing Multi-Tool AI Agent Chains With Dynamic, Real-Time Compositional Policies

    Jul 3, 2026Chris Schneider, Kriti Faujdar, Philipp Schoenegger +1LLM Agent SecurityInformation Flow Control

  3. CONTRA: Red-Teaming Configurations of Personalizable Agents

    Jul 3, 2026Jonathan Nöther, Adish Singla, Goran RadanovicLLM Red TeamingAI Agent Security

  4. Steerability via constraints: a substrate for scalable oversight of coding agents

    Jul 2, 2026Thomas WinningerScalable OversightCoding Agents

  5. Understanding and Evaluating Claw-like Agent Security Through a Computer-Systems Lens

    Jun 29, 2026Peizhi Niu, Wenjie Qu, Shangding Gu +14AI Agent SecurityLLM Agent Security

  6. When Latent Agents Lie: KV-Cache Integrity in Multi-Agent LLM Collaboration

    Jun 27, 2026Luís Brito, Carlos BaqueroLLM Agent SecurityMulti-LLM Collaboration

  7. LLM agents security duality: a comprehensive survey of self-security and empowered cybersecurity

    Jun 26, 2026Yiwei Xu, Yong Zhuang, Xuanming Liu +6Autonomous Cyber DefenseLLM Agent Security

  8. Tool Use Enables Undetectable Steganography in Multi-Agent LLM Systems

    Jun 25, 2026Jimmy Laurence Rippin, Simon C. Marshall, David Demitri Africa +1Multi-Agent LLM SystemsMulti-Agent Coordination

  9. A Deterministic Control Plane for LLM Coding Agents

    Jun 25, 2026Padmaraj MadathaAI Coding AgentsLLM Guardrails

  10. Adaptive Evaluation of Out-of-Band Defenses Against Prompt Injection in LLM Agents

    Jun 25, 2026Praneeth Narisetty, Shiva Nagendra Babu Kore, Uday Kumar Reddy Kattamanchi +1LLM Agent EvaluationLLM Agent Security

  11. Red-Teaming the Agentic Red-Team

    Jun 23, 2026Dario Pasquini, Michal Bazyli, Taras Fedynyshyn +1Software SecurityAI Agent Security

  12. Detecting Malicious Agent Skills in the Wild using Attention

    Jun 22, 2026Bacem Etteib, Daniele Lunghi, Tégawendé F. BissyandéLLM Agent SecurityAgent Skill Security Auditing

  13. Safety in Self-Evolving LLM Agent Systems: Threats, Amplification, and Case Studies

    Jun 22, 2026Ruixiao Lin, Xinhao Deng, Qingming Li +12LLM Agent SecuritySelf-Evolving Agents

  14. Black-Box Forensics for Conversational LLM Agents

    Jun 21, 2026Isadora White, Yasaman Jafari, Taylor Berg-KirkpatrickLanguage Model FingerprintingLLM Agent Security

  15. Harness-MU: A Safe, Governed, and Effective Harness for Multi-User LLM Agents

    Jun 20, 2026Wangxuan Fan, Xiaoyu Nie, Zhongxiang DaiPrivacy-Preserving Language ModelsLLM Agent Security

  16. Safe to Check, Unsafe to Use: Relinking at the Compression Boundary of LLM Agents

    Jun 19, 2026Zesen Liu, Zihan Zhang, Dongdong SheAdversarial Attacks on LLMsLLM Agent Security

  17. Local LLM Agents as Vulnerable Runtimes:A Source-Code Audit of the Agent Runtime Layer

    Jun 19, 2026Zhengsong Zhang, Zongze Li, Jiawei Guo +1AI Agent AuditingSoftware Vulnerability Detection

  18. When Lower Privileges Suffice: Investigating Over-Privileged Tool Selection in LLM Agents

    Jun 18, 2026Kaiyue Yang, Yuyan Bu, Jingwei Yi +5LLM Agent SecurityAI Agent Security Benchmarks

  19. SafeClawBench: Separating Semantic, Audit-Evidence, and Sandbox Harm in Tool-Using LLM Agents

    Jun 16, 2026Yuchuan Tian, Mengyu Zheng, Haocheng Mei +5LLM Safety BenchmarksLLM Agent Security

  20. Seeing Is Not Screening: Multimodal Hidden Instruction Attacks on Agent Skill Scanners

    Jun 16, 2026Xiaojun Jia, Jie Liao, Simeng Qin +5LLM Agent SecurityPrompt Injection Defense

  21. The Proxy Knows Too Much: Sealing LLM API Routers with Attested TEEs

    Jun 15, 2026Sipeng Xie, Qianhong Wu, Hengrun Lu +4Remote AttestationLLM Agent Security

  22. SkillVetBench: LLM-as-Judge for Multi-Dimensional Security Risk Evaluation in Open-Source LLM Agent Skills

    Jun 14, 2026Ismail Hossain, Sai Puppala, Md Jahangir Alam +2LLM-as-a-JudgeLLM Agent Security

  23. FragFuse: Bypassing Access Control of Large Language Model Agents via Memory-Based Query Fragmentation and Fusion

    Jun 14, 2026Zixin Rao, Wentian Zhu, Chan Aristella Lu +5Adversarial Attacks on LLMsLLM Agent Security

  24. Benign in Isolation, Harmful in Composition: Security Risks in Agent Skill Ecosystems

    Jun 13, 2026Yi Xie, Jiawei Du, Yu Cheng +2LLM Agent SecurityAI Agent Security Benchmarks

  25. Securing Multi-Agent GIS Systems: Risk Evaluation and Prompt Hardening Optimization

    Jun 13, 2026Kyle Gao, Pranavi Kotta, Linlin Xu +2Adversarial RobustnessLLM Red Teaming