AI Agent Security

Momentum

43 papers in the last four weeks, up 258% on the four weeks before. 0.4% of all new papers.

Jul 13Week of Sep 28

Latest papers 311

All topics
CardsList
  1. Hallucination as Exploit: Evidence-Carrying Multimodal Agents

    May 18, 2026Guijia Zhang, Hao Zheng, Harry YangAI Agent SecurityMultimodal Hallucination

  2. Trustworthy Agent Network: Trust in Agent Networks Must Be Baked In, Not Bolted On

    May 18, 2026Yixiang Yao, Yuhang Yao, Xinyi Fan +5Multi-Agent LLM SystemsMulti-Agent Coordination

  3. Agent Security is a Systems Problem

    May 18, 2026Mihai Christodorescu, Earlence Fernandes, Ashish Hooda +11AI Agent Security

  4. Overeager Coding Agents: Measuring Out-of-Scope Actions on Benign Tasks

    May 18, 2026Yubin Qu, Ying Zhang, Yanjun Zhang +4Computer-Use Agent BenchmarksCoding Agents

  5. LivePI: More Realistic Benchmarking of Agents Against Indirect Prompt Injection

    May 18, 2026Lei Zhao, Abhay Bhaskar, Edgar DobribanAI Agent SecurityAI Agent Safety

  6. AI Agents May Always Fall for Prompt Injections

    May 17, 2026Sahar Abdelnabi, Eugene BagdasarianContextual IntegrityAI Agent Security

  7. ADR: An Agentic Detection System for Enterprise Agentic AI Security

    May 17, 2026Chenning Li, Pan Hu, Justin Xu +9MCP SecurityAI Agent Security

  8. Towards trustworthy agentic AI: a comprehensive survey of safety, robustness, privacy, and system security

    May 17, 2026Jinhu Qi, Muzhi Li, Jiahong Liu +9AI Agent ReliabilityAI Agent Evaluation

  9. PrivScope: Task-scoped Disclosure Control for Hybrid Agentic Systems

    May 15, 2026Shafizur Rahman Seeam, Zhengxiong Li, Zhiyuan Yu +3Data LeakagePrivacy-Preserving Language Models

  10. SLEIGHT-Bench: A Benchmark of Evasion Attacks Against Agent Monitors

    May 15, 2026Elle Najt, Colin Toft, Tyler Tracy +2Software Engineering AgentsAdversarial Attacks

  11. The End of Trust: How Agentic AI Breaks Security Assumptions

    May 14, 2026Osama Zafar, Alexander Nemecek, Erman AydayAI Agent SecurityAI Agent Governance

  12. Do Coding Agents Understand Least-Privilege Authorization?

    May 14, 2026Zheng Yan, Jingxiang Weng, Charles Chen +9Software Engineering AgentsAI Agent Security

  13. AgentTrap: Measuring Runtime Trust Failures in Third-Party Agent Skills

    May 13, 2026Haomin Zhuang, Hanwen Xing, Yujun Zhou +5AI Agent SecurityLLM Agent Security

  14. Language-Based Agent Control

    May 13, 2026Timothy Zhou, Loris D'Antoni, Nadia PolikarpovaAI Agent SecurityInformation Flow Control

  15. Large Language Models for Agentic NetOps and AIOps: Architectures, Evaluation, and Safety

    May 12, 2026Muhammad Bilal, Jon Crowcroft, Ruizhi Wang +2LLM Agent EvaluationAI Agent Security

  16. Attacks and Mitigations for Distributed Governance of Agentic AI under Byzantine Adversaries

    May 12, 2026Matthew D. Laws, Alina Oprea, Cristina Nita-RotaruAI Agent SecurityAI Agent Monitoring

  17. No More, No Less: Task Alignment in Terminal Agents

    May 12, 2026Sina Mavali, David Pape, Jonathan Evertz +5Computer-Use Agent BenchmarksAI Alignment

  18. Behavioral Integrity Verification for AI Agent Skills

    May 12, 2026Yuhao Wu, Tung-Ling Li, Hongliang LiuAgent ReliabilityAI Agent Auditing

  19. Can a Single Message Paralyze the AI Infrastructure? The Rise of AbO-DDoS Attacks through Targeted Mobius Injection

    May 12, 2026Zi Liang, Ronghua Li, Yanyun Wang +2AI Agent SecurityLLM Agent Security

  20. Under the Hood of SKILL.md: Semantic Supply-chain Attacks on AI Agent Skill Registry

    May 12, 2026Shoumik Saha, Kazem Faghih, Soheil FeiziAI Agent SecurityLLM Agent Skill Retrieval

  21. ExploitGym: Can AI Agents Turn Security Vulnerabilities into Real Attacks?

    May 11, 2026Zhun Wang, Nico Schiller, Hongwei Li +13AI Agent SecurityAI Agent Security Benchmarks

  22. MATRA: Modeling the Attack Surface of Agentic AI Systems -- OpenClaw Case Study

    May 11, 2026Tim Van hamme, Thomas Vissers, Javier Carnerero-Cano +4AI Risk ManagementAI Agent Security

  23. Red-Teaming Agent Execution Contexts: Open-World Security Evaluation on OpenClaw

    May 11, 2026Hongwei Yao, Yiming Liu, Yiling He +1AI Agent SecurityLLM Agent Security