AI Agent Security

Momentum

43 papers in the last four weeks, up 258% on the four weeks before. 0.4% of all new papers.

Jul 13Week of Sep 28

Latest papers 311

All topics
CardsList
  1. ToolHazard: Scaling Adversarial Environments for Security Evaluation and Alignment of LLM-based Agents

    Aug 12, 2026Yutao Mou, Pengfei Yang, Zhe Yin +6AI Agent EvaluationAdversarial Scenario Generation

  2. ColluSkill: Adversarial Cross-Skill Composition for Evading Agent Skill Scanners

    Aug 10, 2026Puyu Zeng, Simeng Qin, Jingzhi Li +3AI Agent SecurityLLM Agent Security

  3. ActBench: Self-Evolving Benchmark of Behavioral Safety in Cowork Agents

    Aug 10, 2026Hongwei Yao, Yiming Liu, Meihui Chen +6AI Agent SecurityAI Agent Safety

  4. Context Is Not Authority: Structured Runtime Governance for Financial Market Agents

    Aug 10, 2026Rui Tang, Qiangqiang Liu, Yichi Zhang +3Financial ServicesAI Agent Security

  5. Not an A11y: How Android Accessibility Exposes Mobile AI Agents to Indirect Prompt Injection

    Aug 9, 2026Rahul Deivasigamani, Sayeda Faatin Alvi, Derqui Andrea +2AI Agent SecurityLLM Agent Security

  6. HarnessSafe: Evaluating Safety Across Persistent Carriers in Agent Harnesses

    Aug 7, 2026Xiao Zhang, Yusheng Wang, Yuhao Fei +5AI Agent SecurityAI Agent Safety Benchmarks

  7. Hardware Keystores for AI Agent Signing Workflows: A Zero-Trust MCP Enforcement Architecture

    Aug 6, 2026Leo Sambrook, Sampo SovioMCP SecurityAI Agent Security

  8. When Experience Becomes Instruction: Trajectory Poisoning in Self-Evolving Agent Skill Systems

    Aug 6, 2026Jialuo Chen, Lingqi Jiang, Xinhao Deng +7Adversarial AttacksAI Agent Security

  9. Combating Knowledge Corruption in Agent Systems: A Byzantine-Tolerant Secure Collaborative RAG Framework

    Aug 5, 2026Zhaoqi Wang, Daqing He, Zijian Zhang +13RAG Poisoning AttacksAI Agent Security

  10. Internalising the Identity Primitive: Cryptographic Individuality for an Autonomous Agent on a Public Blockchain

    Aug 4, 2026Keisuke SuzukiAI Agent Security

  11. Securing Agentic AI: From Per-Action Checks to Trajectory Assurance

    Aug 3, 2026Alireza Lotfi, Subangkar Karmaker Shanto, Imtiaz Karim +1AI Agent SecurityLLM Agent Security

  12. Adversarial Attacks in Multi-Agent LLM Pipelines: Unveiling Structural Vulnerabilities in Agentic AI Architectures

    Aug 1, 2026Faisal Haque Bappy, Tahrim Hossain, Tarannum Shaila Zaman +3Adversarial Attacks on LLMsAI Agent Security

  13. OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution

    Aug 1, 2026Yunhao Chen, Xin Wang, Yixu Wang +6Adversarial AttacksLong-Horizon Agent Evaluation

  14. Skillsets on the Chain: A Blockchain-based Zero-Trust Framework for Agentic AI Networking

    Jul 31, 2026Yayu Gao, Yong Xiao, Hao Hu +5CybersecurityAI Agent Security

  15. AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents

    Jul 29, 2026Ruoyu Wang, Heng Zhao, Renjie Wu +4Autonomous Cyber DefenseCybersecurity

  16. What Does It Take to Detect an AI Agent? Minimal Feature Sets for Behavioral Detection under Browser Automation

    Jul 29, 2026Vishisht Choudhary, Lukas Schmidt, Anne Zoë Kenntner +3AI Agent SecurityFeature Selection

  17. Distributing Security Controls Through Harness Engineering

    Jul 28, 2026William Robert GoreAI Agent SecurityLLM Agent Security

  18. Cyber-Capable AI Agents: Vulnerabilities, Evaluation Containment, and Defensive Response

    Jul 28, 2026Abu Bakar SiddikAI Agent EvaluationCybersecurity

  19. Where Is the Cost of Third-Party API Routers in Agentic Software Development?

    Jul 26, 2026Donghao Fu, Jingxin Li, Xue Jiang +1Coding AgentsAI Agent Security

  20. False Prophets: On the Security of World Models in Agentic Systems

    Jul 25, 2026Erik Imgrund, Anna Wimbauer, Klim Kireev +1Adversarial AttacksAI Agent Security

  21. Agent Security Needs Redefinition through a Holistic Framework

    Jul 24, 2026Vincent Siu, Jingxuan He, Kyle Montgomery +3AI Agent SecurityAI Agent Security Benchmarks

  22. ToolGuardian: Declarative Security for AI Agent-Tool Interactions

    Jul 23, 2026Arun Ravindran, Saurabh DeochakeAI Agent SecurityLLM Agent Security

  23. Cryptographically verifiable authorization for autonomous AI agents: a falsifiable hypothesis and proof of concept

    Jul 23, 2026M. Llambí-Morillas, D. Fernández-FernándezAI Agent SecurityAI Agent Governance

  24. IssueTrojanBench: Benchmarking AI Coding Agents Against Malicious Issue Requests

    Jul 22, 2026Ankur Singh, Jinqiu Yang, Tse-Hsun ChenAI Coding AgentsAI Agent Security

  25. The Ethics of Autonomous AI Agents for Offensive Security

    Jul 22, 2026Andreas Happe, Jürgen Cito, Jasmin WachterAI AccountabilityCybersecurity