AI Agent Security Benchmarks

Momentum

21 papers in the last four weeks, up 91% on the four weeks before. 0.2% of all new papers.

Jul 13Week of Sep 28

Latest papers 200

All topics
CardsList
  1. MemPoison: Uncovering Persistent Memory Threats and Structural Blind Spots in LLM Agents

    Jul 16, 2026Jifeng Gao, Kang Xia, Yi Zhang +5LLM Agent MemoryLLM Agent Security

  2. NetInjectBench: Benchmarking Indirect Prompt Injection in Tool-Using Large Language Model Agents for Network Operations

    Jul 11, 2026Ruksat Khan Shayoni, Muhammad Faraz Shoaib, S M Asif Hossain +1AI Agent Security BenchmarksIndirect Prompt Injection

  3. PiSAs: Benchmarking Contextual Integrity in Multi-User Agentic Systems

    Jul 6, 2026Shubham Gupta, Nazanin Mohammadi Sepahvand, Abhinav Kumar +6Multi-Agent LLM SystemsContextual Integrity

  4. When Claws Remember but Do Not Tell: Stealthy Memory Injection in Persistent Personal Agents

    Jul 6, 2026Yechao Zhang, Shiqian Zhao, Jiawen Zhang +5AI Agent Security BenchmarksAgent Memory Poisoning

  5. CONTRA: Red-Teaming Configurations of Personalizable Agents

    Jul 3, 2026Jonathan Nöther, Adish Singla, Goran RadanovicLLM Red TeamingAI Agent Security

  6. Distributed Attacks in Persistent-State AI Control

    Jul 2, 2026Josh Hills, Ida Caspary, Asa Cooper SticklandAI Agent SecurityAI Agent Security Benchmarks

  7. Understanding and Evaluating Claw-like Agent Security Through a Computer-Systems Lens

    Jun 29, 2026Peizhi Niu, Wenjie Qu, Shangding Gu +14AI Agent SecurityLLM Agent Security

  8. MIRROR: Novelty-Constrained Memory-Guided MCTS Red-Teaming for Agentic RAG

    Jun 25, 2026Inderjeet Singh, Andrés Murillo, Motoyoshi Sekiya +2Multimodal RAGTool-Augmented Language Model Agents

  9. CyberChainBench: Can AI Agents Secure Smart Contracts Against Real-World On-Chain Vulnerabilities?

    Jun 24, 2026Jintao Huang, Fengqing Jiang, Radha Poovendran +1Automated Program RepairSoftware Vulnerability Detection

  10. AI Snitches Get Glitches: Towards Evading Agentic Surveillance

    Jun 24, 2026Hyejun Jeong, Dzung Pham, Amir Houmansadr +1AI Agent SecurityAI Agent Security Benchmarks

  11. Decoupling Reconnaissance and Exploitation: Measuring the Capability Boundaries of LLM-Based Web Penetration Testing

    Jun 24, 2026Liwei Yu, Shuo Li, Ming Zhou +2LLM Agent EvaluationAI Agent Security Benchmarks

  12. RIFT-Bench: Dynamic Red-teaming For Agentic AI Systems

    Jun 22, 2026Yarin Yerushalmi Levi, Roy Betser, Amit Giloni +5AI Agent SecurityAI Agent Security Benchmarks

  13. Local LLM Agents as Vulnerable Runtimes:A Source-Code Audit of the Agent Runtime Layer

    Jun 19, 2026Zhengsong Zhang, Zongze Li, Jiawei Guo +1AI Agent AuditingSoftware Vulnerability Detection

  14. Honeyquest for LLMs: Rethinking Cyber Deception for AI Attackers

    Jun 19, 2026Kerri Prinos, Lilianne Brush, Cameron DentonCybersecurityAI Agent Security

  15. When Lower Privileges Suffice: Investigating Over-Privileged Tool Selection in LLM Agents

    Jun 18, 2026Kaiyue Yang, Yuyan Bu, Jingwei Yi +5LLM Agent SecurityAI Agent Security Benchmarks

  16. GateMem: Benchmarking Memory Governance in Multi-Principal Shared-Memory Agents

    Jun 17, 2026Zhe Ren, Yibo Yang, Yimeng Chen +7LLM Agent MemoryAI Agent Security

  17. SafeClawBench: Separating Semantic, Audit-Evidence, and Sandbox Harm in Tool-Using LLM Agents

    Jun 16, 2026Yuchuan Tian, Mengyu Zheng, Haocheng Mei +5LLM Safety BenchmarksLLM Agent Security

  18. SkillVetBench: LLM-as-Judge for Multi-Dimensional Security Risk Evaluation in Open-Source LLM Agent Skills

    Jun 14, 2026Ismail Hossain, Sai Puppala, Md Jahangir Alam +2LLM-as-a-JudgeLLM Agent Security

  19. Benign in Isolation, Harmful in Composition: Security Risks in Agent Skill Ecosystems

    Jun 13, 2026Yi Xie, Jiawei Du, Yu Cheng +2LLM Agent SecurityAI Agent Security Benchmarks