AI Agent Security Benchmarks

Momentum

21 papers in the last four weeks, up 91% on the four weeks before. 0.2% of all new papers.

Jul 13Week of Sep 28

Latest papers 200

All topics
CardsList
  1. SkillSafetyBench: Evaluating Agent Safety under Skill-Facing Attack Surfaces

    May 12, 2026Chang Jin, An Wang, Zeming Wei +7LLM Agent SecurityAI Agent Security Benchmarks

  2. CTFusion: A CTF-based Benchmark for LLM Agent Evaluation

    May 12, 2026Dongjun Lee, Ga-eun Bae, Insu YunBenchmark ContaminationLLM Agent Evaluation

  3. ExploitGym: Can AI Agents Turn Security Vulnerabilities into Real Attacks?

    May 11, 2026Zhun Wang, Nico Schiller, Hongwei Li +13AI Agent SecurityAI Agent Security Benchmarks

  4. From Controlled to the Wild: Evaluation of Pentesting Agents for the Real-World

    May 11, 2026Pedro Conde, Henrique Branquinho, Valerio Mazzone +3Software Vulnerability DetectionAI Agent Evaluation

  5. LITMUS: Benchmarking Behavioral Jailbreaks of LLM Agents in Real OS Environments

    May 11, 2026Chiyu Zhang, Huiqin Yang, Bendong Jiang +8AI Agent Security BenchmarksLLM Jailbreak Attacks

  6. CrackMeBench: Binary Reverse Engineering for Agents

    May 11, 2026Isaac David, Arthur GervaisSoftware Reverse EngineeringLLM Agent Evaluation

  7. Red-Teaming Agent Execution Contexts: Open-World Security Evaluation on OpenClaw

    May 11, 2026Hongwei Yao, Yiming Liu, Yiling He +1AI Agent SecurityLLM Agent Security

  8. MonitoringBench: Semi-Automated Red-Teaming for Agent Monitoring

    May 10, 2026Monika Jotautaitė, Maria Angelica Martinez, Ollie Matthews +1Software Engineering AgentsAI Agent Security Benchmarks

  9. FORTIS: Benchmarking Over-Privilege in Agent Skills

    May 9, 2026Shawn Li, Chenxiao Yu, Han Wang +8LLM Agent SecurityAI Agent Security Benchmarks

  10. Agent Collectives Should Not Detect Their Own Imposters: A Chess Case Study

    May 9, 2026Alexandre Le Mercier, Chris Develder, Thomas DemeesterAI Agent SecurityAI Agent Security Benchmarks

  11. CyBiasBench: Benchmarking Bias in LLM Agents for Cyber-Attack Scenarios

    May 8, 2026Taein Lim, Seongyong Ju, Munhyeok Kim +2AI Agent Security BenchmarksLLMs for Cybersecurity

  12. DecodingTrust-Agent Platform (DTap): A Controllable and Interactive Red-Teaming Platform for AI Agents

    May 6, 2026Zhaorun Chen, Xun Liu, Haibo Tong +14AI Agent EvaluationAI Agent Security

  13. MOSAIC-Bench: Measuring Compositional Vulnerability Induction in Coding Agents

    May 5, 2026Jonathan Steinberg, Oren GalLanguage Model Safety EvaluationAI Coding Agents

  14. MCPHunt: An Evaluation Framework for Cross-Boundary Data Propagation in Multi-Server MCP Agents

    Apr 30, 2026Haonan Li, Tianjun Sun, Yongqing Wang +1MCP SecurityPrivacy Leakage in Language Models

  15. Autonomous LLM Agents & CTFs: A Second Look

    Apr 29, 2026Youness Bouchari, Matteo Boffa, Marco Mellia +3LLM Agent OrchestrationAI Agent Security Benchmarks

  16. Structured Security Auditing and Robustness Enhancement for Untrusted Agent Skills

    Apr 28, 2026Lijia Lv, Xuehai Tang, Jie Wen +2AI Agent AuditingAI Agent Security

  17. GAMMAF: A Common Framework for Graph-Based Anomaly Monitoring Benchmarking in LLM Multi-Agent Systems

    Apr 27, 2026Pablo Mateo-Torrejón, Alfonso Sánchez-MaciánMulti-Agent LLM SystemsAI Agent Security Benchmarks

  18. CI-Work: Benchmarking Contextual Integrity in Enterprise LLM Agents

    Apr 23, 2026Wenjie Fu, Xiaoting Qin, Jue Zhang +5Data LeakageContextual Integrity

  19. Do Agents Dream of Root Shells? Partial-Credit Evaluation of LLM Agents in Capture the Flag Challenges

    Apr 21, 2026Ali Al-Kaswan, Maksim Plotnikov, Maxim Hájek +3LLM Agent EvaluationAI Agent Security Benchmarks