AI Agent Security

Momentum

43 papers in the last four weeks, up 258% on the four weeks before. 0.4% of all new papers.

Jul 13Week of Sep 28

Latest papers 311

All topics
CardsList
  1. OpenAgenet / OAN Yellow Paper: Technical Architecture for Trust-Governed Resource Identity and Discovery

    Jun 2, 2026Jinliang XuAI Agent SecurityAI Agent Governance

  2. OpenAgenet / OAN White Paper: Open Infrastructure for Trusted Agent Interconnection

    Jun 2, 2026Jinliang XuAI Agent SecurityAI Agent Governance

  3. SkillHarm: Lifecycle-Aware Skill-Based Attacks via Automated Construction

    Jun 1, 2026Yuting Ning, Zhehao Zhang, Yash Kumar Lal +8AI Agent SecurityAI Agent Security Benchmarks

  4. ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree

    May 31, 2026Vincent Koc, Patrick Erichsen, Jacob Tomlinson +3AI Agent SecurityAgent Skill Security Auditing

  5. Benchmarking Security Risk Detection and Verification in Open Agentic Skill Ecosystems

    May 30, 2026Ismail Hossain, Sai Puppala, Zhuoran Lu +2AI Agent SecurityAI Agent Security Benchmarks

  6. When Safe Skills Collide: Measuring Compositional Risk in Agent Skill Ecosystems

    May 30, 2026Su Wang, Pin Qian, Yihang Chen +6Software SecurityAI Agent Security

  7. VisualLeakBench: Reproducible Action-Boundary Propagation Failures in Vision-Language Agents

    May 29, 2026Youting Wang, Yuan Tang, Yitian Qian +1Data LeakageTool Use in VLMs

  8. Provably Secure Agent Guardrail

    May 28, 2026Benlong Wu, Weiming Zhang, Kejiang Chen +2LLM GuardrailsAI Agent Security

  9. The Importance of Out-of-Band Metadata for Safe Autonomous Agents: The Redpanda Agentic Data Plane

    May 27, 2026Tyler Akidau, Tyler Rockwood, Johannes Brüderl +1AI Agent SecurityAI Agent Safety

  10. AIRGuard: Guarding Agent Actions with Runtime Authority Control

    May 27, 2026Suliu Qin, Haomin Zhuang, Yujun Zhou +2AI Agent SecurityTool-Using Agents

  11. Technical Report: Exploring the Emerging Threats of the Agent Skill Ecosystem

    May 27, 2026Luca Beurer-Kellner, Aleksei Kudrinskii, Marco Milanta +3AI Agent SecurityMalware Analysis

  12. Out of Sight, Not Out of Mind: Unveiling Latent Attack in Latent-based Multi-Agent Systems

    May 27, 2026Chenxi Wang, Ruiyang Huang, Jiayan Sun +2Multi-Agent LLM SystemsAdversarial Attacks

  13. SNARE: Adaptive Scenario Synthesis for Eliciting Overeager Behavior in Coding Agents

    May 27, 2026Yubin Qu, Yi Liu, Gelei Deng +4Coding AgentsAI Agent Security

  14. MRMMIA: Membership Inference Attacks on Memory in Chat Agents

    May 27, 2026Kai Chen, Yan Pang, Tianhao WangData LeakageMembership Inference Attacks

  15. HARP: Measuring Harm Amplification in Multi-Agent LLM Systems

    May 26, 2026Md Hafizur Rahman, Zafaryab Haider, Tanzim Mahfuz +1Multi-Agent LLM SystemsAdversarial Attacks on LLMs

  16. Lessons from Penetration Tests on Large-Scale Agent Systems

    May 26, 2026Kevin Eykholt, Dhilung Kirat, Xiaokui Shu +3Software SecurityAI Agent Security

  17. When Agents Control Robots: A Zero Trust Policy Model for Agentic Cyber-Physical Systems

    May 25, 2026Tharindu Ranathunga, Kavishka Fernando, Susan ReaRobotic ControlCyber-Physical Systems

  18. Security of OpenClaw Agents: Fundamentals, Attacks, and Countermeasures

    May 25, 2026Yuntao Wang, Jianle Ba, Han Liu +5AI Agent SecurityLLM Agent Security

  19. Measuring Security Without Fooling Ourselves: Why Benchmarking Agents Is Hard

    May 21, 2026Sahar Abdelnabi, Chris Hicks, Konrad Rieck +1AI Agent EvaluationAI Agent Security

  20. Quality and Security Signals in AI-Generated Python Refactoring Pull Requests

    May 20, 2026Mohamed Almukhtar, Anwar Ghammam, Hua MingAI Coding AgentsAI Agent Security