AI Agent Security

Momentum

43 papers in the last four weeks, up 258% on the four weeks before. 0.4% of all new papers.

Jul 13Week of Sep 28

Latest papers 311

All topics
CardsList
  1. Position: Academic Conferences are Potentially Facing Denominator Gaming Caused by Fully Automated Scientific Agents

    May 11, 2026Rong Shan, Te Gao, Hang Zheng +6Scholarly Peer ReviewAI Agent Security

  2. Skill Description Deception Attack against Task Routing in Internet of Agents

    May 11, 2026Jiayi He, Xiaofeng Luo, Jiawen Kang +3Adversarial AttacksAI Agent Security

  3. Oracle Poisoning: Corrupting Knowledge Graphs to Weaponise AI Agent Reasoning

    May 10, 2026Ben Kereopa-Yorke, Guillermo Diaz, Holly Wright +3AI Agent SecurityLLM Agent Security

  4. Restricting the Model, Missing the System: Measurement and Accountability in Offensive AI Governance

    May 10, 2026Michael Alexander Riegler, Finn Schwall, Annika Willoch Olstad +4AI Agent SecurityAI Safety Evaluation

  5. Don't Click That: Teaching Web Agents to Resist Deceptive Interfaces

    May 10, 2026Yilin Zhang, Yingkai Hua, Chunyu Wei +2Adversarial RobustnessAI Agent Security

  6. The Authorization-Execution Gap Is a Major Safety and Security Problem in Open-World Agents

    May 10, 2026Baoyuan Wu, Qingshan Liu, Adel Bibi +2Agent Failure AnalysisAI Agent Security

  7. Agent Collectives Should Not Detect Their Own Imposters: A Chess Case Study

    May 9, 2026Alexandre Le Mercier, Chris Develder, Thomas DemeesterAI Agent SecurityAI Agent Security Benchmarks

  8. PAAC: Privacy-Aware Agentic Device-Cloud Collaboration

    May 9, 2026Liangqi Yuan, Wenzhi Fang, Shiqiang Wang +1Privacy-Preserving Language ModelsAI Agent Security

  9. OrchJail: Jailbreaking Tool-Calling Text-to-Image Agents by Orchestration-Guided Fuzzing

    May 8, 2026Jianming Chen, Yawen Wang, Junjie Wang +3Adversarial Prompt GenerationT2I Generation

  10. Securing Computer-Use Agents: A Unified Architecture-Lifecycle Framework for Deployment-Grounded Reliability

    May 8, 2026Zejian Chen, Zhanyuan Liu, Chaozhuo Li +6Agent ReliabilityComputer-Use Agents

  11. TeamBench: Evaluating Agent Coordination under Enforced Role Separation

    May 8, 2026Yubin Kim, Chanwoo Park, Taehan Kim +9Multi-Agent CoordinationLLM Agent Evaluation

  12. Securing the Agent: Vendor-Neutral, Multitenant Enterprise Retrieval and Tool Use

    May 6, 2026Francisco Javier Arceo, Varsha Prasad NarsingLLM Agent OrchestrationAI Agent Security

  13. DecodingTrust-Agent Platform (DTap): A Controllable and Interactive Red-Teaming Platform for AI Agents

    May 6, 2026Zhaorun Chen, Xun Liu, Haibo Tong +14AI Agent EvaluationAI Agent Security

  14. Redefining AI Red Teaming in the Agentic Era: From Weeks to Hours

    May 5, 2026Raja Sekhar Rao Dheekonda, Will Pearce, Nick LandersAdversarial Attacks on LLMsLLM Red Teaming

  15. TRAP: Tail-aware Ranking Attack for World-Model Planning

    May 3, 2026Siyuan Duan, Ke Zhang, Xizhao LuoBackdoor AttacksWorld Model-Based Planning

  16. A Low-Latency Fraud Detection Layer for Detecting Adversarial Interaction Patterns in LLM-Powered Agents

    May 1, 2026Sheldon Yu, Yingcheng Sun, Hanqing Guo +1AI Agent SecurityLLM Agent Security

  17. AgentReputation: A Decentralized Agentic AI Reputation Framework

    Apr 30, 2026Mohd Sameen Chishti, Damilare Peter Oyinloye, Jingyue LiAI AccountabilityAI Agent Security

  18. Security Attack and Defense Strategies for Autonomous Agent Frameworks: A Layered Review with OpenClaw as a Case Study

    Apr 30, 2026Luyao Xu, Xiang ChenAI Agent SecurityLLM Agent Security

  19. Ambient Persuasion in a Deployed AI Agent: Unauthorized Escalation Following Routine Non-Adversarial Content Exposure

    Apr 29, 2026Diego F. Cuadros, Abdoul-Aziz MaigaAgent Failure AnalysisAI Agent Security

  20. Agent Name Service (ANS): A Proof-of-Concept Trust Layer for Secure AI Agent Discovery, Identity, and Governance in Kubernetes

    Apr 29, 2026Akshay Mittal, Elyson De La CruzAI Agent SecurityAI Agent Governance

  21. Structured Security Auditing and Robustness Enhancement for Untrusted Agent Skills

    Apr 28, 2026Lijia Lv, Xuehai Tang, Jie Wen +2AI Agent AuditingAI Agent Security