AI Agent Security

Momentum

43 papers in the last four weeks, up 258% on the four weeks before. 0.4% of all new papers.

Jul 13Week of Sep 28

Latest papers 311

All topics
CardsList
  1. A2M: Trace-Optimized Agent Hijacking in the MCP Ecosystem

    Sep 22, 2026Laizhen Li, Xuan Wang, Peicheng Zhao +4MCP SecurityAI Agent Security

  2. Indirect tipping: a social attack surface in AI agent populations

    Sep 21, 2026Ariel Flint, Luca Maria Aiello, Sara M. Constantino +2Multi-Agent CoordinationAI Agent Security

  3. ClashBench: Conflicts Leading Agents to Seize and Harm

    Sep 17, 2026Yuejin Xie, Yu Li, Dadi Guo +6AI Agent SecurityAI Agent Safety

  4. Red-Teaming Auto Mode: Improving Blocking Classifiers Against Malign Coding Agents

    Sep 17, 2026Alex Remedios, Simon Storf, Fabien Roger +1AI Agent SecurityAI Agent Monitoring

  5. Reputation as Community Memory for the Agentic Web

    Sep 16, 2026Ryan Chard, Gus Ellerm, Alexander Brace +4Agent MemoryAI Agent Security

  6. Market Signal Injection: Adversarial Context Manipulation of LLM Pricing Agents

    Sep 16, 2026Dohun Lee, Hyunwoo ParkAdversarial Attacks on LLMsAI Agent Security

  7. Autonomy in Check: Governor-Mediated Adaptive Security at the Edge

    Sep 16, 2026Ijaz Ahmad, Ijaz Ahmad, Flavio Esposito +1Autonomous Cyber DefenseAI Agent Security

  8. Reflections on Trusting Trust, Revisited: Contaminating Self-Modifying AI Coding Agents with Poisoned Benchmarks

    Sep 15, 2026Franziska Roesner, Tadayoshi KohnoAI Coding AgentsAI Agent Security

  9. Vulnerability Localization Benchmark: Measuring Agentic Security Analysis at Repository Scale

    Sep 14, 2026Aman Priyanshu, Supriti Vijay, Kimia Majd +8Software Vulnerability DetectionAI Agent Security

  10. Universal Defenses for Tool-Integrated LLM Agents Against Adversarial Attacks

    Sep 14, 2026Xiaoyan Li, Yunli WangAdversarial Attacks on LLMsBackdoor Defense in LLMs

  11. SoK: Rethinking Jailbreaking in the Era of Agentic AI: Attacks, Defenses, and Practical Consideration

    Sep 14, 2026Md Jueal Mia, Yanzhao Wu, Selcuk Uluagac +1Language Model Safety EvaluationJailbreak Robustness

  12. An Experimental Evaluation of Multimodal Prompt Injection Attacks on Agentic AI Frameworks

    Sep 10, 2026Viet K. Nguyen, Mohammad I. HusainAI Agent SecurityAI Agent Security Benchmarks

  13. From Evidence to Effect: Authority Semantics and Runtime Infrastructure for Stateful Agents

    Sep 8, 2026Yang Li, Zongsi Xu, Sergey Volkov +7AI Agent SecurityTool Access Control for LLM Agents

  14. Do AI Coding Assistants Check Before They Install? A Pre-Registered Demand-Side Audit of Trust Signals in the Research Software Supply Chain

    Sep 7, 2026Pengyin ShanAlgorithmic AuditingAI Coding Agents

  15. CoER: Defending against Adaptive Indirect Prompt Injection via Adversarial Co-Evolution and Refinement

    Sep 7, 2026Boyang Zhang, Qingxin Xiao, Lingwei Dang +1Reinforcement LearningAI Agent Security

  16. Rethinking Indirect Prompt Injection as a Test-Time Search Problem

    Sep 7, 2026Duong M. Nguyen, Joon Sik Kim, Blazej Manczak +1Inference-Time SearchAI Agent Security

  17. A Blind Trust, the Bloody Thrust: When Attacker-Controlled Hook Updates Steer AI Agent Harnesses towards Malicious Behaviors

    Sep 3, 2026Pengxun Li, Litian Zhang, Jianwei Hou +4AI Agent SecurityLLM Agent Harnesses

  18. Defense-as-Skill: Evolving Runtime Guard Skill for Skill-Augmented Agents

    Sep 1, 2026Xiaofang Yang, Ziqi Miao, Dianbo Sui +2AI Agent SecurityAI Agent Security Benchmarks

  19. EvoSkill Injection: Red-Teaming Autonomous Skill Generation and Evolution in Self-Evolving Agents

    Aug 31, 2026Doyun Kim, Chanwoo Kim, Sugyeong Eo +2AI Agent SecurityAI Agent Security Benchmarks

  20. InterSAGE: The Secure and Verifiable Interoperability Protocol for An Internet of Agents

    Aug 13, 2026Zhenhua Zou, Sheng Guo, Qiuyang Zhan +3AI AccountabilityAI Agent Security

  21. PIPES: Securing Agent Perception with Provenance and Priors

    Aug 13, 2026Sanjay Kariyappa, Severin Klingler, G. Edward SuhData ProvenanceAI Agent Security

  22. Rethinking Agent Security as a Networking Problem

    Aug 12, 2026Van Tran, Taveesh Sharma, Tajveer Singh Dhesi +1CybersecurityAI Agent Security