AI Agent Security

Momentum

43 papers in the last four weeks, up 258% on the four weeks before. 0.4% of all new papers.

Jul 13Week of Sep 28

Latest papers 311

All topics
CardsList
  1. Secure-CUA: Controlling Untrusted Influence in Computer-Use Agents

    Oct 7, 2026Sarthak Choudhary, Mihai Christodorescu, Ashish Hooda +3Computer-Use AgentsAI Agent Security

  2. SwarmReconGuard: Black-Box Detection of Distributed Collective Reconnaissance by Individually Benign-Looking Agent Populations

    Oct 6, 2026Vahid Tavakkoli, Kabeh Mohsenzadegan, Kyandoghere KyamakyaAI Agent SecurityNetwork Intrusion Detection

  3. Multi-Aspect Runtime Verification for Simulation-Based V&V of LLM-Enabled Autonomous Agents

    Oct 6, 2026Nikolaos Kekatos, Dimitrios Nikou, Anastasios Temperekidis +4LLM Agent VerificationRuntime Enforcement for AI Agents

  4. Towards a Unified Misuse Monitoring Benchmark

    Oct 5, 2026Aniruddh Pramod, James Oldfield, Adel BibiAI Agent SecurityAI Agent Monitoring

  5. Runaway Reaction: When Benign Skills Compose into Malicious Behavior

    Oct 5, 2026Zunlong Zhou, Ziyuan Yang, Mengyu Sun +1AI Agent SecurityLLM Agent Security

  6. Agentic-ZTA: A Multi-Agent Architecture for Autonomous Zero Trust Enforcement

    Oct 5, 2026Shovan Roy, Lopamudra Praharaj, Maanak Gupta +1AI Agent SecurityMulti-Agent System Security

  7. Who Is Your Agent Serving? Provider-Side Indirect Prompt Injection in Proactive Agents

    Oct 4, 2026Rui Wang, Chao Wang, Xinchen Wang +3AI Agent SecurityIndirect Prompt Injection

  8. APEX: Active Protection at Execution Boundaries for LLM Agents

    Oct 3, 2026Xinran Zheng, Xin Fan Guo, Zhiqiang Hao +7AI Agent SecurityPrompt Injection Defense

  9. SoK: Decentralized Agent Economic Infrastructure

    Oct 1, 2026Rui Sun, Xihan Xiong, Qin Wang +5AI Agent AuditingAgentic Workflows

  10. AuraForge: Scaling Security Supervision for Training Coding Agents

    Oct 1, 2026Danqing Wang, Songwen Zhao, Harsh Sharma +4AI Coding AgentsSoftware Security

  11. Sapien: A Stateful Policy Engine for Autonomous AI Agents

    Sep 30, 2026Corinn Tiffany, Wen Zhang, Eugene Bagdasarian +1AI Agent SecurityRuntime Enforcement for AI Agents

  12. Pretext: Defeating Malicious Skill Detection Frameworks for AI Agents

    Sep 30, 2026Tobias Kaiser, Aritra DharAI Agent SecurityAgent Skill Security Auditing

  13. MiniRep: Robust Reputation-Based Aggregation for Multi-Agent Debate

    Sep 30, 2026Jiaming Zhang, Yuwan Liu, Yue Huang +1Multi-Agent DebateAI Agent Security

  14. AgBench: Agentic AI Benchmarks for Personal AI Devices

    Sep 29, 2026Yizhou Han, Di Wu, Dhananjay Saikumar +1AI Agent ReliabilityAI Agent Evaluation

  15. Evaluating Whether GPT-6 Astra Performs Unsanctioned Supply-Chain Attacks

    Sep 29, 2026Alexandra Souly, Kai Fronsdal, Abby D'Cruz +2LLM AuditingAI Agent Security

  16. Share-Borne AI Virus: Memory-Hopping Attacks Across LLM Agents

    Sep 28, 2026Sidharth Pulipaka, Ansh Sharma, Stanislau Hlebik +4AI Agent SecurityAgent Memory Poisoning

  17. SEAD: A State-Based Perspective on Attack and Defense in Tool-Using Agents

    Sep 28, 2026Xinjie Shen, Junran Wang, Rongzhe Wei +1AI Agent SecurityAI Agent Security Benchmarks

  18. Despite Instructions: Frontier Agents Improvise Covert Channels at Test Time

    Sep 26, 2026Jacob Dineen, Silei Ren, Muhao Chen +2AI Agent SecurityPrivacy Leakage in Language Models

  19. LLM Agents Can Easily Tamper With Their Own Traces

    Sep 24, 2026Jeremy Qin, David Schmotz, Derck Prinzhorn +3AI Agent AuditingAI Agent Security

  20. On the Effectiveness of Kernel-Level Evidence for Agent Security

    Sep 24, 2026Spencer King, Zhilu Zhang, Mikhail Kuznetsov +3AI Agent SecurityAI Agent Security Benchmarks

  21. Persistent Billable State: Denial-of-Wallet Attacks and Defenses in Tool-Calling LLM Agents

    Sep 23, 2026Jinqian Zhang, Haojun Xia, Shujiang Wu +4AI Agent SecurityLLM Agent Security