AI Agent Security

Momentum

43 papers in the last four weeks, up 258% on the four weeks before. 0.4% of all new papers.

Jul 13Week of Sep 28

Latest papers 311

All topics
CardsList
  1. A Comparative Evaluation of AI Agent Security Guardrails

    Apr 27, 2026Qi Li, Jiu Li, Pingtao Wei +8LLM GuardrailsAI Agent Security

  2. Poster: ClawdGo: Endogenous Security Awareness Training for Autonomous AI Agents

    Apr 27, 2026Jiaqi Li, Yang Zhao, Bin Sun +3AI Agent SecurityLLM Agent Training

  3. AI Identity: Standards, Gaps, and Research Directions for AI Agents

    Apr 25, 2026Takumi Otsuka, Kentaroh Toyoda, Alex LeungAI AccountabilityAI Agent Security

  4. Semantic Denial of Service in LLM-controlled robots

    Apr 25, 2026Jonathan Steinberg, Oren GalLanguage Model Safety EvaluationAdversarial Attacks on LLMs

  5. UNSEEN: A Cross-Stack LLM Unlearning Defense against AR-LLM Social Engineering Attacks

    Apr 25, 2026Tianlong Yu, Yang Yang, Xiao Luo +6Privacy-Preserving Language ModelsLLM Guardrails

  6. Breaking MCP with Function Hijacking Attacks: Novel Threats for Function Calling and Agentic Models

    Apr 22, 2026Yannis Belkhiter, Giulio Zizzo, Sergio Maffeis +2Adversarial Attacks on LLMsAI Agent Security

  7. AgentSOC: A Multi-Layer Agentic AI Framework for Security Operations Automation

    Apr 22, 2026Joyjit Roy, Samaresh Kumar SinghAutonomous Cyber DefenseCybersecurity

  8. An AI Agent Execution Environment to Safeguard User Data

    Apr 21, 2026Robert Stanley, Avi Verma, Lillian Tsai +2AI Agent SecurityPrivacy Leakage in Language Models

  9. If you're waiting for a sign... that might not be it! Mitigating Trust Boundary Confusion from Visual Injections on Vision-Language Agentic Systems

    Apr 21, 2026Jiamin Chang, Minhui Xue, Ruoxi Sun +3VLM RobustnessAdversarial Attacks on VLMs

  10. How Adversarial Environments Mislead Agentic AI?

    Apr 20, 2026Zhonghao Zhan, Huichi Zhou, Zhenhao Li +3Adversarial RobustnessAI Agent Evaluation

  11. Owner-Harm: A Missing Threat Model for AI Agent Safety

    Apr 20, 2026Dongcheng Zhang, Yiqing JiangAI Agent SecurityAI Agent Safety

  12. From Craft to Kernel: A Governance-First Execution Architecture and Semantic ISA for Agentic Computers

    Apr 20, 2026Xiangyu Wen, Yuang Zhao, Xiaoyu Xu +9AI Agent SecurityAI Agent Safety

  13. SafeAgent: A Runtime Protection Architecture for Agentic Systems

    Apr 19, 2026Hailin Liu, Eugene Ilyushin, Jie Ni +1Adversarial Attacks on LLMsAI Agent Security

  14. CapSeal: Capability-Sealed Secret Mediation for Secure Agent Execution

    Apr 18, 2026Shutong Jin, Ruiyi Guo, Ray C. C. CheungAI Agent SecurityLLM Agent Security

  15. Hardening x402: PII-Safe Agentic Payments via Pre-Execution Metadata Filtering

    Apr 13, 2026Vladimir StantchevAI Agent SecuritySafety Filtering

  16. Are GUI Agents Focused Enough? Automated Distraction via Semantic-level UI Element Injection

    Apr 9, 2026Wenkui Yang, Chao Jin, Haisu Zhu +7GUI AgentsAI Agent Security

  17. Auditable Agents

    Apr 7, 2026Yi Nian, Aojie Yuan, Haiyue Zhang +7AI AccountabilityAI Agent Auditing

  18. T-MAP: Red-Teaming LLM Agents with Trajectory-aware Evolutionary Search

    Mar 21, 2026Hyomin Lee, Sangwoo Park, Yumin Choi +3Adversarial Prompt GenerationLLM Red Teaming