AI Agent Security

Momentum

43 papers in the last four weeks, up 258% on the four weeks before. 0.4% of all new papers.

Jul 13Week of Sep 28

Latest papers 311

All topics
CardsList
  1. From Shield to Target: Denial-of-Service Attacks on LLM-Based Agent Guardrails

    Jun 12, 2026Yuguang Zhou, Xunguang Wang, Pingchuan Ma +3LLM GuardrailsAdversarial Attacks on LLMs

  2. A Security Analysis of Long-Horizon Agentic AI Systems: Threats, Evaluation, and Framework Development

    Jun 12, 2026Ahmed Mohammed Almalki, Mehedi MasudAI Agent Security

  3. Same-Origin Policy for Agentic Browsers

    Jun 12, 2026Xilong Wang, Xiaoxing Chen, Patrick Li +2Data LeakageAI Agent Security

  4. Minim: Privacy-Aware Minimal View for Agents via Trusted Local Sanitization

    Jun 11, 2026Hexuan Yu, Chaoyu Zhang, Heng Jin +4Data LeakageAI Agent Security

  5. SEVRA-BENCH: Social Engineering of Vulnerabilities in Review Agents

    Jun 11, 2026Rui Melo, Riccardo Fogliato, Sean Zhou +2Software Vulnerability DetectionAI Agent Security

  6. The Containment Gap: How Deployed Agentic AI Frameworks Fail Public-Facing Safety Requirements

    Jun 11, 2026Md Jafrin Hossain, Mohammad Arif Hossain, Weiqi Liu +1AI Agent SecurityAI Agent Safety

  7. PI-Hunter: Automated Red-Teaming for Exposing and Localizing Prompt Injections

    Jun 10, 2026Pengfei He, Lesly Miculicich, Vishesh Sharma +5AI Agent AuditingLLM Auditing

  8. Smarter Saboteurs, Better Fixers: Scaling & Security in Linear Multi-Agent Workflows

    Jun 10, 2026Timothy McAllister, Sina Abdidizaji, Ivan Garibay +1AI Agent SecurityLLM Agent Security

  9. Understanding and mitigating the risks of OpenClaw for non-technical users: A practical guide with Skill

    Jun 9, 2026Junchang Zheng, Junfeng Tan, Jialiang LinCybersecurityAI Agent Security

  10. RedAct: Redacting Agent Capability Traces for Procedural Skill Protection

    Jun 9, 2026Shuwen Xu, Zhitao He, Yi R. FungAI Agent SecurityAI Agent Security Benchmarks

  11. MIRAGE: A Polarity-Flipping Encoding Subspace in LLM Agents

    Jun 9, 2026Pratibha Revankar, Kargi Chauhan, Jihye Kim +3LLM InterpretabilityAI Agent Security

  12. GitInject: Real-World Prompt Injection Attacks in AI-Powered CI/CD Pipelines

    Jun 7, 2026Jafar Isbarov, Umid Suleymanov, Ilia Shumailov +1AI Agent SecurityAI Agent Security Benchmarks

  13. An AI Security Agent for University ACMIS: Multi-Vector Threat Detection and Automated Response

    Jun 6, 2026Joseph Walusimbi, Joshua Benjamin SsentongoAutonomous Cyber DefenseNetwork Intrusion Detection

  14. VATS: Exploiting Implicit Authority in Error-Path Injection via Systematic Mutation

    Jun 6, 2026Harshil Patel, Kunal PaiLLM Red TeamingAI Agent Security

  15. POISE: Position-Aware Undetectable Skill Injection on LLM Agents

    Jun 6, 2026Haochang Hao, Dehai Min, Zhifang Zhang +4AI Agent SecurityPrompt Injection Attacks

  16. TRACE: Trajectory Reasoning through Adaptive Cross-Step Evidence Aggregation for LLM Agents

    Jun 5, 2026Vijitha Mittapalli, Shreyaa Jayant Dani, Satya Srujana Pilli +7Long-Horizon Agent EvaluationAI Agent Security

  17. Beyond Similarity: Trustworthy Memory Search for Personal AI Agents

    Jun 4, 2026Jiawen Zhang, Kejia Chen, Jiachen Ma +7LLM Agent MemoryAI Agent Security

  18. Data Flow Control: Data Safety Policies for AI Agents

    Jun 4, 2026Charlie Summers, Eugene WuAI Agent SecurityInformation Flow Control

  19. CyberGym-E2E: Scalable Real-World Benchmark for AI Agents' End-to-End Cybersecurity Capabilities

    Jun 3, 2026Tianneng Shi, Robin Rheem, Dongwei Jiang +13Automated Program RepairSoftware Vulnerability Detection

  20. What If Prompt Injection Never Left? Exploring Cross-Session Stored Prompt Injection in Agentic Systems

    Jun 3, 2026Yuanbo Xie, Tianyun Liu, Yingjie Zhang +4Prompt InjectionAI Agent Security

  21. Agent libOS: A Runtime Substrate for Capability-Controlled Self-Evolving LLM Agents

    Jun 2, 2026Yingqi ZhangAI Agent SecurityLLM Agents