Prompt Injection Attacks on AI Agents

Latest papers 108

All topics
CardsList
  1. NetInjectBench: Benchmarking Indirect Prompt Injection in Tool-Using Large Language Model Agents for Network Operations

    Jul 11, 2026Ruksat Khan Shayoni, Muhammad Faraz Shoaib, S M Asif Hossain +1AI Agent Security BenchmarksIndirect Prompt Injection

  2. Agent Data Injection Attacks are Realistic Threats to AI Agents

    Jul 6, 2026Woohyuk Choi, Juhee Kim, Taehyun Kang +3AI Agent SecurityIndirect Prompt Injection

  3. DualView: Preventing Indirect Prompt Injection in Personal AI Agents

    Jul 4, 2026Juhee Kim, Woohyuk Choi, Taehyun Kang +2LLM Agent SecurityIndirect Prompt Injection

  4. Adaptive Evaluation of Out-of-Band Defenses Against Prompt Injection in LLM Agents

    Jun 25, 2026Praneeth Narisetty, Shiva Nagendra Babu Kore, Uday Kumar Reddy Kattamanchi +1LLM Agent EvaluationLLM Agent Security

  5. When AUC 0.998 Is Not Enough: A Candidate Evaluation Protocol for Hidden-State Probes of Indirect Prompt Injection in Multimodal Computer-Use Agents

    Jun 22, 2026Yanhang Li, Zhichao Fan, Zexin ZhuangAI Agent EvaluationIndirect Prompt Injection

  6. Analyzing Defensive Misdirection Against Model-Guided Automated Attacks on Agentic AI Systems

    Jun 18, 2026Reza Soosahabi, Vivek NamsaniAdversarial RobustnessJailbreak Attacks

  7. A Layered Security Framework Against Prompt Injection in RAG-Based Chatbots

    Jun 17, 2026Gulshan Saleem, Nisar Ahmed, Muhammad Imran Zaman +1LLM GuardrailsIndirect Prompt Injection

  8. SafeClawBench: Separating Semantic, Audit-Evidence, and Sandbox Harm in Tool-Using LLM Agents

    Jun 16, 2026Yuchuan Tian, Mengyu Zheng, Haocheng Mei +5LLM Safety BenchmarksLLM Agent Security

  9. Defending against Adaptive Prompt Injection Attacks via Reasoning-enabled Task Alignment

    Jun 13, 2026Lipeng He, Yihan Wang, Jiawen Zhang +1Adversarial TrainingAI Agent Safety

  10. AutoDojo: A Generative Benchmark for Evaluating Prompt Injection Defenses in LLM Agents

    Jun 13, 2026Xinhang Ma, Taoran Li, Chaowei Xiao +3AI Agent Security BenchmarksIndirect Prompt Injection

  11. Who Pays the Price? Stakeholder-Centric Prompt Injection Benchmarking for Real-world Web Agents

    Jun 11, 2026Zihao Wang, Yiming Li, Yutong Wu +8AI Agent EvaluationLLM Agent Security

  12. PI-Hunter: Automated Red-Teaming for Exposing and Localizing Prompt Injections

    Jun 10, 2026Pengfei He, Lesly Miculicich, Vishesh Sharma +5AI Agent AuditingLLM Auditing

  13. Smarter Saboteurs, Better Fixers: Scaling & Security in Linear Multi-Agent Workflows

    Jun 10, 2026Timothy McAllister, Sina Abdidizaji, Ivan Garibay +1AI Agent SecurityLLM Agent Security

  14. Assessing Automated Prompt Injection Attacks in Agentic Environments

    Jun 9, 2026David Hofer, Edoardo Debenedetti, Florian TramèrAdversarial Prompt GenerationLLM Agent Security

  15. GitInject: Real-World Prompt Injection Attacks in AI-Powered CI/CD Pipelines

    Jun 7, 2026Jafar Isbarov, Umid Suleymanov, Ilia Shumailov +1AI Agent SecurityAI Agent Security Benchmarks

  16. What If Prompt Injection Never Left? Exploring Cross-Session Stored Prompt Injection in Agentic Systems

    Jun 3, 2026Yuanbo Xie, Tianyun Liu, Yingjie Zhang +4Prompt InjectionAI Agent Security

  17. Send a SCOUT First: Pre-hoc Reasoning for Adaptive Detector Allocation in Prompt-Injection Defense

    May 29, 2026Shuhao Zhang, Jiarui Li, Qi Cao +2AI Agent Security BenchmarksLLM Security

  18. Investigating Detection and Obfuscation of Prompt Injection Attacks Against Software Reverse Engineering AI Agents

    May 29, 2026Brian Crawford, Patrick McClureSoftware Reverse EngineeringIndirect Prompt Injection

  19. The Surface You Test Is Not the Surface That Breaks

    May 28, 2026Syed Nazmus Sakib, Nafiul Haque, Shahrear Bin Amin +1LLM Agent SecurityAI Agent Security Benchmarks

  20. MIRAGE: Context-Aware Prompt Injection against Mobile GUI Agents via User-Generated Content

    May 27, 2026Ruoqi Guo, Yi Liu, Gelei Deng +7Mobile GUI AutomationGUI Agents

  21. IterInject: Indirect Prompt Injection Against LLM Agents via Feedback-Guided Iterative Optimization

    May 23, 2026Zixuan Chen, Jiaxiang Chen, Li Luo +4LLM Agent SecurityIndirect Prompt Injection

  22. LivePI: More Realistic Benchmarking of Agents Against Indirect Prompt Injection

    May 18, 2026Lei Zhao, Abhay Bhaskar, Edgar DobribanAI Agent SecurityAI Agent Safety

  23. AI Agents May Always Fall for Prompt Injections

    May 17, 2026Sahar Abdelnabi, Eugene BagdasarianContextual IntegrityAI Agent Security

  24. ASPI: Seeking Ambiguity Clarification Amplifies Prompt Injection Vulnerability in LLM Agents

    May 17, 2026Udari Madhushani Sehwag, Zhengyang Shan, Heming Liu +3LLM Agent SecurityAI Agent Security Benchmarks