Prompt Injection Attacks

Latest papers 70

All topics
CardsList
  1. Defending against Adaptive Prompt Injection Attacks via Reasoning-enabled Task Alignment

    Jun 13, 2026Lipeng He, Yihan Wang, Jiawen Zhang +1Adversarial TrainingAI Agent Safety

  2. AutoDojo: A Generative Benchmark for Evaluating Prompt Injection Defenses in LLM Agents

    Jun 13, 2026Xinhang Ma, Taoran Li, Chaowei Xiao +3AI Agent Security BenchmarksIndirect Prompt Injection

  3. Who Pays the Price? Stakeholder-Centric Prompt Injection Benchmarking for Real-world Web Agents

    Jun 11, 2026Zihao Wang, Yiming Li, Yutong Wu +8AI Agent EvaluationLLM Agent Security

  4. Does AI Reviewer See the Full Picture? Attacking and Defending Multimodal Peer Review

    Jun 10, 2026Xinyu Zhao, Rana Muhammad Shahroz Khan, Zhen Xu +2Adversarial Attacks on VLMsMultimodal Robustness

  5. Assessing Automated Prompt Injection Attacks in Agentic Environments

    Jun 9, 2026David Hofer, Edoardo Debenedetti, Florian TramèrAdversarial Prompt GenerationLLM Agent Security

  6. GitInject: Real-World Prompt Injection Attacks in AI-Powered CI/CD Pipelines

    Jun 7, 2026Jafar Isbarov, Umid Suleymanov, Ilia Shumailov +1AI Agent SecurityAI Agent Security Benchmarks

  7. VATS: Exploiting Implicit Authority in Error-Path Injection via Systematic Mutation

    Jun 6, 2026Harshil Patel, Kunal PaiLLM Red TeamingAI Agent Security

  8. POISE: Position-Aware Undetectable Skill Injection on LLM Agents

    Jun 6, 2026Haochang Hao, Dehai Min, Zhifang Zhang +4AI Agent SecurityPrompt Injection Attacks

  9. What If Prompt Injection Never Left? Exploring Cross-Session Stored Prompt Injection in Agentic Systems

    Jun 3, 2026Yuanbo Xie, Tianyun Liu, Yingjie Zhang +4Prompt InjectionAI Agent Security

  10. From Prompt Injection to Persistent Control: Defending Agentic Harness Against Trojan Backdoors

    May 29, 2026Jiejun Tan, Zhicheng Dou, Xinyu Yang +4Prompt InjectionLLM Agent Security

  11. Investigating Detection and Obfuscation of Prompt Injection Attacks Against Software Reverse Engineering AI Agents

    May 29, 2026Brian Crawford, Patrick McClureSoftware Reverse EngineeringIndirect Prompt Injection

  12. Automatically Attacking Software Reverse Engineering AI Agents

    May 28, 2026Brian Crawford, Justin Phillips, Patrick McClureSoftware EngineeringAdversarial Prompt Generation

  13. The Surface You Test Is Not the Surface That Breaks

    May 28, 2026Syed Nazmus Sakib, Nafiul Haque, Shahrear Bin Amin +1LLM Agent SecurityAI Agent Security Benchmarks

  14. Measuring Real-World Prompt Injection Attacks in LLM-based Resume Screening

    May 27, 2026Mohan Zhang, Yuqi Jia, Zhen Tan +4LLM SecurityIndirect Prompt Injection

  15. MIRAGE: Context-Aware Prompt Injection against Mobile GUI Agents via User-Generated Content

    May 27, 2026Ruoqi Guo, Yi Liu, Gelei Deng +7Mobile GUI AutomationGUI Agents

  16. Localization then Neutralization: Gradient-guided Token Suppression against Visual Prompt Injection Attack

    May 24, 2026Dongpeng Zhang, Ke Ma, Yangbangyan Jiang +4Gradient-Based AttributionVLM Robustness

  17. IterInject: Indirect Prompt Injection Against LLM Agents via Feedback-Guided Iterative Optimization

    May 23, 2026Zixuan Chen, Jiaxiang Chen, Li Luo +4LLM Agent SecurityIndirect Prompt Injection

  18. An Empirical Study of Privacy Leakage Chains via Prompt Injection in Black-Box Chatbot Environments

    May 18, 2026Hongjang Yang, Hyunsik Na, Daeseon ChoiData LeakagePrivacy Leakage in Language Models

  19. AI Agents May Always Fall for Prompt Injections

    May 17, 2026Sahar Abdelnabi, Eugene BagdasarianContextual IntegrityAI Agent Security

  20. ASPI: Seeking Ambiguity Clarification Amplifies Prompt Injection Vulnerability in LLM Agents

    May 17, 2026Udari Madhushani Sehwag, Zhengyang Shan, Heming Liu +3LLM Agent SecurityAI Agent Security Benchmarks

  21. WARD: Adversarially Robust Defense of Web Agents Against Prompt Injections

    May 14, 2026Tri Cao, Yulin Chen, Hieu Cao +8Adversarial TrainingPrompt Injection