Prompt Injection Attacks on LLMs

LLM: Large Language Model

Latest papers 95

All topics
CardsList
  1. Send a SCOUT First: Pre-hoc Reasoning for Adaptive Detector Allocation in Prompt-Injection Defense

    May 29, 2026Shuhao Zhang, Jiarui Li, Qi Cao +2AI Agent Security BenchmarksLLM Security

  2. The Surface You Test Is Not the Surface That Breaks

    May 28, 2026Syed Nazmus Sakib, Nafiul Haque, Shahrear Bin Amin +1LLM Agent SecurityAI Agent Security Benchmarks

  3. Measuring Real-World Prompt Injection Attacks in LLM-based Resume Screening

    May 27, 2026Mohan Zhang, Yuqi Jia, Zhen Tan +4LLM SecurityIndirect Prompt Injection

  4. Prompt Injection Detection is Regime-Dependent: A Deployment-Aware Evaluation with Interpretable Structural Signals

    May 26, 2026Akindoyin Akinrele, Shreyank N GowdaLanguage Model Safety EvaluationPrompt Injection Attacks on LLMs

  5. IterInject: Indirect Prompt Injection Against LLM Agents via Feedback-Guided Iterative Optimization

    May 23, 2026Zixuan Chen, Jiaxiang Chen, Li Luo +4LLM Agent SecurityIndirect Prompt Injection

  6. LivePI: More Realistic Benchmarking of Agents Against Indirect Prompt Injection

    May 18, 2026Lei Zhao, Abhay Bhaskar, Edgar DobribanAI Agent SecurityAI Agent Safety

  7. AI Agents May Always Fall for Prompt Injections

    May 17, 2026Sahar Abdelnabi, Eugene BagdasarianContextual IntegrityAI Agent Security

  8. ASPI: Seeking Ambiguity Clarification Amplifies Prompt Injection Vulnerability in LLM Agents

    May 17, 2026Udari Madhushani Sehwag, Zhengyang Shan, Heming Liu +3LLM Agent SecurityAI Agent Security Benchmarks

  9. AgentShield: Deception-based Compromise Detection for Tool-using LLM Agents

    May 10, 2026Yassin H. Rassul, Tarik A. RashidLLM Agent SecurityAI Agent Monitoring

  10. LoopTrap: Termination Poisoning Attacks on LLM Agents

    May 7, 2026Huiyu Xu, Zhibo Wang, Wenhui Zhang +4Adversarial Attacks on LLMsLLM Red Teaming

  11. A Sentence Relation-Based Approach to Sanitizing Malicious Instructions

    May 1, 2026Soumil Datta, Melissa Umble, Daniel S. Brown +1Prompt Injection DefenseRAG Security

  12. CleanBase: Detecting Malicious Documents in RAG Knowledge Databases

    May 1, 2026Weifei Jin, Xilong Wang, Wei Zou +2RAG Poisoning AttacksRAG Security

  13. Evaluation of Prompt Injection Defenses in Large Language Models

    Apr 26, 2026Priyal Deep, Shane Emmons, Amy Fox +4Data LeakagePrivacy Leakage in Language Models

  14. RouteGuard: Internal-Signal Detection of Skill Poisoning in LLM Agents

    Apr 24, 2026Wenjie Xiao, Xuehai Tang, Biyu Zhou +2LLM Agent SecurityPrompt Injection Attacks on LLMs

  15. SafeAgent: A Runtime Protection Architecture for Agentic Systems

    Apr 19, 2026Hailin Liu, Eugene Ilyushin, Jie Ni +1Adversarial Attacks on LLMsAI Agent Security

  16. Prompt Injection as Role Confusion

    Feb 22, 2026Charles Ye, Jasmine Cui, Dylan Hadfield-MenellAdversarial Attacks on LLMsPrompt Injection Attacks