Prompt Injection Detection

Latest papers 39

All topics
CardsList
  1. RAG-PIBench: A Leakage-Aware Benchmark for Prompt-Injection Detection in Trustworthy RAG Systems

    Oct 6, 2026Niveen O. Jaffal, Ahmet Yuksel, David MohaisenRAG SecurityPrompt Injection Attacks on LLMs

  2. PIDS-Bench: Evaluating Prompt-Injection Detectors Under Over-Defense, Obfuscation, and Distribution Shift

    Sep 14, 2026Yusuf Khalid Shire, Sang-Chul KimDistribution Shift RobustnessPrompt Injection Detection

  3. DriftNet: A Dual-Head Trajectory Transformer for Detecting and Localizing Prompt Injection in LLM Agents

    Sep 12, 2026Asif Pinjari, Mithun Paul Saint-GermainAI Agent Security BenchmarksAI Agent Monitoring

  4. No-Box Vulnerability Analysis: Description-only Detection of Indirect Prompt Injection Vulnerabilities in MCP Servers

    Sep 9, 2026Zehua Zhang, Jie Hu, Pratham Hegde +13MCP SecuritySoftware Vulnerability Detection

  5. BASIS: Breach-Aware Selective Prompt Injection Shielding with Prefill Attention Probes

    Aug 8, 2026Laiqiao Qin, Tianqing Zhu, Longxiang Gao +1Prompt Injection DefensePrompt Injection Attacks on LLMs

  6. Robust Context-Aware Detection of Malicious Instructions in Text

    Aug 5, 2026Buzhao Liu, Xinhang Ma, Yevgeniy VorobeychikAdversarial TrainingPrompt Injection Defense

  7. Your Agentic LLMs Secretly Encode Latent Signals of Indirect Prompt-Injection Exposure

    Aug 1, 2026Jianshuo Dong, Yiming Liu, Maosen Zhang +6Indirect Prompt InjectionPrompt Injection Defense

  8. When AUC 0.998 Is Not Enough: A Candidate Evaluation Protocol for Hidden-State Probes of Indirect Prompt Injection in Multimodal Computer-Use Agents

    Jun 22, 2026Yanhang Li, Zhichao Fan, Zexin ZhuangAI Agent EvaluationIndirect Prompt Injection

  9. Confidently Wrong: Severity-Aware Calibration of Prompt-Injection Detectors under Attack Shift

    Jun 21, 2026Md Anas BiswasIndirect Prompt InjectionModel Calibration

  10. A Layered Security Framework Against Prompt Injection in RAG-Based Chatbots

    Jun 17, 2026Gulshan Saleem, Nisar Ahmed, Muhammad Imran Zaman +1LLM GuardrailsIndirect Prompt Injection

  11. GuardNet: Ensemble Strategies of Shallow Neural Networks for Robust Prompt Injection and Jailbreak Detection

    Jun 4, 2026Paulo Ricardo Ferreira Neves, Edson Rodrigues da Cruz Filho, Paulo Henrique Eleuterio Falsetti +7LLM GuardrailsNeural Network Robustness

  12. Send a SCOUT First: Pre-hoc Reasoning for Adaptive Detector Allocation in Prompt-Injection Defense

    May 29, 2026Shuhao Zhang, Jiarui Li, Qi Cao +2AI Agent Security BenchmarksLLM Security

  13. Investigating Detection and Obfuscation of Prompt Injection Attacks Against Software Reverse Engineering AI Agents

    May 29, 2026Brian Crawford, Patrick McClureSoftware Reverse EngineeringIndirect Prompt Injection

  14. Measuring Real-World Prompt Injection Attacks in LLM-based Resume Screening

    May 27, 2026Mohan Zhang, Yuqi Jia, Zhen Tan +4LLM SecurityIndirect Prompt Injection

  15. Prompt Injection Detection is Regime-Dependent: A Deployment-Aware Evaluation with Interpretable Structural Signals

    May 26, 2026Akindoyin Akinrele, Shreyank N GowdaLanguage Model Safety EvaluationPrompt Injection Attacks on LLMs

  16. When Prompts Become Payloads: A Framework for Mitigating SQL Injection Attacks in Large Language Model-Driven Applications

    May 11, 2026Farzad Nourmohammadzadeh Motlagh, Mehrdad Hajizadeh, Mehryar Majd +3LLM SecurityText-to-SQL

  17. AgentShield: Deception-based Compromise Detection for Tool-using LLM Agents

    May 10, 2026Yassin H. Rassul, Tarik A. RashidLLM Agent SecurityAI Agent Monitoring

  18. A Sentence Relation-Based Approach to Sanitizing Malicious Instructions

    May 1, 2026Soumil Datta, Melissa Umble, Daniel S. Brown +1Prompt Injection DefenseRAG Security

  19. CleanBase: Detecting Malicious Documents in RAG Knowledge Databases

    May 1, 2026Weifei Jin, Xilong Wang, Wei Zou +2RAG Poisoning AttacksRAG Security