Realistic Threat Model

Momentum

3 papers in the last four weeks, against 1 the four weeks before. 0.0% of all new papers.

Jul 13Week of Sep 28

Latest papers 63

All topics
CardsList
  1. Distillation Defenses Easily Break After Reinforcement Learning

    Sep 28, 2026Shidan Javaheri, Alexander Panfilov, Oliver Britton +2Probe-Logit DistillationLLM Defense Mechanisms

  2. On Identifying Adversarial Intent Injection in AI-Native 6G Networks

    Sep 14, 2026Nilesh Chakraborty, Petar Djukic, Burak KantarciIntrusion DetectionRealistic Threat Model

  3. Protocol effects on feature-based hardware-Trojan detection across Trust-Hub families

    Sep 7, 2026Hang Xiao, Chuhong Xu, Kainan Zhou +2MalwareGating

  4. Mind the Gap: Robustness Risks in PII Detection Systems

    Sep 3, 2026Adeel Zafar, Slawomir NowaczykRealistic Threat ModelNamed-Entity Recognition

  5. Fairis: Fairness-Aware Aggregation with Provable Influence Containment against Fairness Poisoning Attacks in Collaborative Machine Learning

    Aug 6, 2026Devharsh Trivedi, Nesrine Kaaniche, Nikos Triandopoulos +2Algorithmic FairnessFederated Learning

  6. Breadcrumbing Search Agents

    Aug 5, 2026Xuebin Li, Hanqing Zhao, Siyuan Liang +4Search AgentsInstruction Injection Attacks

  7. ThreatForest: Multi-Agent Attack Tree Generation with Pluggable TTP Framework Mapping

    Jul 29, 2026Cristian Leo, Anton Dykyi, Danny Cortegaca +2Realistic Threat ModelThreat

  8. Agentic Permissions Policy Algebra for Taint Confinement in LLM Agents

    Jul 27, 2026Arseny Kravchenko, Vadim Liventsev, Innokentii Konstantinov +2Realistic Threat ModelInformation-Flow Control

  9. The Chronos Vulnerability: A Taxonomy of Temporal Persistence and Memory-Based Deception in Agentic AI

    Jul 20, 2026Om Narayan, Ramkinker Singh, Praveen BaskarRealistic Threat ModelLLM Defense Mechanisms

  10. An Early Warning of Emerging Biosecurity Risks in Frontier LLMs

    Jul 20, 2026Zhida He, Xia Hu, Baichen Le +20Large Language Model SafetyLife Sciences Research

  11. Signal-based Model Access Risk Analysis for AI System Operations Security

    Jul 17, 2026Maria Mahbub, Steven Young, Amir Sadovnik +4Artificial Intelligence RiskRealistic Threat Model

  12. Rethinking Penetration Testing for AI-Enabled Systems: From Resource Compromise to Behavioral Objective Violation

    Jul 15, 2026Mohammad Allahbakhsh, Mohammad Hassan Bahari, Moslem Attar-RaoufPenetration TestingAdversarial Evaluation

  13. A Threshold Exceedance Framework for CBRN Uplift Evaluation in Frontier Language Models

    Jul 13, 2026Rahul Gupta, Abhinav Mohanty, Payal Motwani +8Realistic Threat ModelRise

  14. Large Language Models in Misinformation Ecosystems: Misuse, Defense, and Vulnerability

    Jul 11, 2026Lingwei Wei, Dou Hu, Wei Zhou +2MisinformationRealistic Threat Model

  15. From Forgeries to Foundation Models: A Systematic Survey of Identity Document Attack and Detection

    Jul 1, 2026Gourab Das, Pavan Kumar C, Raghavendra RamachandraForgeriesRealistic Threat Model

  16. Falcon: Functional Assembly and Language for Compositional Reasoning in X-ray

    Jun 24, 2026Yonathan Michael, Mohamad Alansari, Natnael Takele +2X-RayRecent Vision-Language Models

  17. Security and Privacy in Retrieval-Augmented Generation: Architectures, Threats, Defenses, and Future Directions for Building Trustworthy Systems

    Jun 24, 2026Balamurugan Palanisamy, G S S Chalapathi, Vikas Hassija +1Agentic Retrieval-Augmented Generation SystemsLarge Language Model Safety

  18. What Does It Mean to Break a Distillation Defense?

    Jun 23, 2026Lena Libon, Pura Peetathawatchai, Michael Aerni +2Realistic Threat ModelLLM Defense Mechanisms

  19. TRAP: Benchmark for Task-completion and Resistance to Active Privacy-extraction

    Jun 17, 2026Moon Ye-Bin, Nam Hyeon-Woo, Baek Seong-Eun +2PrivacyLeakage

  20. CIAware-Bench: Benchmarking Control Intervention Awareness Across Frontier LLMs

    Jun 9, 2026Joachim Schaeffer, Alexander Panfilov, Thomas Jiralerspong +4Adversarial EvaluationRealistic Threat Model

  21. Assessing Automated Prompt Injection Attacks in Agentic Environments

    Jun 9, 2026David Hofer, Edoardo Debenedetti, Florian TramèrIndirect Prompt InjectionLarge Language Model Agents

  22. Partially Observable Adversarial Patch Attacks on Vision-Language-Action Models in Robotics

    Jun 2, 2026Xiaofei Wang, Mingliang Han, Tianyu Hao +3Diffusion-Based Vision-Language-ActionsRealistic Threat Model

  23. AI Model Extraction Attacks: Bypassing Single-Client Assumptions in Defenses

    Jun 2, 2026Maxime Schwarzer, Johannes F. Loevenich, Gustavo Sánchez +5Realistic Threat ModelLLM Defense Mechanisms

  24. A Full-Pipeline Framework for Evaluating Membership Inference Attacks in Machine Learning

    May 28, 2026Ding Chen, Xinwen Cheng, Xuyang Zhong +3Membership Inference AttacksModel Evaluation

  25. When Muon Optimizer Meets Adversarial Training: A Theoretical and Empirical Study

    May 26, 2026Jun Yan, Weiquan Huang, Jiankai Zuo +4Adversarial TrainingMuon

  26. When Agents Control Robots: A Zero Trust Policy Model for Agentic Cyber-Physical Systems

    May 25, 2026Tharindu Ranathunga, Kavishka Fernando, Susan ReaCyber-Physical SystemRealistic Threat Model

  27. Security in the Fine-Tuning Lifecycle of Large Language Models: Threats, Defenses,Evaluation, and Future Directions

    May 24, 2026Wenjuan Li, Yitao Liu, Runze Chen +1Large Language Model Fine-TuningModel Fine-Tuning

  28. Test-Time Training Undermines Safety Guardrails

    May 21, 2026Simone Antonelli, Sadegh Akhondzadeh, Aleksandar BojchevskiInference-Time DefenseTest-Time Training

  29. MV-Gate: Insider Threat Detection via Multi-View Behavioral Statistics and Semantic Modeling

    May 18, 2026Kaichuan Kong, Dongjie Liu, Xiaobo Jin +1Intrusion DetectionRealistic Threat Model

  30. STRIDE-AI: A Threat Modeling Framework for Generative AI Security Assessment

    May 16, 2026Tsafac Nkombong Regine Cyrille, Franziska SchwarzArtificial Intelligence RiskRealistic Threat Model

  31. Training on Documents About Monitoring Leads to CoT Obfuscation

    May 14, 2026Reilly Haskins, Bilal Chughtai, Joshua EngelsSabotageObfuscation

  32. Watermarking Should Be Treated as a Monitoring Primitive

    May 13, 2026Toluwani Aremu, Nils Lukas, Jie ZhangWatermarkingRealistic Threat Model

  33. CoT-Guard: Small Models for Strong Monitoring

    May 12, 2026Nirav Diwan, Han Wang, Berkcan Kapusuzoglu +6Realistic Threat ModelLLM Defense Mechanisms

  34. MATRA: Modeling the Attack Surface of Agentic AI Systems -- OpenClaw Case Study

    May 11, 2026Tim Van hamme, Thomas Vissers, Javier Carnerero-Cano +4Realistic Threat ModelArtificial Intelligence Risk

  35. PRISM: Generation-Time Detection and Mitigation of Secret Leakage in Multi-Agent LLM Pipelines

    May 11, 2026Riya Tapwal, Abhishek Kumar, Carsten MapleAttacker Large Language ModelRealistic Threat Model

  36. Privacy-Preserving Distributed Learning in IoT Systems: A Unified Threat Model and Evaluation Framework

    May 10, 2026John Cartmell, Alexander WilliamsPrivacy-Preserving Machine LearningInternet Of Thing

  37. ORACLE: Anticipating Scams from Partial Trajectories in Streaming App Usage

    May 9, 2026Wenbo Gao, Songbai Tan, Zhongan Wang +6Fraud DetectionThreat Detection

  38. HBEE: Human Behavioral Entropy Engine -- Pre-Registered Multi-Agent LLM Simulation of Peer-Suspicion-Based Detection Inversion

    May 8, 2026Vickson FerrelRealistic Threat ModelAdversarial Robustness

  39. Narrow Secret Loyalty Dodges Black-Box Audits

    May 7, 2026Alfie Lamerton, Fabien RogerRealistic Threat ModelPoisoning

  40. Gray-Box Poisoning of Continuous Malware Ingestion Pipelines

    May 6, 2026Jan Dolejš, Martin Jureček, Róbert LórenczMalwarePoisoning

  41. On the Privacy of LLMs: An Ablation Study

    May 4, 2026Karima Makhlouf, Lamiaa Basyoni, Syed Khaderi +4Membership Inference AttacksAttacker Large Language Model

  42. Trojan Hippo Bench: A Dynamic Benchmark for Persistent Memory Attacks and Defenses in LLM Agents

    May 3, 2026Debeshee Das, Julien Piet, Darya Kaviani +3Agentic MemoryRealistic Threat Model

  43. From Prompt to Physical Actuation: Holistic Threat Modeling of LLM-Enabled Robotic Systems

    Apr 29, 2026Neha Nagaraja, Hayretdin Bahsi, Carlo R. da CunhaRealistic Threat ModelThreat

  44. Risk Reporting for Developers' Internal AI Model Use

    Apr 27, 2026Oscar Delaney, Sambhav Maheshwari, Joe O'Brien +2Artificial Intelligence RiskPre-Deployment Safety Assessments

  45. Agentic Adversarial Rewriting Exposes Architectural Vulnerabilities in Black-Box NLP Pipelines

    Apr 26, 2026Mazal Bethany, Kim-Kwang Raymond Choo, Nishant Vishwamitra +1Attacker Large Language ModelEvasion