LLMs for Cybersecurity

LLM: Large Language Model

Momentum

19 papers in the last four weeks, up 280% on the four weeks before. 0.2% of all new papers.

Jul 13Week of Sep 28

Latest papers 166

All topics
CardsList
  1. Post-Hoc Trajectory-Risk Certification for Modular LLM-Based Security Agents

    Aug 4, 2026Zhenpeng LiIoT Intrusion DetectionCybersecurity

  2. DiagChain: A Diagnostic Benchmark for Evaluating LLM Agents on Evidence-Grounded Attack Chain Reconstruction

    Aug 4, 2026Xuyang Liu, Yibin Han, Zhenwei Zhang +8Retrieval-Augmented GenerationLLM Agent Evaluation

  3. Antares: Foundation Models for Agentic Vulnerability Localization

    Aug 3, 2026Supriti Vijay, Aman Priyanshu, Didier Chapoteau +8Software Vulnerability DetectionSmall Language Models

  4. EntailLLM: Verifying LLM-Generated Vulnerability Discovery Paths with Domain Knowledge via Logic Programming

    Aug 3, 2026Kaustuv Mukherji, Jaikrishna Manojkumar Patil, Colton Payne +4Logic ProgrammingSoftware Vulnerability Detection

  5. MITRE-SAGE: A Multi-Agent Cybersecurity Question-Answering Model

    Aug 3, 2026Ali Habibzadeh, Farid Feyzi, Reza Ebrahimi AtaniRetrieval-Augmented GenerationMulti-Agent LLM Systems

  6. From Chasing Ghosts to Missed Attacks: Perspectives and Perceptions of SOC Practitioners on LLM Integration, Risks, and Readiness

    Aug 1, 2026Jonas Thurner, Nadine Jost, Stefan Albert Horstmann +4CybersecurityLLMs for Cybersecurity

  7. Distilling Knowledge from Large Language Models into Lightweight Reinforcement Learning Agents for Autonomous Cyber Operations

    Jul 30, 2026Konur Tholl, François Rivest, Mariam El Mezouar +2Autonomous Cyber DefenseLLM-Guided RL

  8. ThreatForest: Multi-Agent Attack Tree Generation with Pluggable TTP Framework Mapping

    Jul 29, 2026Cristian Leo, Anton Dykyi, Danny Cortegaca +2LLMs for Cybersecurity

  9. HoF-Bench: Rediscovering Real AI-Discovered CVEs Without Frontier Models

    Jul 29, 2026Petr Simecek, Elnaz Babayeva, Jiri Balhar +23Software Vulnerability DetectionLLMs for Cybersecurity

  10. KQFuzz: Knowledge-Guided Fuzzing for Quantum Libraries via Large Language Models

    Jul 28, 2026Fuyuan Xia, Qixin Zhang, Chenhao Ying +5Automated Software TestingLLMs for Cybersecurity

  11. The Disruptive Impact of Large Language Models on Capture the Flag Competitions and the Path Toward Fair Play

    Jul 28, 2026Michael Macaulay, Harmony Bouabid, Guo Gen Ang +1CybersecurityLLMs for Cybersecurity

  12. DeepFaith: Evidence-Grounded LLMs for Faithful Incident Reporting in Multi-Stage APT Defense

    Jul 27, 2026Trung V. Phan, Tri Gia Nguyen, Thomas BauschertCybersecurityLLMs for Cybersecurity

  13. Every Model Cheats: Prompt-Level Mitigation of Cheating on Offensive Cyber Tasks

    Jul 23, 2026Michael Kouremetis, Ads Dawson, Raja Sekhar Rao Dheekonda +1LLM AuditingDeception in Language Models

  14. From Evaluation to Optimisation: Hierarchy-Aware Training Signals for CWE Prediction in Python

    Jul 23, 2026Muntasir Adnan, Manile Srun, Carlos C. N. KuhnSoftware Vulnerability DetectionLLMs for Cybersecurity

  15. Beyond Heavy Log Curation: Perplexity-Based APT Detection via Unsupervised, Context-Augmented Language Models

    Jul 23, 2026Shoya Otsu, Kei Suzuki, Toshiaki Koike-Akino +2Online Anomaly DetectionLLMs for Cybersecurity

  16. Evaluating Large Language Models for Symbolic Security Protocol Analysis

    Jul 22, 2026Paolo Modesti, Syed Ahmed, Ioannis Sfyrakis +1LLM EvaluationCybersecurity

  17. Find Before You Fine-Tune: A Diagnostic Study of Small LLMs for Cybersecurity QA

    Jul 21, 2026Shaswata Mitra, Subash Neupane, Trisha Chakraborty +4LLM EvaluationLLM Fine-Tuning

  18. Semi-Automated Detection of Gaps in LLM Security Knowledge

    Jul 20, 2026Shufan Chai, Liangliang Sun, Jessica StaddonLLM AuditingLLMs for Cybersecurity

  19. Evaluating Open-Weight LLMs for Generating Structured Threat Information for Autonomous Vehicle Vulnerabilities

    Jul 17, 2026Md Erfan, Ahmed Ryan, Md Kamal Hossain Chowdhury +1CybersecurityLLMs for Cybersecurity

  20. AutoTrace: From Patches to Triggers via Agentic Interprocedural Exploration

    Jul 13, 2026Arastoo Zibaeirad, Marco Vieira, Thomas ZimmermannSoftware Vulnerability DetectionStatic Code Analysis

  21. SMETA-ZSL:Semantic Meta-Alignment for Zero-Shot Threat Classification

    Jul 10, 2026Ivan Alejandro Montoya Sanchez, Anantaa Kotal, Aritran PiplaiZero-Shot LearningCyber Threat Intelligence

  22. From Legacy Documentation to OSCAL: An MCP-Based Agent Pipeline for Threat-Informed Continuous Compliance in Critical Infrastructure

    Jul 9, 2026Lea Roxanne Muth, Marian MargrafCybersecurityLLMs for Cybersecurity

  23. Large Language Models (LLMs) and Generative AI in Cybersecurity and Privacy: A Survey of Dual-Use Risks, AI-Generated Malware, Explainability, and Defensive Strategies

    Jul 8, 2026Kiarash Ahi, Saeed ValizadehLLM SecurityLLMs for Cybersecurity

  24. Beyond Refusal: A Same-Lineage Study of Aligned and Abliterated LLMs for Vulnerability Analysis

    Jul 7, 2026Mingchen Li, Meikang Qiu, Zifan Peng +4Software Vulnerability DetectionLLMs for Cybersecurity

  25. TACTIC-KG: Toward Small Agent Teams for Cyber Threat Intelligence Knowledge Graph Construction

    Jul 6, 2026Mouhamed Amine Bouchiha, Gregory BlancKG ConstructionMulti-Agent LLM Systems