LLMs for Cybersecurity

LLM: Large Language Model

Momentum

19 papers in the last four weeks, up 280% on the four weeks before. 0.2% of all new papers.

Jul 13Week of Sep 28

Latest papers 166

All topics
CardsList
  1. Evaluating LLM-Generated Obfuscated XSS Payloads for Machine Learning-Based Detection

    Apr 21, 2026Divyesh Gabbireddy, Suman SahaSoftware Vulnerability DetectionAdversarial Examples

  2. Do Agents Dream of Root Shells? Partial-Credit Evaluation of LLM Agents in Capture the Flag Challenges

    Apr 21, 2026Ali Al-Kaswan, Maksim Plotnikov, Maxim Hájek +3LLM Agent EvaluationAI Agent Security Benchmarks

  3. Systematic Capability Benchmarking of Frontier Large Language Models for Offensive Cyber Tasks

    Apr 18, 2026Tyler H. Merves, Michael H. Conaway, Joseph M. Escobar +2LLM Agent EvaluationCybersecurity

  4. Analyzing Chain of Thought (CoT) Approaches in Control Flow Code Deobfuscation Tasks

    Apr 16, 2026Seyedreza Mohseni, Sarvesh Baskar, Edward Raff +1Software Reverse EngineeringCoT Reasoning

  5. Quantifying Frontier LLM Capabilities for Container Sandbox Escape

    Mar 1, 2026Rahul Marchand, Art O Cathain, Jerome Wynne +5LLM Safety BenchmarksLLM Agent Security

  6. TrojanGYM: A Detector-in-the-Loop LLM for Adaptive RTL Hardware Trojan Insertion

    Jan 23, 2026Saideep Sreekumar, Zeng Wang, Akashdeep Saha +6Backdoor AttacksAdversarial Robustness

  7. Toward Cybersecurity-Expert Small Language Models

    Oct 15, 2025Matan Levi, Daniel Ohayon, Ariel Blobstein +3Cyber Threat IntelligenceCybersecurity

  8. What's on My Network? Using Large Language Models to Identify Real-World IoT Devices at Scale

    Sep 24, 2025Rameen Mahmood, Tousif Ahmed, Sai Teja Peddinti +1Internet of ThingsLLMs for Cybersecurity

  9. FALCON: Transforming Cyber Threat Intelligence into Deployable IDS Rules with Self-Reflection

    Aug 26, 2025Shaswata Mitra, Subash Neupane, Martin Duclos +5Cyber Threat IntelligenceNetwork Intrusion Detection

  10. Less Data, More Security: Advancing Cybersecurity LLMs Specialization via Resource-Efficient Domain-Adaptive Continuous Pre-training with Minimal Tokens

    Jun 30, 2025Salahuddin Salahuddin, Ahmed Hussain, Jussi Löppönen +1Language Model PretrainingDomain-Adaptive Pretraining

  11. DMind Benchmark: Toward a Holistic Assessment of LLM Capabilities across the Web3 Domain

    Apr 18, 2025Enhao Huang, Pengyu Sun, Shuxun Wang +13LLM EvaluationSoftware Vulnerability Detection

  12. Large Language Models for Cryptocurrency Transaction Analysis: A Bitcoin Case Study

    Jan 30, 2025Yuchen Lei, Yuexin Xiang, Rafael Dowsley +5LLM EvaluationLLMs for Cybersecurity

  13. ANVIL: Anomaly-based Vulnerability Identification without Labelled Training Data

    Aug 28, 2024Weizhou Wang, Eric Liu, Xiangyu Guo +3Software Vulnerability DetectionLLMs for Cybersecurity