Agent Skill Security Auditing

Momentum

8 papers in the last four weeks, against 1 the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 37

All topics
CardsList
  1. Runaway Reaction: When Benign Skills Compose into Malicious Behavior

    Oct 5, 2026Zunlong Zhou, Ziyuan Yang, Mengyu Sun +1AI Agent SecurityLLM Agent Security

  2. Pretext: Defeating Malicious Skill Detection Frameworks for AI Agents

    Sep 30, 2026Tobias Kaiser, Aritra DharAI Agent SecurityAgent Skill Security Auditing

  3. Hiding in Plain Sight: Decoupling Pretext from Actuation for Skill Poisoning in LLM Agents

    Sep 30, 2026Wenxin Wu, Lingyong Yan, Lei Sha +2Adversarial Attacks on LLMsLLM Agent Security

  4. Can Agents Trust Their Skills? Uncovering Unsafe Chains of Trust in Skill-Based LLM Agents

    Sep 30, 2026Yan Wang, Zhihao Zhang, Ke Chen +5LLM Agent SecurityPrompt Injection Attacks on AI Agents

  5. SKILLLITE: Evidence-Guided Malicious Skill Auditing with Compact LLMs

    Sep 29, 2026Haoran Ou, Gelei Deng, Xuanye Zhang +3Small Language ModelsLLM Agent Security

  6. MMSkillRisk: Can Agents Stay Safe When Multimodal Skills Become Traps?

    Sep 28, 2026Lingqi Jiang, Jialuo Chen, Jianan Ma +8Prompt Injection Attacks on AI AgentsAgent Skill Security Auditing

  7. SkillAtlas: An Attack Trace Library for Agent Skills

    Sep 11, 2026Yuxin Tian, Zenghao Duan, Liang Pang +2LLM Agent SecurityAgent Skill Security Auditing

  8. AgentLeak: Cloning Stronger LLM Agent Capabilities onto Weaker Agents Beyond Skill Stealing

    Sep 7, 2026Xiaoting Lyu, Yuhong Wu, Yufei Han +6LLM Agent SecurityLLM Agent Skill Learning

  9. ColluSkill: Adversarial Cross-Skill Composition for Evading Agent Skill Scanners

    Aug 10, 2026Puyu Zeng, Simeng Qin, Jingzhi Li +3AI Agent SecurityLLM Agent Security

  10. SkillsMetric: Mapping the Detection Boundary of Static Analysis for Malicious Agent Skills

    Aug 9, 2026Xinze Chen, Chi Zhang, Ping Ji +1Static Code AnalysisAI Agent Security Benchmarks

  11. SkillConsist: Detecting Inconsistencies in Agent Skills via Bidirectional Graph Alignment

    Aug 7, 2026Chaofan Meng, Yuhang Zheng, Yingnan Zhou +1Agent EvaluationAgent Skill Security Auditing

  12. SkillTrace: Multi-Trace Provenance Auditing for LLM-Agent Skill Reuse

    Aug 5, 2026Jialuo Chen, Minghe Wang, Lingqi Jiang +7Data ProvenanceAI Agent Auditing

  13. When Agents Learn to Be You: Benchmarking Privacy Leakage, Impersonation Risk, and Defenses in Persona Skills

    Aug 4, 2026Yongli Xiang, Zhifang Zhang, Bojun Yang +4Privacy Leakage in Language ModelsAI Agent Security Benchmarks

  14. Skillsets on the Chain: A Blockchain-based Zero-Trust Framework for Agentic AI Networking

    Jul 31, 2026Yayu Gao, Yong Xiao, Hao Hu +5CybersecurityAI Agent Security

  15. OpenSkillRisk: Benchmarking Agent Safety When Using Real-World Risky Third-Party Skills

    Jul 22, 2026Qiyuan Liu, Tingfeng Hui, Kun Zhan +2LLM Agent EvaluationAI Agent Safety

  16. CONTRA: Red-Teaming Configurations of Personalizable Agents

    Jul 3, 2026Jonathan Nöther, Adish Singla, Goran RadanovicLLM Red TeamingAI Agent Security

  17. SkillFuzz: Fuzzing Skill Composition for Implicit Intents Discovery in Open Skill Marketplaces

    Jul 2, 2026Jinwei Hu, Yi Dong, Youcheng Sun +1LLM AgentsAgent Skill Security Auditing

  18. Skills Are Not Islands: Measuring Dependency and Risk in Agent Skill Supply Chains

    Jul 1, 2026Changguo Jia, Tianqi Zhao, Runzhi He +1LLM AgentsAgent Skill Security Auditing

  19. Understanding and Evaluating Claw-like Agent Security Through a Computer-Systems Lens

    Jun 29, 2026Peizhi Niu, Wenjie Qu, Shangding Gu +14AI Agent SecurityLLM Agent Security

  20. Detecting Malicious Agent Skills in the Wild using Attention

    Jun 22, 2026Bacem Etteib, Daniele Lunghi, Tégawendé F. BissyandéLLM Agent SecurityAgent Skill Security Auditing

  21. SkillAudit: From Fixed-Suite Benchmarking to Skill-Centered Assessment

    Jun 21, 2026Dexu Yu, Youhua Li, Zhaoyang Guan +12AI Agent AuditingAutomated Evaluation

  22. Seeing Is Not Screening: Multimodal Hidden Instruction Attacks on Agent Skill Scanners

    Jun 16, 2026Xiaojun Jia, Jie Liao, Simeng Qin +5LLM Agent SecurityPrompt Injection Defense

  23. SkillVetBench: LLM-as-Judge for Multi-Dimensional Security Risk Evaluation in Open-Source LLM Agent Skills

    Jun 14, 2026Ismail Hossain, Sai Puppala, Md Jahangir Alam +2LLM-as-a-JudgeLLM Agent Security

  24. Benign in Isolation, Harmful in Composition: Security Risks in Agent Skill Ecosystems

    Jun 13, 2026Yi Xie, Jiawei Du, Yu Cheng +2LLM Agent SecurityAI Agent Security Benchmarks

  25. ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree

    May 31, 2026Vincent Koc, Patrick Erichsen, Jacob Tomlinson +3AI Agent SecurityAgent Skill Security Auditing

  26. Benchmarking Security Risk Detection and Verification in Open Agentic Skill Ecosystems

    May 30, 2026Ismail Hossain, Sai Puppala, Zhuoran Lu +2AI Agent SecurityAI Agent Security Benchmarks

  27. When Safe Skills Collide: Measuring Compositional Risk in Agent Skill Ecosystems

    May 30, 2026Su Wang, Pin Qian, Yihang Chen +6Software SecurityAI Agent Security