cs.CRSep 30, 2026

Hiding in Plain Sight: Decoupling Pretext from Actuation for Skill Poisoning in LLM Agents

Authors: Wenxin Wu, Lingyong Yan, Lei Sha, Shuaiqiang Wang, Jiashu Zhao

Organizations: Beihang University · Baidu Inc · Wilfrid Laurier University

Abstract

LLM agents increasingly rely on reusable Skills for complex, multi-step tasks, creating a critical supply-chain attack surface where poisoned Skill content steers agent decision loops under benign requests. Existing skill poisoning attacks either colocate actuation with its contextual pretext or distribute actuation across multiple Skills, but do not explicitly separate the rationale for execution from the operation itself. In this work, we reveal that untrusted agent decisions fundamentally depend on two conceptually distinct Risk-Realization Factors (RRFs): an actuation factor (specifying what concrete operation is performed) and a pretext factor (providing the situational rationale for why the agent must perform it). Guided by this abstraction, we propose a coordination-based attack paradigm: decoupling pretext from actuation. Rather than fragmenting the malicious actuation, we preserve it as an intact operation within a downstream Steering Skill, while delegating the pretext factor to an upstream Grounding Skill that subtly alters persistent environment artifacts through routine utility operations. The intact actuation thus hides in plain sight, appearing completely legitimate and task-driven only when evaluated against the fabricated pretext. Building on this formulation, we develop an automated framework that discovers authentic execution dependencies, synthesizes coordinated pretext-actuation skill pairs, and iteratively refines poisoned skill instructions via runtime closed-loop feedback. Extensive evaluations across single-session and persistent cross-lifecycle scenarios demonstrate that decoupled skill poisoning achieves high attack success, exposing a critical blind spot in isolated Skill security audits. Our automated framework code is available at https://github.com/Wenxin-buaa/CoordPoison.git.

Figures & tables

Appendix figures & tables4 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. RouteGuard: Internal-Signal Detection of Skill Poisoning in LLM Agents

    Apr 24, 2026Wenjie Xiao, Xuehai Tang, Biyu Zhou +2Malicious AgentsPoisoning

  2. POISE: Position-Aware Undetectable Skill Injection on LLM Agents

    Jun 6, 2026Haochang Hao, Dehai Min, Zhifang Zhang +4PoisoningMalicious Agents

  3. Pretext: Defeating Malicious Skill Detection Frameworks for AI Agents

    Sep 30, 2026Tobias Kaisar, Aritra DharMalicious AgentsAttacker Large Language Model