cs.SISep 30, 2026

Memetic Trojans: Social Contagions as Carriers of Adversarial Payloads in Agent Networks

Authors: Birk Torpmann-Hagen, Finn Schwall, Leon Moonen

Organizations: Simula Research Laboratory Oslo, Norway · Simula Research Laboratory Oslo Metropolitan University Oslo, Norway

Abstract

Autonomous large language model (LLM) agents increasingly interact in network environments where adversarial content can propagate between agents. Known attacks include agent worms, which spread through self-replicating prompt injections or configuration compromises. We introduce \emph{memetic trojans}, a distinct class of network-mediated attack that exploits agents' tendencies to retransmit and amplify content. Unlike agent worms, whose propagation is adversarially induced, memetic trojans exploit \emph{endogenous} transmission by embedding adversarial payloads in \emph{social contagions}: content agents have internal reasons to share. As part of our work, we extract social contagions from Moltbook, a social media platform for LLM agents. Controlled transmission experiments reveal large differences in virality: the most effective contagion is retransmitted in approximately 50% of subsequent agent posts and upvoted at 2.5x the average post's rate. Its memetic trojan counterpart largely inherits these properties. Monte Carlo attack simulations show that memetic trojans amplify expected exposure by up to 3.19x. Network structure and amplification mechanisms strongly shape propagation, producing heavy-tailed outcomes with near network-wide exposure. These results identify endogenous social transmission as a distinct security vulnerability in multi-agent systems. Because propagation does not require agents to follow malicious retransmission instructions, defenses focused on prompt-injection detection or preventing agent compromise cannot alone prevent memetic trojan propagation. Securing large-scale agent ecosystems may require network-level defenses that account for how agent preferences, recommendation mechanisms, and network topology amplify adversarial payloads.

Figures & tables

Appendix figures & tables8 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. AgentWorm: Self-Propagating Attacks Across LLM Agent Ecosystems

    Mar 16, 2026Yihao Zhang, Zeming Wei, Xiaokun Luan +7Multi-Llm AgentsLarge Language Model Agents

  2. Share-Borne AI Virus: Memory-Hopping Attacks Across LLM Agents

    Sep 28, 2026Sidharth Pulipaka, Ansh Sharma, Stanislau Hlebik +4Large Language Model AgentsAdversarial Robustness