Language Model Evasion Attacks

Recent momentum

-32%

13 papers in the last 28 days · 0.2% of indexed attention

Twelve weeks of publication activity for this topic as it is defined today.

Weekly history

Recent digests

What was published in this topic, kept on the site without email delivery.

Period ending 2026-09-21

10 new papers

A weekly snapshot of new work published in Language Model Evasion Attacks.

Period ending 2026-09-14

1 new paper

A weekly snapshot of new work published in Language Model Evasion Attacks.

Period ending 2026-09-07

1 new paper

A weekly snapshot of new work published in Language Model Evasion Attacks.

202 papers

Latest in Language Model Evasion Attacks

Open your feed →
CardsList
  1. Test-Time Unlearning via Sparse Autoencoder

    Sep 14, 2026Pingzhi Li, Jinhao Duan, Vaishnav Tadiparthi +6Language Model Evasion AttacksForgetting

  2. LaCache: Robust Semantic Caching for LLM Serving

    Aug 3, 2026Jiacheng Liang, Yuhui Wang, Tanqiu Jiang +1CacheLanguage Model Evasion Attacks

  3. GPT-Red: Automated Red Teaming via Self-Play at Scale

    Jul 28, 2026Eric Wallace, Christopher A. Choquette-Choo, Nikhil Kandpal +15Red-TeamingLanguage Model Evasion Attacks