Exploitation

Recent momentum

-56%

4 papers in the last 28 days · 0.1% of indexed attention

Twelve weeks of publication activity for this topic as it is defined today.

Weekly history

Recent digests

What was published in this topic, kept on the site without email delivery.

Period ending 2026-09-14

3 new papers

A weekly snapshot of new work published in Exploitation.

Period ending 2026-09-07

2 new papers

A weekly snapshot of new work published in Exploitation.

83 papers

Latest in Exploitation

Open your feed →
CardsList
  1. Multimodal Reward Hacking in Reinforcement Learning

    Jul 10, 2026Jiayu Yao, Yiwei Wang, Anmeng Zhang +5Exploitation

  2. Mitigating LLM-based p-Hacking by Preregistering for the Next LLM

    Jun 26, 2026Maria Thomas, Kristina Gligoric, Nihar B. ShahLarge LanguageMitigation

  3. Cheap Reward Hacking Detection

    Jun 8, 2026Iván Belenky, Joaquín Itria, Steven JohnsExploitationEncoding Models

  4. unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning

    May 27, 2026Geoffrey Bradway, Roger Creus Castanyer, Lorenz Wolf +3ExploitationRobotwin

  5. Imperfect World Models are Exploitable

    May 15, 2026Logan Mondal Bhamidipaty, Esmeralda S. Whitammer, David Abel +2ExploitationWorld Models