Jailbreak Evaluations

Recent momentum

emerging

0 papers in the last 28 days · 0.0% of indexed attention

Twelve weeks of publication activity for this topic as it is defined today.

Weekly history

Recent digests

What was published in this field, kept on the site without email delivery.

Period ending 2026-09-21

2 new papers

A weekly snapshot of new work published in Jailbreak Evaluations.

Period ending 2026-09-14

6 new papers

A weekly snapshot of new work published in Jailbreak Evaluations.

Period ending 2026-09-07

6 new papers

A weekly snapshot of new work published in Jailbreak Evaluations.

Inside this field

Focused directions

184 papers

Latest in Jailbreak Evaluations

Open your feed →
CardsList
  1. Stealing Reasoning Traces from Proprietary LLM APIs

    Aug 10, 2026Alexander Panfilov, David Schmotz, Ilia Shumailov +5Large Language Model ReasoningReasoning Traces

  2. Asymmetric Collapse in Model Merging: When Refusal Over- writes Recognition

    Jul 26, 2026Aarnav Choudhary, Matheus Fonseca Rocha, Jiwon Seo +2Model MergingRefusals