Period ending 2026-09-21
27 new papers
A weekly snapshot of new work published in Large Language Model Safety.
Twelve weeks of publication activity for this topic as it is defined today.
Weekly history
What was published in this field, kept on the site without email delivery.
Period ending 2026-09-21
A weekly snapshot of new work published in Large Language Model Safety.
Period ending 2026-09-14
A weekly snapshot of new work published in Large Language Model Safety.
Inside this field
Within Large Language Model Safety
Within Large Language Model Safety
Within Large Language Model Safety
Within Large Language Model Safety
855 papers
Cognitive Overload'', hypothesizing that the effort required to decipher degraded inputs diverts attentional resources from safety auditing. This phenomenon is consistent across various visual perturbations, including noise and geometric distortion. To address this, we propose a simple Structured Cognitive Offloading'' strategy that mitigates these risks by enforcing a serialized pipeline to decouple visual transcription from safety assessment. Our work exposes a significant risk in vision-based compression and provides critical insights for the secure design of future MLLMs.