Harms

Recent momentum

-42%

7 papers in the last 28 days · 0.1% of indexed attention

Twelve weeks of publication activity for this topic as it is defined today.

Weekly history

Recent digests

What was published in this topic, kept on the site without email delivery.

Period ending 2026-09-21

3 new papers

A weekly snapshot of new work published in Harms.

Period ending 2026-09-14

2 new papers

A weekly snapshot of new work published in Harms.

Period ending 2026-09-07

3 new papers

A weekly snapshot of new work published in Harms.

70 papers

Latest in Harms

  1. The Role of Fine-grained Harm Signals in LLM Safety

    Sep 16, 2026Soyeon Park, Seogyeong Jeong, Sunwoo Kim +1Large Language Model SafetyHarms

  2. AI and Consumer Rights in India Working Paper

    Aug 13, 2026Omir Kumar, Sriya Sridhar, Vibhav Mithal +1Artificial Intelligence SystemsGhana

  3. Sound Probabilistic Safety Bounds for Large Language Models

    Jul 22, 2026Mahdi Nazeri, Anne-Kathrin Schmuck, Sadegh Soudjani +1Large Language Model SafetyHarms

  4. One Year Later...The Harms Persist, But So Do We!

    Jun 22, 2026Annika Marie Schoene, Cansu Canca, Gautham Vijay Kumar +1Mental HealthLarge Language Model Safety

  5. On Defining Erasure Harms for NLP

    Jun 14, 2026Yu Lu Liu, Arnav Goel, Jackie Chi Kit Cheung +3Concept ErasureHarms

  6. First, do no harm: Breaking suicidogenic echo chambers in media recommendation

    May 24, 2026Alberto Díaz-Álvarez, Raúl Lara-Cabrera, Fernando Ortega-Requena +1Suicide RiskRecommendation

  7. The Moltbook Files: A Harmless Slopocalypse or Humanity's Last Experiment

    May 8, 2026William Brach, Federico Torrielli, Stine Lyngsø Beltoft +3OpenclawEmergence

  8. Characterizing the Consistency of the Emergent Misalignment Persona

    Apr 30, 2026Anietta Weckauff, Yuchen Zhang, Maksym AndriushchenkoMisalignmentPersona