Harms

Recent momentum

emerging

0 papers in the last 28 days · 0.0% of indexed attention

Twelve weeks of publication activity for this topic as it is defined today.

Weekly history

Recent digests

What was published in this field, kept on the site without email delivery.

Period ending 2026-09-21

4 new papers

A weekly snapshot of new work published in Harms.

Period ending 2026-09-14

2 new papers

A weekly snapshot of new work published in Harms.

Period ending 2026-09-07

3 new papers

A weekly snapshot of new work published in Harms.

Inside this field

Focused directions

129 papers

Latest in Harms

  1. The Role of Fine-grained Harm Signals in LLM Safety

    Sep 16, 2026Soyeon Park, Seogyeong Jeong, Sunwoo Kim +1Large Language Model SafetyHarms

  2. Detoxifying Toxic Communication: A Design Science Approach to Responsible AI

    Aug 31, 2026Hossein Arshadi Soufiani, Henry M. Kim, Hjalmar Turesson +2Harmful ContentDetoxification

  3. AI and Consumer Rights in India Working Paper

    Aug 13, 2026Omir Kumar, Sriya Sridhar, Vibhav Mithal +1Artificial Intelligence SystemsGhana

  4. ToxScreen: Detecting Whether an LLM Has Been Poisoned

    Jul 29, 2026Anthony Hughes, Nicole Xing, Collin Francel +2Backdoor AttacksToxicity

  5. Sound Probabilistic Safety Bounds for Large Language Models

    Jul 22, 2026Mahdi Nazeri, Anne-Kathrin Schmuck, Sadegh Soudjani +1Large Language Model SafetyHarms

  6. ToxiREX: A Dataset on Toxic REasoning in ConteXt

    Jun 26, 2026Stefan F. Schouten, Ilia Markov, Piek VossenToxicityHarmful Content

  7. One Year Later...The Harms Persist, But So Do We!

    Jun 22, 2026Annika Marie Schoene, Cansu Canca, Gautham Vijay Kumar +1Mental HealthLarge Language Model Safety

  8. On Defining Erasure Harms for NLP

    Jun 14, 2026Yu Lu Liu, Arnav Goel, Jackie Chi Kit Cheung +3Concept ErasureHarms

  9. The Tone of Awareness: Topic, Sentiment, and Toxicity Maps During Mental Health Month on TikTok

    Jun 11, 2026Henrique Ferraz de Arruda, Andreia Sofia Teixeira, Pranay Gundala Reddy +3Mental HealthToxicity