cs.CLSep 28, 2026

In-game Toxic Detection: Bi-directional Representations with Attention Residuals

Authors: Yuanzhe Jia

Organizations: University of Sydney, Australia

Abstract

In-game toxic language has emerged as a critical concern in the gaming industry and community. While several frameworks and models for online game toxicity analysis have been proposed, detecting toxicity in player chat utterances remains a formidable challenge: stemming not only from the extremely short length of such utterances but also from the heavy reliance on game slang, abbreviations, and domain-specific jargon, which generic language models are poorly suited to recognize. This paper presents a shared task for in-game toxic language detection built upon real-world in-game chat data, and proposes the best-preforming model for the toxic language slot filling: Bi-directional Representations with Attention Residuals (BRAR). Experimental results demonstrate that BRAR effectively captures the global context and outperforms the existing baselines on slot filling.

Figures & tables

Explore similar work

CardsList
  1. Where Does Toxicity Live? Mechanistic Localization and Targeted Suppression in Language Models

    May 27, 2026Himanshu Beniwal, Mayank SinghToxicityLarge Language Model Safety

  2. Toxicity in Twitch Chats: An LLM-Based Analysis Across Gaming Communities

    May 18, 2026Ronja Fuchs, Florian Rupp, Timo Bertram +2Toxicity