cs.CLMay 22, 2026

Emotion Recognition in Sign Language Conversation

Authors: Yusong WangKeyu MaoTakao ObiMinghao ShaoKotaro Funakoshi

Organizations: Institute of Science Tokyo, Japan · New York University, New York, 11201, USA

Abstract

Emotion Recognition in Conversation is a core component of affective computing, while current sign language emotion datasets primarily focus on isolated sentences and lack conversational context. Models trained exclusively on these isolated utterances demonstrate degraded performance in real world scenarios because they cannot utilize historical dialogue flow. To address this structural limitation, we introduce the ERC task to sign language video analysis and propose the eJSL Dialog dataset. Constructed using the scripts from the STUDIES corpus, the dataset contains 1,920 video samples organized into 480 unique dialogues. We conduct systematic benchmarking on this dataset using models ranging from isolated visual networks to multimodal conversational architectures. The results reveal a domain gap when applying generic multimodal conversational emotion recognition models to sign language. These findings demonstrate the explicit need for context-aware visual extractors specific to sign language and indicate that constructing larger conversational datasets to support large-scale pre-training is a necessary next step for future research.

Explore similar work

CardsList
  1. Emotion Recognition in Signers

    Dec 17, 2025Kotaro Funakoshi, Yaoxiong ZhuSign LanguageEmotion Recognition