cs.CLOct 8, 2026

SignRAG: Unified Retrieval-Augmented Gloss-Free Sign Language Translation

Authors: Zhi Rao, Yucheng Zhou, Qianran Sun, Yiqing Huang, Longcan Yuan, Jiayi Hou, Chengwen Yao, Lin Cheng, +3 more

Organizations: Faculty of Innovation Engineering, Macau University of Science and Technology, Macau, China · SKL-IOTSC, CIS, University of Macau, Macau, China · MAIS, Institute of Automation, Chinese Academy of Sciences, Beijing, China · Yale University, New Haven, CT, USA · VIVO AI Lab, China

Abstract

Contemporary decoder-only large language models (LLMs) have demonstrated strong capabilities across a wide range of domains. However, existing pretraining paradigms for gloss-free sign language translation (SLT) are largely designed around conventional encoder-decoder pretrained language models, which limits their direct applicability to decoder-only LLMs. To address this limitation, we propose SignRAG, a unified framework combining hierarchical pretraining, target-domain retrieval augmentation, and retrieval-aware reinforcement fine-tuning. Hierarchical pretraining first learns linguistically grounded sign representations and then jointly aligns the sign encoder with an LLM, mitigating cross-modal optimization imbalance. For downstream adaptation, SignRAG complements parameter-based fine-tuning with a target-domain retrieval gallery that provides instance-specific translation cues. To ensure that retrieved contexts are used appropriately, we further introduce Retrieval Utility-Guided Reinforcement Fine-Tuning (RUG-RFT), which combines translation-quality and retrieval-utility rewards to encourage beneficial retrieval use while suppressing harmful reliance. Experiments on multiple SLT benchmarks establish new state-of-the-art performance. In particular, to the best of our knowledge, SignRAG is the first gloss-free approach to outperform gloss-supervised methods across all reported metrics on CSL-Daily. Our code has been released at GitHub, together with models of different sizes to support future academic research.

Explore similar work

CardsList
  1. SignLlama: Enhancing Gloss-free Sign Language Translation by Prioritizing Visual Features for LLMs

    Aug 10, 2026Shiwei Gan, Xiao Liu, Yafeng Yin +5Sign Language TranslationVLM Adaptation

  2. Think in Latent Thoughts: A New Paradigm for Gloss-Free Sign Language Translation

    Apr 16, 2026Yiyang Jiang, Li Zhang, Xiao-Yong Wei +1Sign Language TranslationVideo-Language Models

  3. DualAnchor: Preserving Language Priors and Improving Lexical Fidelity in Gloss-Free Sign Language Translation

    Jul 30, 2026Hongbin Zhang, Junhao Liu, Xuefeng Bai +3Sign Language TranslationVision-Language Alignment