cs.CVOct 8, 2026

Beyond Visual Enhancement: Adaptive Multi-Context Steering to Mitigate LVLM Hallucinations

Authors: Shuran Ma, JiaLe Li, Yuxin Dong, Shan Zheng, Qingyun Jiang, Xiang Chen, Qi Zhu, Deyi Ji, +4 more

Organizations: Shanghai Jiao Tong University · Peking University · Beijing University of Chemical Technology · Nanjing University of Aeronautics and Astronautics · University of Science and Technology of China · Tsinghua University

Abstract

Hallucination remains a significant challenge in Large Vision-Language Models (LVLMs). Existing training-free methods generally mitigate hallucinations through contrastive decoding or visual enhancement, often increasing the relative influence of visual evidence during generation. This raises a fundamental question: Can LVLMs dynamically regulate the contributions of different context sources to suppress hallucinations? In this work, we investigate and quantify how LVLMs coordinate multiple context sources during decoding and examine how this intrinsic behavior can guide hallucination mitigation. We find that LVLMs exhibit an intrinsic vision-attending tendency that can guide adaptive visual steering, while textual contexts can also contribute to hallucination mitigation. Motivated by these findings, we propose AIMS (Adaptive Information Multi-source Steering), a lightweight training-free framework that adaptively coordinates visual, prefilled textual, and generated contexts during decoding. Specifically, AIMS constructs compact prototypes for the three context domains and estimates their affinities with the current query to determine head-wise steering weights. The resulting multi-source steering direction is applied to the query representation, enabling adaptive context integration without additional model training or auxiliary forward passes. Extensive experiments across multiple LVLMs and decoding strategies demonstrate that AIMS effectively mitigates object hallucination while maintaining competitive general-purpose multimodal capabilities.

Figures & tables

Explore similar work

CardsList
  1. See Only When Needed: Context-Aware Attention Intervention for Mitigating Hallucinations in LVLMs

    Jun 29, 2026Yuqing Lei, Wenbo Lyu, Yingjun Du +3Visual AttentionLarge Vision-Language Models

  2. Steer Where It Matters: Token-Level Visual-Sensitivity Steering for LVLMs Hallucination Mitigation

    Jun 2, 2026Ruipeng Zhang, Zhihao Li, C. L. Philip Chen +1VLM HallucinationLarge Vision-Language Models

  3. Look Clearly Before Answering: Mitigating Hallucinations in LVLMs via Saliency-Driven Perceptual Realignment

    Jul 18, 2026Pengxu Chen, Yao Zhu, Guangming Zhu +4Large Vision-Language ModelsVLM Hallucination Mitigation