cs.CVSep 24, 2026

Domain Recentering and Confidence-Weighted Prior Calibration for Vision-Language Models

Authors: Youngeun Seol, Jimin Shin, Heeseo Yoon, Uiwon Hwang

Organizations: Department of Computer Science and Engineering, Ewha Womans University

Abstract

Vision-language models such as CLIP achieve strong zero-shot classification, yet under distribution shift, visual embeddings drift from fixed text embeddings. Training-free calibration avoids the per-sample optimization of prompt learning, but prior feature calibration gives each image the full bias of one hard cluster. We propose Domain Recentering with Confidence Calibration (DRC), a training-free method adapting CLIP from a set of unlabeled target images. DRC fits a Gaussian mixture once and subtracts from each embedding a posterior-weighted average of component means. It then removes residual class preference with a log-prior correction, estimating the prior from confidence-weighted predictions. Among compared methods, DRC achieves the highest average accuracy on cross-domain datasets, exceeding zero-shot CLIP by 4.13 and 5.07 points with ViT-B/16 and ResNet-50, with gains over CLIP also holding under ImageNet distribution shifts.

Figures & tables

Explore similar work

CardsList
  1. Improving CLIP Adaptation by Breaking Tail Alignment for Source-Free Cross-Domain Few-Shot Learning

    May 28, 2026Shuai Yi, Yixiong Zou, Yuhua Li +1Vision-Language Model AdaptationContrastive Language-Image Pre-Training Model

  2. Respect Your Zero-Shot Uncertainty: Conservative Calibration for Test-Time-Adapted Vision-Language Models

    Aug 6, 2026Jingyan Jiang, Yaru Sun, Xiao Chen +5Vision-Language Model AdaptationZero-Shot Robustness

  3. CoDoL: Conditional Domain Prompt Learning for Out-of-Distribution Generalization

    Sep 18, 2025Min Zhang, Yuyin Wang, Zhongxiang Dai +4Contrastive Language-Image Pre-Training ModelMultiple Prompt Learning Frameworks