cs.LGJul 2, 2026

Bayesian Sparse Low-Rank Adaptation for Large Language Model Uncertainty Estimation

Authors: Jijie ZhangZhe RenQuan ZhangDandan Guo

Organizations: School of Artificial Intelligence, Jilin University · Michigan State University

Abstract

Large language models (LLMs) exhibit remarkable reasoning capabilities, but their task-specific fine-tuning is notoriously plagued by overconfidence, severely hindering trustworthy deployment. We propose Data-Adaptive Lower-Rank Adaptation (DALorRA), a simple and effective variational Bayesian sparse framework that shifts the paradigm of uncertainty quantification from the dense parameter space to the lightweight rank level of low-rank adaptation (LoRA). With the insight that LoRA essentially aggregates multiple rank-one components that may provide superfluous model capacity, DALorRA imposes stochastic masking on rank dimensions, enabling Bayesian regularization of model capacity during training and ensemble-like calibration during inference. Extensive experiments demonstrate DALorRA's excellent calibration of LLMs without compromising reasoning accuracy.

Explore similar work

CardsList
  1. Bayesian Fine-tuning in Projected Subspaces

    May 8, 2026Viktar Dubovik, Patryk Marszałek, Jacek Tabor +1Low-Rank AdaptationParameter-Efficient Fine-Tuning