cs.AISep 29, 2026

Beyond Low-Rank Parameterization: Narrowing the Gap Between LoRA and Full Fine-Tuning via Gradient Decomposition

Authors: Yihao Ouyang, Shiwei Li, Haozhao Wang, Xiandi Luo, Zhuoqi Hu, Jinglun Yu, Yichen Li, Ruixuan Li

Organizations: Huazhong University of Science and Technology, Wuhan, China · Hebei University of Technology, Tianjin, China

Abstract

Low-Rank Adaptation (LoRA) is a widely used approach to parameter-efficient fine-tuning (PEFT), yet a performance gap can remain relative to full fine-tuning (FFT). Many LoRA variants improve the initialization or optimization of low-rank factors. At each training step, however, their first-order weight-space directions are constrained by the current parameterization. We characterize the corresponding LoRA-accessible gradient space and show that it coincides with the tangent space induced by the current LoRA parameterization. This characterization yields an orthogonal decomposition of the full weight gradient at the current model parameters. We term the component orthogonal to this space the normal gradient. Based on this decomposition, we propose GDLoRA (Gradient-Decomposed Low-Rank Adaptation). GDLoRA reconstructs the full weight gradient from forward activations and backward signals, extracts its normal component, and directly updates the base weights with this component, while retaining standard AdamW optimization for the LoRA factors. GDLoRA incorporates complementary normal gradients without increasing standard LoRA's optimizer-state memory budget under matched adapter and optimizer configurations. Experiments on natural language understanding, mathematical reasoning, commonsense reasoning, and image classification show that GDLoRA consistently improves over LoRA and narrows the performance gap to FFT. The code is available at https://anonymous.4open.science/r/GDLoRA.

Figures & tables

Appendix figures & tables10 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Between Gradient and Natural Gradient: A Continuum of LoRA Initializations

    Jul 28, 2026Dianze Liu, Farshid GhezelbashModel Fine-TuningGradient

  2. SDS-LoRA: Overcoming Anisotropic Gradient Scaling in Low-Rank Adaptation

    Jun 15, 2026Junghun Oh, Sungyong Baik, Kyoung Mu LeeLow-Rank StructureHip-Lora

  3. LoRA-GA2^2: Low Rank Adaptation with Multi-step Gradient Adaptive Alignment

    Aug 20, 2026Haonan He, Xinyue FanMulti-LoraModel Fine-Tuning