cs.LGOct 5, 2026

RoSA: Rotational Sparse Adaptation for Memory-Efficient Fine-Tuning

Authors: Muhammad Azeem Lodhi, Chao Zhou, Rebekka Burkholz

Organizations: Saarland University, Saarbrücken, Germany · CISPA Helmholtz Center for Information Security, Saarbrücken, Germany

Abstract

Parameter-efficient fine-tuning (PEFT) reduces the cost of adapting foundation models by focusing training on a small parameter subset. Complementary to this idea, we introduce RoSA (Rotational Sparse Adaptation), which narrows adaptation to a subset of layers at a time. RoSA freezes lower layers close to the input throughout training and rotates a trainable block over later layers, progressively increasing the number of frozen layers close to the input. This design reduces optimizer-state memory, shortens backpropagation, and even forward propagation if activations at the last frozen layer are cached. Because RoSA is orthogonal to the choice of trainable parameterization, it can be combined with PEFT methods or sparse optimizers within each active block. Experiments across multiple LLM architectures and tasks show that RoSA reduces peak memory while maintaining strong fine-tuning performance.

Figures & tables

Appendix figures & tables5 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. FuRA: Full-Rank Parameter-Efficient Fine-Tuning with Spectral Preconditioning

    May 19, 2026Yequan Zhao, Ruijie Zhang, Liyan Tan +3Parameter-Efficient Fine-Tuning MethodsModel Fine-Tuning

  2. CARE-LoRA: Compressed Activation REconstruction for Memory-Efficient LoRA

    Jul 11, 2026Gengyu Zhang, Haiyin Ran, Zhengbao He +4Low-Rank Adaptation FrameworkLarge Language Model Compression