TensorX
返回文献探索

Paper · arXiv 2507.12142

RiemannLoRA: A Unified Riemannian Framework for Ambiguity-Free LoRA Optimization

Vladimir Bogachev, Vladimir Aletov, Alexander Molozhavenko, Denis Bobkov, Vera Soboleva, Aibek Alanov, Maxim Rakhuba

36 upvotesJuly 16, 2025arXiv 预印本
AI 摘要

RiemannLoRA addresses initialization and overparametrization in LoRA by treating LoRA matrices as a smooth manifold, improving convergence speed and performance in LLMs and diffusion models.

Low-Rank AdaptationLoRAparameter-efficient fine-tuninglarge language modelslow-rank matrix factorizationsmooth manifoldRiemannian optimizationRiemannLoRAconvergence speedfinal performancediffusion models

Abstract

Low-Rank Adaptation (LoRA) has become a widely adopted standard for parameter-efficient fine-tuning of large language models (LLMs), significantly reducing memory and computational demands. However, challenges remain, including finding optimal initialization strategies or mitigating overparametrization in low-rank matrix factorization. In this work, we propose a novel approach that addresses both of the challenges simultaneously within a unified framework. Our method treats a set of fixed-rank LoRA matrices as a smooth manifold. Considering adapters as elements on this manifold removes overparametrization, while determining the direction of the fastest loss decrease along the manifold provides initialization. Special care is taken to obtain numerically stable and computationally efficient implementation of our method, using best practices from numerical linear algebra and Riemannian optimization. Experimental results on LLM and diffusion model architectures demonstrate that RiemannLoRA consistently improves both convergence speed and final performance over standard LoRA and its state-of-the-art modifications.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
RiemannLoRA: A Unified Riemannian Framework for Ambiguity-Free LoRA Optimization | TensorX