TensorX
返回文献探索

Paper · arXiv 2307.13269

LoraHub: Efficient Cross-Task Generalization via Dynamic LoRA Composition

Chengsong Huang, Qian Liu, Bill Yuchen Lin, Tianyu Pang, Chao Du, Min Lin

34 upvotesJuly 25, 2023arXiv 预印本
AI 摘要

LoraHub, a framework for combining LoRA modules, enables few-shot cross-task generalization and facilitates community sharing of LRA modules for large language models.

Low-rank adaptationsLoRAlarge language modelsLLMscross-task generalizationLoraHubBig-Bench HardBBH benchmarkfew-shot scenariosin-context learninggeneral intelligence

Abstract

Low-rank adaptations (LoRA) are often employed to fine-tune large language models (LLMs) for new tasks. This paper investigates LoRA composability for cross-task generalization and introduces LoraHub, a strategic framework devised for the purposive assembly of LoRA modules trained on diverse given tasks, with the objective of achieving adaptable performance on unseen tasks. With just a few examples from a novel task, LoraHub enables the fluid combination of multiple LoRA modules, eradicating the need for human expertise. Notably, the composition requires neither additional model parameters nor gradients. Our empirical results, derived from the Big-Bench Hard (BBH) benchmark, suggest that LoraHub can effectively mimic the performance of in-context learning in few-shot scenarios, excluding the necessity of in-context examples alongside each inference input. A significant contribution of our research is the fostering of a community for LoRA, where users can share their trained LoRA modules, thereby facilitating their application to new tasks. We anticipate this resource will widen access to and spur advancements in general intelligence as well as LLMs in production. Code will be available at https://github.com/sail-sg/lorahub.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号