TensorX
返回文献探索

Paper · arXiv 2403.13372

LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models

Yaowei Zheng, Richong Zhang, Junhao Zhang, Yanhan Ye, Zheyan Luo

187 upvotesMarch 20, 2024arXiv 预印本
AI 摘要

LlamaFactory is a unified framework enabling efficient fine-tuning of large language models across various tasks using a web-based user interface.

efficient fine-tuninglarge language modelsLLaMALlamaFactoryLlamaBoardlanguage modelingtext generation

Abstract

Efficient fine-tuning is vital for adapting large language models (LLMs) to downstream tasks. However, it requires non-trivial efforts to implement these methods on different models. We present LlamaFactory, a unified framework that integrates a suite of cutting-edge efficient training methods. It allows users to flexibly customize the fine-tuning of 100+ LLMs without the need for coding through the built-in web UI LlamaBoard. We empirically validate the efficiency and effectiveness of our framework on language modeling and text generation tasks. It has been released at https://github.com/hiyouga/LLaMA-Factory and already received over 13,000 stars and 1,600 forks.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models | TensorX