TensorX
返回文献探索

Paper · arXiv 2410.14745

SemiEvol: Semi-supervised Fine-tuning for LLM Adaptation

Junyu Luo, Xiao Luo, Xiusi Chen, Zhiping Xiao, Wei Ju, Ming Zhang

47 upvotesOctober 17, 2024arXiv 预印本
AI 摘要

A semi-supervised fine-tuning framework named SemiEvol enhances LLM adaptation using both labeled and unlabeled data, showing improved performance through bi-level knowledge propagation and collaborative learning.

supervised fine-tuninglarge language modelssemi-supervised fine-tuningdata-efficient frameworkknowledge propagationbi-level approachin-weightin-contextknowledge selectioncollaborative learningpseudo-response sampleshybrid data scenarios

Abstract

Supervised fine-tuning (SFT) is crucial in adapting large language models (LLMs) to a specific domain or task. However, only a limited amount of labeled data is available in practical applications, which poses a severe challenge for SFT in yielding satisfactory results. Therefore, a data-efficient framework that can fully exploit labeled and unlabeled data for LLM fine-tuning is highly anticipated. Towards this end, we introduce a semi-supervised fine-tuning framework named SemiEvol for LLM adaptation from a propagate-and-select manner. For knowledge propagation, SemiEvol adopts a bi-level approach, propagating knowledge from labeled data to unlabeled data through both in-weight and in-context methods. For knowledge selection, SemiEvol incorporates a collaborative learning mechanism, selecting higher-quality pseudo-response samples. We conducted experiments using GPT-4o-mini and Llama-3.1 on seven general or domain-specific datasets, demonstrating significant improvements in model performance on target data. Furthermore, we compared SemiEvol with SFT and self-evolution methods, highlighting its practicality in hybrid data scenarios.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
SemiEvol: Semi-supervised Fine-tuning for LLM Adaptation | TensorX