TensorX
返回文献探索

Paper · arXiv 2410.08102

Multi-Agent Collaborative Data Selection for Efficient LLM Pretraining

Tianyi Bai, Ling Yang, Zhen Hao Wong, Jiahui Peng, Xinlin Zhuang, Chi Zhang, Lijun Wu, Qiu Jiantao, Wentao Zhang, Binhang Yuan, Conghui He

21 upvotesOctober 10, 2024arXiv 预印本
AI 摘要

A multi-agent collaborative mechanism enhances data selection for pretraining large language models, improving efficiency and performance.

multi-agentcollaborative data selectionlarge language modelsLLM pretrainingdata efficiencyconvergenceperformance gain

Abstract

Efficient data selection is crucial to accelerate the pretraining of large language models (LLMs). While various methods have been proposed to enhance data efficiency, limited research has addressed the inherent conflicts between these approaches to achieve optimal data selection for LLM pretraining. To tackle this problem, we propose a novel multi-agent collaborative data selection mechanism. In this framework, each data selection method serves as an independent agent, and an agent console is designed to dynamically integrate the information from all agents throughout the LLM training process. We conduct extensive empirical studies to evaluate our multi-agent framework. The experimental results demonstrate that our approach significantly improves data efficiency, accelerates convergence in LLM training, and achieves an average performance gain of 10.5% across multiple language model benchmarks compared to the state-of-the-art methods.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
Multi-Agent Collaborative Data Selection for Efficient LLM Pretraining | TensorX