TensorX
返回文献探索

Paper · arXiv 2503.02682

MPO: Boosting LLM Agents with Meta Plan Optimization

Weimin Xiong, Yifan Song, Qingxiu Dong, Bingchan Zhao, Feifan Song, Xun Wang, Sujian Li

29 upvotesMarch 4, 2025arXiv 预印本
AI 摘要

MPO framework enhances LLM-based agent planning with explicit high-level guidance, improving task completion and generalization through meta plans.

large language modelsLLM-based agentsMeta Plan OptimizationMPOplanning hallucinationsmeta plansfeedbacktask executiongeneralization

Abstract

Recent advancements in large language models (LLMs) have enabled LLM-based agents to successfully tackle interactive planning tasks. However, despite their successes, existing approaches often suffer from planning hallucinations and require retraining for each new agent. To address these challenges, we propose the Meta Plan Optimization (MPO) framework, which enhances agent planning capabilities by directly incorporating explicit guidance. Unlike previous methods that rely on complex knowledge, which either require significant human effort or lack quality assurance, MPO leverages high-level general guidance through meta plans to assist agent planning and enables continuous optimization of the meta plans based on feedback from the agent's task execution. Our experiments conducted on two representative tasks demonstrate that MPO significantly outperforms existing baselines. Moreover, our analysis indicates that MPO provides a plug-and-play solution that enhances both task completion efficiency and generalization capabilities in previous unseen scenarios.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
MPO: Boosting LLM Agents with Meta Plan Optimization | TensorX