TensorX
返回文献探索

Paper · arXiv 2405.06650

Large Language Models as Planning Domain Generators

James Oswald, Kavitha Srinivas, Harsha Kokel, Junkyu Lee, Michael Katz, Shirin Sohrabi

11 upvotesApril 2, 2024arXiv 预印本
AI 摘要

Large language models show moderate proficiency in generating automated planning domain models from natural language descriptions.

domain modelsAI planninglarge language models (LLMs)automated evaluationplanning domain modelsnatural language descriptions

Abstract

Developing domain models is one of the few remaining places that require manual human labor in AI planning. Thus, in order to make planning more accessible, it is desirable to automate the process of domain model generation. To this end, we investigate if large language models (LLMs) can be used to generate planning domain models from simple textual descriptions. Specifically, we introduce a framework for automated evaluation of LLM-generated domains by comparing the sets of plans for domain instances. Finally, we perform an empirical analysis of 7 large language models, including coding and chat models across 9 different planning domains, and under three classes of natural language domain descriptions. Our results indicate that LLMs, particularly those with high parameter counts, exhibit a moderate level of proficiency in generating correct planning domains from natural language descriptions. Our code is available at https://github.com/IBM/NL2PDDL.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
Large Language Models as Planning Domain Generators | TensorX