TensorX
返回文献探索

Paper · arXiv 2402.04494

Grandmaster-Level Chess Without Search

Anian Ruoss, Grégoire Delétang, Sourabh Medapati, Jordi Grau-Moya, Li Kevin Wenliang, Elliot Catt, John Reid, Tim Genewein

70 upvotesFebruary 7, 2024arXiv 预印本
AI 摘要

A large-scale transformer model trained on a vast dataset of chess games outperforms traditional chess engines and other state-of-the-art models without domain-specific tweaks.

transformer modelsupervised learningaction-valuesLichess blitz Elochess puzzlesAlphaZeroMCTSGPT-3.5-turbo-instructmodel sizedataset size

Abstract

The recent breakthrough successes in machine learning are mainly attributed to scale: namely large-scale attention-based architectures and datasets of unprecedented scale. This paper investigates the impact of training at scale for chess. Unlike traditional chess engines that rely on complex heuristics, explicit search, or a combination of both, we train a 270M parameter transformer model with supervised learning on a dataset of 10 million chess games. We annotate each board in the dataset with action-values provided by the powerful Stockfish 16 engine, leading to roughly 15 billion data points. Our largest model reaches a Lichess blitz Elo of 2895 against humans, and successfully solves a series of challenging chess puzzles, without any domain-specific tweaks or explicit search algorithms. We also show that our model outperforms AlphaZero's policy and value networks (without MCTS) and GPT-3.5-turbo-instruct. A systematic investigation of model and dataset size shows that strong chess performance only arises at sufficient scale. To validate our results, we perform an extensive series of ablations of design choices and hyperparameters.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
Grandmaster-Level Chess Without Search | TensorX