Aligning Superhuman AI with Human Behavior: Chess as a Model System
Reid McIlroy-Young, Siddhartha Sen, Jon M. Kleinberg, Ashton Anderson
Abstract
As artificial intelligence becomes increasingly intelligent-in some cases, achieving superhuman performance-there is growing potential for humans to learn from and collaborate with algorithms. However, the ways in which AI systems approach problems are often different from the ways people do, and thus may be uninterpretable and hard to learn from. A crucial step in bridging this gap between human and artificial intelligence is modeling the granular actions that constitute human behavior, rather than simply matching aggregate human performance.
We pursue this goal in a model system with a long history in artificial intelligence: chess. The aggregate performance of a chess player unfolds as they make decisions over the course of a game. The hundreds of millions of games played online by players at every skill level form a rich source of data in which these decisions, and their exact context, are recorded in minute detail. Applying existing chess engines to this data, including an open-source implementation of AlphaZero, we find that they do not predict human moves well.
We develop and introduce Maia, a customized version of Alpha-Zero trained on human chess games, that predicts human moves at a much higher accuracy than existing engines, and can achieve maximum accuracy when predicting decisions made by players at a specific skill level in a tuneable way. For a dual task of predicting whether a human will make a large mistake on the next move, we develop a deep neural network that significantly outperforms competitive baselines. Taken together, our results suggest that there is substantial promise in designing artificial intelligence systems with human collaboration in mind by first accurately modeling granular human decision-making.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ff9bad98-fff7-48c7-9527-5e80e42ff9a8Cited by top-tier papers16
- Can You Learn an Algorithm? Generalizing from Easy to Hard Problems with Recurrent NetworksAvi Schwarzschild, Eitan Borgnia, Arjun Gupta, Furong Huang et al.NeurIPS 2021 · 133 citations
- Modeling Strong and Human-Like Gameplay with KL-Regularized SearchAthul Paul Jacob, David J. Wu, Gabriele Farina, Adam Lerer et al.ICML 2022 · 69 citations
- Imitation Learning by Estimating Expertise of DemonstratorsMark Beliaev, Andy Shih, Stefano Ermon, Dorsa Sadigh et al.ICML 2022 · 60 citations
- Amortized Planning with Large-Scale Transformers: A Case Study on ChessAnian Ruoss, Grégoire Delétang, Sourabh Medapati, Jordi Grau-Moya et al.NeurIPS 2024 · 57 citations
- Maia-2: A Unified Model for Human-AI Alignment in ChessZhenwei Tang, Difan Jiao, Reid McIlroy-Young, Jon M. Kleinberg et al.NeurIPS 2024 · 39 citations
Related papers
- Learning Models of Individual Behavior in ChessReid McIlroy-Young, Russell Wang, Siddhartha Sen, Jon M. Kleinberg et al.KDD 2022 · 17 citations
- Human-Aligned Chess With a Bit of SearchYiming Zhang, Athul Paul Jacob, Vivian Lai, Daniel Fried et al.ICLR 2025 · 1 citation
- Designing Skill-Compatible AI: Methodologies and Frameworks in ChessKarim Hamade, Reid McIlroy-Young, Siddhartha Sen, Jon M. Kleinberg et al.ICLR 2024 · 12 citations
- Chessformer: A Unified Architecture for Chess ModelingDaniel Monroe, George Eilender, Philip Chalmers, Zhenwei Tang et al.ICLR 2026 · 8 citations
- Detecting Individual Decision-Making Style: Exploring Behavioral Stylometry in ChessReid McIlroy-Young, Yu Wang, Siddhartha Sen, Jon M. Kleinberg et al.NeurIPS 2021 · 38 citations
