Out-of-Distribution Evaluation of Rule-Based and Strategic Reasoning in Chess Transformers
Modern decision transformers, trained similarly to LLMs, can achieve strong in-distribution performance in complex sequential domains like chess, but it remains unclear to what extent they reason systematically about rules and strategy. We study the reasoning capabilities of a 270M-parameter chess t…