The Imitation Game: Turing Machine Imitator is Length Generalizable Reasoner

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Hua, Zhouqi, Zhang, Wenwei, Lyu, Chengqi, Gu, Yuzhe, Gao, Songyang, Liu, Kuikun, Lin, Dahua, Chen, Kai
Natura: Preprint
Pubblicazione: 2025
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866911403110039552
author Hua, Zhouqi
Zhang, Wenwei
Lyu, Chengqi
Gu, Yuzhe
Gao, Songyang
Liu, Kuikun
Lin, Dahua
Chen, Kai
author_facet Hua, Zhouqi
Zhang, Wenwei
Lyu, Chengqi
Gu, Yuzhe
Gao, Songyang
Liu, Kuikun
Lin, Dahua
Chen, Kai
contents Length generalization, the ability to solve problems of longer sequences than those observed during training, poses a core challenge of Transformer-based large language models (LLM). Although existing studies have predominantly focused on data-driven approaches for arithmetic operations and symbolic manipulation tasks, these approaches tend to be task-specific with limited overall performance. To pursue a more general solution, this paper focuses on a broader case of reasoning problems that are computable, i.e., problems that algorithms can solve, thus can be solved by the Turing Machine. From this perspective, this paper proposes Turing MAchine Imitation Learning (TAIL) to improve the length generalization ability of LLMs. TAIL synthesizes chain-of-thoughts (CoT) data that imitate the execution process of a Turing Machine by computer programs, which linearly expands the reasoning steps into atomic states to alleviate shortcut learning and explicit memory fetch mechanism to reduce the difficulties of dynamic and long-range data access in elementary operations. To validate the reliability and universality of TAIL, we construct a challenging synthetic dataset covering 8 classes of algorithms and 18 tasks. Without bells and whistles, TAIL significantly improves the length generalization ability as well as the performance of Qwen2.5-7B on various tasks using only synthetic data, surpassing previous methods and DeepSeek-R1. The experimental results reveal that the key concepts in the Turing Machine, instead of the thinking styles, are indispensable for TAIL for length generalization, through which the model exhibits read-and-write behaviors consistent with the properties of the Turing Machine in their attention layers. This work provides a promising direction for future research in the learning of LLM reasoning from synthetic data.
format Preprint
id arxiv_https___arxiv_org_abs_2507_13332
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle The Imitation Game: Turing Machine Imitator is Length Generalizable Reasoner
Hua, Zhouqi
Zhang, Wenwei
Lyu, Chengqi
Gu, Yuzhe
Gao, Songyang
Liu, Kuikun
Lin, Dahua
Chen, Kai
Computation and Language
Length generalization, the ability to solve problems of longer sequences than those observed during training, poses a core challenge of Transformer-based large language models (LLM). Although existing studies have predominantly focused on data-driven approaches for arithmetic operations and symbolic manipulation tasks, these approaches tend to be task-specific with limited overall performance. To pursue a more general solution, this paper focuses on a broader case of reasoning problems that are computable, i.e., problems that algorithms can solve, thus can be solved by the Turing Machine. From this perspective, this paper proposes Turing MAchine Imitation Learning (TAIL) to improve the length generalization ability of LLMs. TAIL synthesizes chain-of-thoughts (CoT) data that imitate the execution process of a Turing Machine by computer programs, which linearly expands the reasoning steps into atomic states to alleviate shortcut learning and explicit memory fetch mechanism to reduce the difficulties of dynamic and long-range data access in elementary operations. To validate the reliability and universality of TAIL, we construct a challenging synthetic dataset covering 8 classes of algorithms and 18 tasks. Without bells and whistles, TAIL significantly improves the length generalization ability as well as the performance of Qwen2.5-7B on various tasks using only synthetic data, surpassing previous methods and DeepSeek-R1. The experimental results reveal that the key concepts in the Turing Machine, instead of the thinking styles, are indispensable for TAIL for length generalization, through which the model exhibits read-and-write behaviors consistent with the properties of the Turing Machine in their attention layers. This work provides a promising direction for future research in the learning of LLM reasoning from synthetic data.
title The Imitation Game: Turing Machine Imitator is Length Generalizable Reasoner
topic Computation and Language
url https://arxiv.org/abs/2507.13332