Toward Adaptive Reasoning in Large Language Models with Thought Rollback
Sijia Chen, Baochun Li
Abstract
Large language models (LLMs) have been routinely used to solve various tasks using step-bystep reasoning. However, the structure of intermediate reasoning steps, or thoughts, is rigid and unidirectional, such as chains, trees, or acyclicdirected graphs. Consequently, the resulting inflexible and forward-only reasoning may not address challenging tasks and fail when the LLM frequently gives false responses, i.e., "hallucinations". This paper proposes a new reasoning framework, called Thought Rollback (TR), allowing LLMs to adaptively build thought structure while maintaining effective reasoning toward problem-solving under "hallucinations". The core mechanism of TR is rolling back thoughts, which allows LLMs to perform error analysis on thoughts, and thus roll back to any previously mistaken thought for revision. Subsequently, by including such trial-and-error in the prompt to guide the LLM, each rollback leads to one more reliable reasoning path. Therefore, starting with a simple prompt without human annotations, LLM with TR adaptively and gradually explores thoughts for a correct solution. Comprehensive experiments on mathematical problems and multi-task reasoning demonstrate the state-of-the-art performance of TR in terms of problem-solving rate and interaction cost. For instance, the solving rate of GPT-4 with TR outperforms the current best by 9% on the MATH dataset. The source code is available under the folder examples/ThoughtRollback of https:// github.com/iQua/llmpebase .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 67cc5d79-979f-471e-8af3-fff02e896ffaCited by top-tier papers9
- Counterfactual VLA: Self-Reflective Vision-Language-Action Model with Adaptive ReasoningZhenghao Peng, Wenhao Ding, Yurong You, Yuxiao Chen et al.CVPR 2026 · 25 citations
- Rethinking Optimal Verification Granularity for Compute-Efficient Test-Time ScalingHao Mark Chen, Guanxi Lu, Yasuyuki Okoshi, Zhiwen Mo et al.NeurIPS 2025 · 8 citations
- AdaSTaR: Adaptive Data Sampling for Training Self-Taught ReasonersReiss Koh, Wonbeen Oh, Jaein Jang, Minhyung Lee et al.NeurIPS 2025 · 8 citations
- Hydra-Nav: Object Navigation via Adaptive Dual-Process ReasoningZixuan Wang, Huang Fang, Shaoan Wang, Yuanfei Luo et al.ICML 2026 · 4 citations
- Generator-Assistant Stepwise Rollback Framework for Large Language Model AgentXingzuo Li, Kehai Chen, Yunfei Long, Xuefeng Bai et al.EMNLP 2025 · 1 citation
Builds on21
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- Large Language Models are Zero-Shot ReasonersTakeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo et al.NeurIPS 2022 · 8,168 citations
- Measuring Massive Multitask Language UnderstandingDan Hendrycks, Collin Burns, Steven Basart, Andy Zou et al.ICLR 2021 · 7,905 citations
- Tree of Thoughts: Deliberate Problem Solving with Large Language ModelsShunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran et al.NeurIPS 2023 · 5,068 citations
Related papers
- Boosting of Thoughts: Trial-and-Error Problem Solving with Large Language ModelsSijia Chen, Baochun Li, Di NiuICLR 2024 · 24 citations
- Hallucination Detection from Structural Reasoning ModelJianbo Sun, Pengkun YangICML 2026
- Thought Propagation: an Analogical Approach to Complex Reasoning with Large Language ModelsJunchi Yu, Ran He, Zhitao YingICLR 2024 · 44 citations
- Reversal of Thought: Enhancing Large Language Models with Preference-Guided Reverse Reasoning Warm-upJiahao Yuan, Dehui Du, Hao Zhang, Zixiang Di et al.ACL 2025 · 12 citations
- Reasoning-as-Logic-Units: Scaling Test-Time Reasoning in Large Language Models Through Logic Unit AlignmentCheryl Li, Tianyuan Xu, Steven Y. GuoICML 2025
