Chain-of-Reasoning: Towards Unified Mathematical Reasoning in Large Language Models via a Multi-Paradigm Perspective
Yiyao Yu, Yuxiang Zhang, Dongdong Zhang, Xiao Liang, Hengyuan Zhang, Xingxing Zhang, Mahmoud Khademi, Hany Hassan Awadalla, Junjie Wang, Yujiu Yang, Furu Wei
Abstract
Large Language Models (LLMs) have made notable progress in mathematical reasoning, yet they often rely on single-paradigm reasoning that limits their effectiveness across diverse tasks. In this paper, we introduce Chain-of-Reasoning (CoR), a novel unified framework that integrates multiple reasoning paradigms -Natural Language Reasoning (NLR), Algorithmic Reasoning (AR), and Symbolic Reasoning (SR) -to enable synergistic collaboration. CoR generates multiple potential answers using different reasoning paradigms and synthesizes them into a coherent final solution. We propose a Progressive Paradigm Training (PPT) strategy that allows models to progressively master these paradigms, culminating in the development of CoR-Math-7B. Experimental results demonstrate that CoR-Math-7B significantly outperforms current SOTA models, achieving up to a 41.0% absolute improvement over GPT-4o in theorem proving tasks and a 15% improvement over RLbased methods on the MATH benchmark in arithmetic tasks. These results show the enhanced mathematical comprehensive ability of our model, enabling zero-shot generalization across tasks. The code is available at https: //github.com/microsoft/CoR .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1d0a3d14-858f-4ace-b6a4-5afedd821803Cited by top-tier papers10
- Beyond Pass@ 1: Self-Play with Variational Problem Synthesis Sustains RLVRXiao Liang, Zhong-Zhi Li, Yeyun Gong, Yelong Shen et al.ICLR 2026 · 57 citations
- SwS: Self-aware Weakness-driven Problem Synthesis in Reinforcement Learning for LLM ReasoningXiao Liang, Zhong-Zhi Li, Yeyun Gong, Yang Wang et al.NeurIPS 2025 · 41 citations
- Unlocking Multimodal Mathematical Reasoning via Process Reward ModelRuilin Luo, Zhuofan Zheng, Lei Wang, Yifan Wang et al.NeurIPS 2025 · 38 citations
- Learning to Reason via Mixture-of-Thought for Logical ReasoningTong Zheng, Lichang Chen, Simeng Han, R. Thomas McCoy et al.ICLR 2026 · 23 citations
- The Quest for Efficient Reasoning: A Data-Centric Benchmark to CoT DistillationRuichen Zhang, Rana Muhammad Shahroz Khan, Zhen Tan, Dawei Li et al.ICLR 2026 · 7 citations
Builds on15
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- Tree of Thoughts: Deliberate Problem Solving with Large Language ModelsShunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran et al.NeurIPS 2023 · 5,068 citations
- Graph of Thoughts: Solving Elaborate Problems with Large Language ModelsMaciej Besta, Nils Blach, Ales Kubicek, Robert Gerstenberger et al.AAAI 2024 · 1,292 citations
- ZeRO: memory optimizations toward training trillion parameter modelsSamyam Rajbhandari, Jeff Rasley, Olatunji Ruwase, Yuxiong HeSC 2020 · 852 citations
- Self-Consistency Improves Chain of Thought Reasoning in Language ModelsXuezhi Wang, Jason Wei, Dale Schuurmans, Quoc V. Le et al.ICLR 2023 · 681 citations
Related papers
- Parrot: A Training Pipeline Enhances Both Program CoT and Natural Language CoT for ReasoningSenjie Jin, Lu Chen, Zhiheng Xi, Yuhui Wang et al.EMNLP 2025
- Plan, Verify and Switch: Integrated Reasoning with Diverse X-of-ThoughtsTengxiao Liu, Qipeng Guo, Yuqing Yang, Xiangkun Hu et al.EMNLP 2023 · 7 citations
- UR² : Unify RAG and Reasoning through Reinforcement LearningWeitao Li, Boran Xiang, Xiaolong Wang, Jingyi Ren et al.ACL 2026 · 1 citation
- In-Token Rationality Optimization: Towards Accurate and Concise LLM Reasoning via Self-FeedbackMingye Zhu, Yi Liu, Zheren Fu, Quan Wang et al.AAAI 2026 · 1 citation
- PaCoRe: Learning to Scale Test-Time Compute with Parallel Coordinated ReasoningJingcheng Hu, Yinmin Zhang, Shijie Shang, Xiaobo Yang et al.ACL 2026 · 15 citations
