Reforming the Mechanism: Editing Reasoning Patterns in LLMs with Circuit Reshaping
Zhenyu Lei, Qiong Wu, JIANXIONG DONG, Yinhan He, Emily Dodwell, Yushun Dong, Jundong Li
Abstract
Large language models (LLMs) often exhibit flawed reasoning ability that undermines reliability. Existing approaches to improving reasoning typically treat it as a general and monolithic skill, applying broad training which is inefficient and unable to target specific reasoning errors. We introduce Reasoning Editing, a paradigm for selectively modifying specific reasoning patterns in LLMs while preserving other reasoning pathways. This task presents a fundamental trade-off between Generality, the ability of an edit to generalize across different tasks sharing the same reasoning pattern, and Locality, the ability to preserve other reasoning capabilities. Through systematic investigation, we uncover the Circuit-Interference Law: Edit interference between reasoning patterns is proportional to the overlap of their neural circuits. Guided by this principle, we propose REdit, the first framework to actively reshape neural circuits before editing, thereby modulating interference between reasoning patterns and mitigating the trade-off. REdit integrates three components: (i) Contrastive Circuit Reshaping, which directly addresses the generality-locality trade-off by disentangling overlapping circuits; (ii) Meta-Contrastive Learning, which extends transferability to novel reasoning patterns; and (iii) Dual-Level Protection, which preserves preexisting abilities by constraining reshaping update directions and regularizing tasklevel predictions. Extensive experiments with Qwen-2.5-3B on propositional logic reasoning tasks across three difficulty levels demonstrate that REdit consistently achieves superior generality and locality compared to baselines, with additional validation in mathematics showing broader potential. Our code is available at https://github.com/LzyFischer/REdit . * This work was initiated and completed while Zhenyu was an intern with AT&T CDO. Both Qiong Wu and Jundong Li are the corresponding authors.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 80e05abb-08d6-41b0-be37-174d49bf8d24Builds on32
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Locating and Editing Factual Associations in GPTKevin Meng, David Bau, Alex Andonian, Yonatan BelinkovNeurIPS 2022 · 3,415 citations
- Fast Model Editing at ScaleEric Mitchell, Charles Lin, Antoine Bosselut, Chelsea Finn et al.ICLR 2022 · 527 citations
- Memory-Based Model Editing at ScaleEric Mitchell, Charles Lin, Antoine Bosselut, Christopher D. Manning et al.ICML 2022 · 520 citations
- Does Localization Inform Editing? Surprising Differences in Causality-Based Localization vs. Knowledge Editing in Language ModelsPeter Hase, Mohit Bansal, Been Kim, Asma GhandehariounNeurIPS 2023 · 307 citations
Related papers
- CaKE: Circuit-aware Editing Enables Generalizable Knowledge LearnersYunzhi Yao, Jizhan Fang, Jia-Chen Gu, Ningyu Zhang et al.EMNLP 2025 · 1 citation
- Keys to Robust Edits: From Theoretical Insights to Practical AdvancesJianhao Yan, Futing Wang, Yun Luo, Yafu Li et al.ACL 2025 · 3 citations
- Correcting in Hindsight: Editing Past Key-Value States for Robust LLM ReasoningMengfei Zhang, Yu Mi, Leijing ZhouICML 2026
- MIND: Multi-rationale INtegrated Discriminative Reasoning Framework for Multi-modal Large ModelsChuang Yu, Jinmiao Zhao, Mingxuan Zhao, Yunpeng Liu et al.ICML 2026
- Tokens to Types: Context Editing with Selective Entity Abstraction for Grounded GenerationRounak Sharma, Debabrata Mahapatra, Shiv Kumar SainiSIGIR 2026
