Improving Neural Logic Machines via Failure Reflection
Zhiming Li, Yushi Cao, Yan Zheng, Xu Liu, Bozhi Wu, Tianlin Li, Xiufeng Xu, Junzhe Jiang, Yon Shin Teo, Shang-Wei Lin, Yang Liu
摘要
Reasoning is a fundamental ability towards artificial general intelligence (AGI). Fueled by the success of deep learning, the neural logic machines models (NLMs) have introduced novel neural-symbolic structures and demonstrate great performance and generalization on reasoning and decision-making tasks. However, the original training approaches of the NLMs are still far from perfect, the models would repeat similar mistakes during the training process which leads to sub-optimal performance. To mitigate this issue, we present a novel framework named Failure Reflection Guided Regularizer (FRGR). FRGR first dynamically identifies and summarizes the root cause if the model repeats similar mistakes during training. Then it penalizes the model if it makes similar mistakes in future training iterations. In this way, the model is expected to avoid repeating errors of similar root causes and converge faster to a betterperformed optimum. Experimental results on multiple relational reasoning and decision-making tasks demonstrate the effectiveness of FRGR in improving performance, generalization, training efficiency, and data efficiency. Our code is available at https://sites.google.com/ view/frgr-icml24.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper4
- Learning to Synthesize Programs as Interpretable and Generalizable PoliciesDweep Trivedi, Jesse Zhang, Shao-Hua Sun, Joseph J. LimNeurIPS 2021 · 被引用 104 次
- Provenance-guided synthesis of Datalog programsMukund Raghothaman, Jonathan Mendelson, David Zhao, Mayur Naik 等POPL 2020 · 被引用 49 次
- GALOIS: Boosting Deep Reinforcement Learning via Generalizable Logic SynthesisYushi Cao, Zhiming Li, Tianpei Yang, Hao Zhang 等NeurIPS 2022 · 被引用 23 次
- FAIRER: Fairness as Decision Rationale AlignmentTianlin Li, Qing Guo, Aishan Liu, Mengnan Du 等ICML 2023 · 被引用 20 次
相关 Paper
- Learning Robust Rule Representations for Abstract Reasoning via Internal InferencesWenbo Zhang, Likai Tang, Site Mo, Xianggen Liu 等NeurIPS 2022 · 被引用 7 次
- Adversarial Training for Process Reward ModelsGurusha Juneja, Deepak Nathani, William WangICML 2026 · 被引用 2 次
- Efficient Rectification of Neuro-Symbolic Reasoning Inconsistencies by Abductive ReflectionWen-Chao Hu, Wang-Zhou Dai, Yuan Jiang, Zhi-Hua ZhouAAAI 2025 · 被引用 14 次
- LogiGAN: Learning Logical Reasoning via Adversarial Pre-trainingXinyu Pi, Wanjun Zhong, Yan Gao, Nan Duan 等NeurIPS 2022 · 被引用 19 次
- Selection, Reflection and Self-Refinement: Revisit Reasoning Tasks via a Causal LensYunlong Deng, Boyang Sun, Yan Li, Zeyu Tang 等ICLR 2026 · 被引用 2 次
