LECTOR: Joint Learning of Scientific Reasoning Graphs and Introduction Generation
Jiabei Xiao, Yizhou Wang, Chen Tang, Pengze Li, Wanli Ouyang, SHIXIANG TANG
Abstract
AI Scientists have shown promising progress across multiple stages of the research pipeline, among which automatic scientific paper writing remains a formidable challenge. The Introduction writing is especially challenging, which demands not only linguistic fluency, but logical soundness and verifiable faithfulness. Most AI-assisted methods treat the task as text generation instead of reasoning and structuring, leading to severe drawbacks, e.g. , hallucinating citations. To address this, we first formulate the Content-Conditional Introduction Generation (CCIG) task, which requires grounding the Introduction in the paper's core evidence. We then propose LECTOR , a novel Logic-Expression Co-Reinforcement Learning framework that can strictly follow the scientist's logic, add high-quality citations and keep structured expressions. LECTOR first constructs a logic-reasoning graph from the paper's main body to serve as a verifiable logical blueprint. Subsequently, it employs a Logic-Expression Co-Rewarding mechanism to jointly optimize for both the graph's structural fidelity and the final narrative's quality. We conduct a dataset from Nature Communications papers to assess our method. Extensive experiments show consistent improvements in both logic fidelity and Introduction generation quality metrics, e.g. , Graph Quality (+26.7%) , Citation Quality (+8.6%) , and Paper Consistency (+3.3%) . Code and data are available at: https://github.com/Xiao-Youth/LECTOR
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on6
- AutoSurvey: Large Language Models Can Automatically Write SurveysYidong Wang, Qi Guo, Wenjin Yao, Hongbo Zhang et al.NeurIPS 2024 · 151 citations
- AI-Researcher: Autonomous Scientific InnovationJiabin Tang, Lianghao Xia, Zhonghang Li, Chao HuangNeurIPS 2025 · 101 citations
- LongBench: A Bilingual, Multitask Benchmark for Long Context UnderstandingYushi Bai, Xin Lv, Jiajie Zhang, Hongchang Lyu et al.ACL 2024 · 94 citations
- SciDQA: A Deep Reading Comprehension Dataset over Scientific PapersShruti Singh, Nandan Sarkar, Arman CohanEMNLP 2024 · 2 citations
- ARCHE: A Novel Task to Evaluate LLMs on Latent Reasoning Chain ExtractionPengze Li, Jiaqi Liu, Junchi Yu, Lihao Liu et al.AAAI 2026 · 1 citation
Related papers
- CiteGuard: Faithful Citation Attribution for LLMs via Retrieval-Augmented ValidationYee Man Choi, Xuehang Guo, Yi R. Fung, Qingyun WangACL 2026 · 7 citations
- ReviewRL: Towards Automated Scientific Review with RLSihang Zeng, Kai Tian, Kaiyan Zhang, Yuru Wang et al.EMNLP 2025
- HLM-Cite: Hybrid Language Model Workflow for Text-based Scientific Citation PredictionQianyue Hao, Jingyang Fan, Fengli Xu, Jian Yuan et al.NeurIPS 2024 · 23 citations
- Advancing Abductive Reasoning in Knowledge Graphs through Complex Logical Hypothesis GenerationJiaxin Bai, Yicheng Wang, Tianshi Zheng, Yue Guo et al.ACL 2024 · 5 citations
- Leveraging Outline-Optimized Generative Interactions and Critique for Self-Refining Outlines with Reinforcement LearningHengwei Liu, Haoyuan Ma, Qingqing Lyu, Daoxin Zhang et al.ACL 2026
