Expectation-Guided Self-Verification for Aligning Large Reasoning Models with Domain Knowledge
Han Zhang, Chen Zhao, Shasha Wang, Gang Chen, Chenghong Zhang
2026Year
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 8897ccce-25b2-43a3-bd5f-2d4d3169fc5cRelated papers
- Thinking Out Loud: Do Reasoning Models Know When They're Right?Qingcheng Zeng, Weihao Xuan, Leyang Cui, Rob VoigtEMNLP 2025 · 1 citation
- Beyond Oracle: Verifier-Supervision for Instruction Hierarchy in Reasoning and Instruction-Tuned LLMsSian-Yao Huang, Li-Hsien Chang, Che-Yu Lin, Cheng-Lin YangNeurIPS 2025 · 4 citations
- Experience is the Best Teacher: Augmenting LLM Reasoning with Knowledge Learned from the PastMingjun Zhou, Weixin Zeng, Xiang ZhaoWWW 2026
- Incentivizing Agentic Reasoning Capability with Outcome Supervision for Knowledge Base Question AnsweringZhuo Chen, Fei Wang, Zixuan Li, Zhao Zhang et al.WWW 2026
- Entailer: Answering Questions with Faithful and Truthful Chains of ReasoningOyvind Tafjord, Bhavana Dalvi Mishra, Peter ClarkEMNLP 2022 · 28 citations
