HypER: Literature-grounded Hypothesis Generation and Distillation with Provenance
Rosni Vasu, Chandrayee Basu, Bhavana Dalvi Mishra, Cristina Sarasua, Peter Clark, Abraham Bernstein
摘要
Large Language models have demonstrated promising performance in research ideation across scientific domains. Hypothesis development, the process of generating a highly specific declarative statement connecting a research idea with empirical validation, has received relatively less attention. Existing approaches trivially deploy retrieval augmentation and focus only on the quality of the final output ignoring the underlying reasoning process behind ideation. We present HypER (Hypothesis Generation with Explanation and Reasoning), a small language model (SLM) trained for literature-guided reasoning and evidence-based hypothesis generation. HypER is trained in a multi-task setting to discriminate between valid and invalid scientific reasoning chains in presence of controlled distractions. We find that HypER outperforms the base model, distinguishing valid from invalid reasoning chains (+22% average absolute F1), generates better evidence-grounded hypotheses (0.327 vs. 0.305 base model) with high feasibility and impact as judged by human experts (>3.5 on 5-point Likert scale). Resource at .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- AstaBench: Rigorous Benchmarking of AI Agents with a Scientific Research SuiteJonathan Bragg, Mike D'Arcy, Nishant Balepur, Dan Bareket 等ICLR 2026 · 被引用 51 次
- From Automation to Autonomy: A Survey on Large Language Models in Scientific DiscoveryTianshi Zheng, Zheye Deng, Hong Ting Tsang, Weiqi Wang 等EMNLP 2025 · 被引用 5 次
- Generating Literature-Driven Scientific Theories at ScalePeter Jansen, Peter Clark, Doug Downey, Daniel S. WeldACL 2026 · 被引用 3 次
它引用的顶会 Paper6
- Selective Annotation Makes Language Models Better Few-Shot LearnersHongjin Su, Jungo Kasai, Chen Henry Wu, Weijia Shi 等ICLR 2023 · 被引用 63 次
- SciMON: Scientific Inspiration Machines Optimized for NoveltyQingyun Wang, Doug Downey, Heng Ji, Tom HopeACL 2024 · 被引用 22 次
- Exploring and Verbalizing Academic Ideas by Concept Co-occurrenceYi Xu, Shuqian Sheng, Bo Xue, Luoyi Fu 等ACL 2023 · 被引用 5 次
- Related Work and Citation Text Generation: A SurveyXiangci Li, Jessica OuyangEMNLP 2024 · 被引用 3 次
- DiscoveryBench: Towards Data-Driven Discovery with Large Language ModelsBodhisattwa Prasad Majumder, Harshit Surana, Dhruv Agarwal, Bhavana Dalvi Mishra 等ICLR 2025
相关 Paper
- Hypothesis-Driven Reasoning for Large Language ModelsAakash Kumar Agarwal, Moyuru YamadaAAAI 2026
- IdeaSynth: Iterative Research Idea Development Through Evolving and Composing Idea Facets with Literature-Grounded FeedbackKevin Pu, K. J. Kevin Feng, Tovi Grossman, Tom Hope 等CHI 2025 · 被引用 15 次
- On the Role of Model Prior in Real-World Inductive ReasoningZhuo Liu, Ding Yu, Hangfeng HeEMNLP 2025
- Literature Meets Data: A Synergistic Approach to Hypothesis GenerationHaokun Liu, Yangqiaoyu Zhou, Mingxuan Li, Chenfei Yuan 等ACL 2025 · 被引用 17 次
- DeepRAG: Thinking to Retrieve Step by Step for Large Language ModelsXinyan Guan, Jiali Zeng, Fandong Meng, Chunlei Xin 等ICLR 2026 · 被引用 30 次
