MatSci-NLP: Evaluating Scientific Language Models on Materials Science Language Tasks Using Text-to-Schema Modeling
Yu Song, Santiago Miret, Bang Liu
摘要
We present MatSci-NLP, a natural language benchmark for evaluating the performance of natural language processing (NLP) models on materials science text. We construct the benchmark from publicly available materials science text data to encompass seven different NLP tasks, including conventional NLP tasks like named entity recognition and relation classification, as well as NLP tasks specific to materials science, such as synthesis action retrieval which relates to creating synthesis procedures for materials. We study various BERT-based models pretrained on different scientific text corpora on MatSci-NLP to understand the impact of pretraining strategies on understanding materials science text. Given the scarcity of high-quality annotated data in the materials science domain, we perform our fine-tuning experiments with limited training data to encourage the generalize across MatSci-NLP tasks. Our experiments in this low-resource training setting show that language models pretrained on scientific text outperform BERT trained on general text. Mat-BERT, a model pretrained specifically on materials science journals, generally performs best for most tasks. Moreover, we propose a unified text-to-schema for multitask learning on MatSci-NLP and compare its performance with traditional fine-tuning methods. In our analysis of different training methods, we find that our proposed text-to-schema methods inspired by question-answering consistently outperform single and multitask NLP fine-tuning methods. The code and datasets are publicly available 1 . * Equal contribution. † Corresponding author. Canada CIFAR AI Chair.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- LLaMP: Large Language Model Made Powerful for High-fidelity Materials Knowledge RetrievalYuan Chiang, Elvis Hsieh, Chia-Hong Chou, Janosh RiebesellEMNLP 2025 · 被引用 7 次
- MatExpert: Decomposing Materials Discovery By Mimicking Human ExpertsQianggang Ding, Santiago Miret, Bang LiuICLR 2025 · 被引用 3 次
- Zero-Shot Learning for Materials Science Texts: Leveraging Duck Typing PrinciplesXin Zhang, Peiliang Zhang, Jingling Yuan, Lin LiAAAI 2025 · 被引用 1 次
- ActionIE: Action Extraction from Scientific Literature with Programming LanguagesXianrui Zhong, Yufeng Du, Siru Ouyang, Ming Zhong 等ACL 2024
- SciCustom: A Framework for Custom Evaluation of Scientific Capabilities in Large Language ModelsYiyang Gu, Junwei Yang, Junyu Luo, Ye Yuan 等ACL 2026
它引用的顶会 Paper3
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- The SOFC-Exp Corpus and Neural Approaches to Information Extraction in the Materials Science DomainAnnemarie Friedrich, Heike Adel, Federico Tomazic, Johannes Hingerl 等ACL 2020 · 被引用 17 次
- Text2Event: Controllable Sequence-to-Structure Generation for End-to-end Event ExtractionYaojie Lu, Hongyu Lin, Jin Xu, Xianpei Han 等ACL 2021
相关 Paper
- Can Machines Read Coding Manuals Yet? - A Benchmark for Building Better Language Models for Code UnderstandingIbrahim Abdelaziz, Julian Dolby, Jamie P. McCusker, Kavitha SrinivasAAAI 2022 · 被引用 7 次
- A Multi-Task Learning Framework for Reading Comprehension of Scientific Tabular DataXu Yang, Meihui Zhang, Ju Fan, Zeyu Luo 等ICDE 2024 · 被引用 1 次
- MS-Mentions: Consistently Annotating Entity Mentions in Materials Science Procedural TextTim O'Gorman, Zach Jensen, Sheshera Mysore, Kevin Huang 等EMNLP 2021 · 被引用 13 次
- AtomWorld: A Benchmark for Evaluating Spatial Reasoning in Large Language Models on Material StructuresTaoyuze Lv, Alexander Chen, Fengyu Xie, Chu Wu 等ICML 2026 · 被引用 3 次
- Improving AMR Parsing with Sequence-to-Sequence Pre-trainingDongqin Xu, Junhui Li, Muhua Zhu, Min Zhang 等EMNLP 2020 · 被引用 57 次
