SG-Net: Syntax-Guided Machine Reading Comprehension
Zhuosheng Zhang, Yuwei Wu, Junru Zhou, Sufeng Duan, Hai Zhao, Rui Wang
Abstract
For machine reading comprehension, the capacity of effectively modeling the linguistic knowledge from the detail-riddled and lengthy passages and getting ride of the noises is essential to improve its performance. Traditional attentive models attend to all words without explicit constraint, which results in inaccurate concentration on some dispensable words. In this work, we propose using syntax to guide the text modeling by incorporating explicit syntactic constraints into attention mechanism for better linguistically motivated word representations. In detail, for self-attention network (SAN) sponsored Transformer-based encoder, we introduce syntactic dependency of interest (SDOI) design into the SAN to form an SDOI-SAN with syntax-guided self-attention. Syntax-guided network (SG-Net) is then composed of this extra SDOI-SAN and the SAN from the original Transformer encoder through a dual contextual architecture for better linguistics inspired representation. To verify its effectiveness, the proposed SG-Net is applied to typical pre-trained language model BERT which is right based on a Transformer encoder. Extensive experiments on popular benchmarks including SQuAD 2.0 and RACE show that the proposed SG-Net design helps achieve substantial performance improvement over strong baselines.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f4ba4a64-4e20-47ed-a19e-960a3c24e4deCited by top-tier papers24
- Semantics-Aware BERT for Language UnderstandingZhuosheng Zhang, Yuwei Wu, Hai Zhao, Zuchao Li et al.AAAI 2020 · 396 citations
- Retrospective Reader for Machine Reading ComprehensionZhuosheng Zhang, Junjie Yang, Hai ZhaoAAAI 2021 · 237 citations
- Neural Module Networks for Reasoning over TextNitish Gupta, Kevin Lin, Dan Roth, Sameer Singh et al.ICLR 2020 · 134 citations
- Topic-Aware Multi-turn Dialogue ModelingYi Xu, Hai Zhao, Zhuosheng ZhangAAAI 2021 · 93 citations
- Bipartite Flat-Graph Network for Nested Named Entity RecognitionYing Luo, Hai ZhaoACL 2020 · 84 citations
Related papers
- Learning Disentangled Semantic Representations for Zero-Shot Cross-Lingual Transfer in Multilingual Machine Reading ComprehensionLinjuan Wu, Shaojuan Wu, Xiaowang Zhang, Deyi Xiong et al.ACL 2022 · 18 citations
- Zero-Shot Cross-Lingual Machine Reading Comprehension via Inter-sentence Dependency GraphLiyan Xu, Xuchao Zhang, Bo Zong, Yanchi Liu et al.AAAI 2022 · 5 citations
- Saliency-Guided Attention Network for Image-Sentence MatchingZhong Ji, Haoran Wang, Jungong Han, Yanwei PangICCV 2019 · 96 citations
- Semantics-Aware Inferential Network for Natural Language UnderstandingShuailiang Zhang, Hai Zhao, Junru Zhou, Xi Zhou et al.AAAI 2021 · 4 citations
- Syntactically Look-Ahead Attention Network for Sentence CompressionHidetaka Kamigaito, Manabu OkumuraAAAI 2020 · 22 citations
