Towards Equipping Transformer with the Ability of Systematic Compositionality
Chen Huang, Peixin Qin, Wenqiang Lei, Jiancheng Lv
摘要
One of the key factors in language productivity and human cognition is the ability of Systematic Compositionality, which refers to understanding composed, unseen examples of seen primitives. However, recent evidence reveals that the Transformers have difficulty in generalizing the composed context based on the seen primitives. To this end, we take the first step to propose a compositionality-aware Transformer called CAT and two novel pre-training tasks to facilitate the systematic compositionality. We tentatively provide a successful implementation of a multi-layer CAT on the basis of the especially popular BERT. The experimental results demonstrate that CAT outperforms baselines on compositionality-aware tasks with minimal impact on effectiveness on standardized language understanding tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Can Large Language Models Understand Internet Buzzwords Through User-Generated ContentChen Huang, Junkai Luo, Xinzuo Wang, Wenqiang Lei 等ACL 2025
- Composition-Incremental Learning for Compositional GeneralizationZhen Li, Yuwei Wu, Chenchen Jing, Che Sun 等AAAI 2026
它引用的顶会 Paper11
- Faith and Fate: Limits of Transformers on CompositionalityNouha Dziri, Ximing Lu, Melanie Sclar, Xiang Lorraine Li 等NeurIPS 2023 · 被引用 728 次
- Measuring Compositional Generalization: A Comprehensive Method on Realistic DataDaniel Keysers, Nathanael Schärli, Nathan Scales, Hylke Buisman 等ICLR 2020 · 被引用 401 次
- Permutation Equivariant Models for Compositional Generalization in LanguageJonathan Gordon, David Lopez-Paz, Marco Baroni, Diane BouchacourtICLR 2020 · 被引用 112 次
- Assessing Phrasal Representation and Composition in TransformersLang Yu, Allyson EttingerEMNLP 2020 · 被引用 60 次
- Discrete-Valued Neural CommunicationDianbo Liu, Alex Lamb, Kenji Kawaguchi, Anirudh Goyal 等NeurIPS 2021 · 被引用 55 次
相关 Paper
- Dissecting Chain-of-Thought: Compositionality through In-Context Filtering and LearningYingcong Li, Kartik Sreenivasan, Angeliki Giannou, Dimitris Papailiopoulos 等NeurIPS 2023 · 被引用 12 次
- Compositional Task Representations for Large Language ModelsNan Shao, Zefan Cai, Hanwei Xu, Chonghua Liao 等ICLR 2023
- Inducing Transformer's Compositional Generalization Ability via Auxiliary Sequence Prediction TasksYichen Jiang, Mohit BansalEMNLP 2021
- Attention as a HypernetworkSimon Schug, Seijin Kobayashi, Yassir Akram, João Sacramento 等ICLR 2025
- How Do In-Context Examples Affect Compositional Generalization?Shengnan An, Zeqi Lin, Qiang Fu, Bei Chen 等ACL 2023 · 被引用 15 次
