Hierarchical Phrase-Based Sequence-to-Sequence Learning
Bailin Wang, Ivan Titov, Jacob Andreas, Yoon Kim
摘要
We describe a neural transducer that maintains the flexibility of standard sequence-to-sequence (seq2seq) models while incorporating hierarchical phrases as a source of inductive bias during training and as explicit constraints during inference. Our approach trains two models: a discriminative parser based on a bracketing transduction grammar whose derivation tree hierarchically aligns source and target phrases, and a neural seq2seq model that learns to translate the aligned phrases one-by-one. We use the same seq2seq model to translate at all phrase scales, which results in two inference modes: one mode in which the parser is discarded and only the seq2seq component is used at the sequence-level, and another in which the parser is combined with the seq2seq model. Decoding in the latter mode is done with the cube-pruned CKY algorithm, which is more involved but can make use of new translation rules during inference. We formalize our model as a sourceconditioned synchronous grammar and develop an efficient variational inference algorithm for training. When applied on top of both randomly initialized and pretrained seq2seq models, we find that both inference modes performs well compared to baselines on small scale machine translation benchmarks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Grammar Prompting for Domain-Specific Language Generation with Large Language ModelsBailin Wang, Zi Wang, Xuezhi Wang, Yuan Cao 等NeurIPS 2023 · 被引用 138 次
- Non-autoregressive Machine Translation with Probabilistic Context-free GrammarShangtong Gui, Chenze Shao, Zhengrui Ma, Xishan Zhang 等NeurIPS 2023 · 被引用 16 次
- Unsupervised Discontinuous Constituency Parsing with Mildly Context-Sensitive GrammarsSonglin Yang, Roger Levy, Yoon KimACL 2023 · 被引用 1 次
它引用的顶会 Paper10
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- COGS: A Compositional Generalization Challenge Based on Semantic InterpretationNajoung Kim, Tal LinzenEMNLP 2020 · 被引用 149 次
- Sequence-to-Sequence Learning with Latent Neural GrammarsYoon KimNeurIPS 2021 · 被引用 44 次
- Learning to Recombine and Resample Data For Compositional GeneralizationEkin Akyürek, Afra Feyza Akyürek, Jacob AndreasICLR 2021 · 被引用 36 次
- Visually Grounded Compound PCFGsYanpeng Zhao, Ivan TitovEMNLP 2020 · 被引用 35 次
相关 Paper
- Structured Reordering for Modeling Latent Alignments in Sequence TransductionBailin Wang, Mirella Lapata, Ivan TitovNeurIPS 2021 · 被引用 20 次
- Explicit Syntactic Guidance for Neural Text GenerationYafu Li, Leyang Cui, Jianhao Yan, Yongjing Yin 等ACL 2023 · 被引用 4 次
- Non-autoregressive Machine Translation with Disentangled Context TransformerJungo Kasai, James Cross, Marjan Ghazvininejad, Jiatao GuICML 2020 · 被引用 113 次
- HTKG: Deep Keyphrase Generation with Neural Hierarchical Topic GuidanceYuxiang Zhang, Tao Jiang, Tianyu Yang, Xiaoli Li 等SIGIR 2022 · 被引用 14 次
- A Conditional Splitting Framework for Efficient Constituency ParsingThanh-Tung Nguyen, Xuan-Phi Nguyen, Shafiq R. Joty, Xiaoli LiACL 2021
