Nonlinear Mixup: Out-Of-Manifold Data Augmentation for Text Classification
Hongyu Guo
摘要
Data augmentation with Mixup (Zhang et al. 2018) has shown to be an effective model regularizer for current art deep classification networks. It generates out-of-manifold samples through linearly interpolating inputs and their corresponding labels of random sample pairs. Despite its great successes, Mixup requires convex combination of the inputs as well as the modeling targets of a sample pair, thus significantly limits the space of its synthetic samples and consequently its regularization effect. To cope with this limitation, we propose “nonlinear Mixup”. Unlike Mixup where the input and label pairs share the same, linear, scalar mixing policy, our approach embraces nonlinear interpolation policy for both the input and label pairs, where the mixing policy for the labels is adaptively learned based on the mixed input. Experiments on benchmark sentence classification datasets indicate that our approach significantly improves upon Mixup. Our empirical studies also show that the out-of-manifold samples generated by our strategy encourage training samples in each class to form a tight representation cluster that is far from others.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper22
- G-Mixup: Graph Data Augmentation for Graph ClassificationXiaotian Han, Zhimeng Jiang, Ninghao Liu, Xia HuICML 2022 · 被引用 251 次
- i-Mix: A Domain-Agnostic Strategy for Contrastive Representation LearningKibok Lee, Yian Zhu, Kihyuk Sohn, Chun-Liang Li 等ICLR 2021 · 被引用 133 次
- Semi-supervised Vision Transformers at ScaleZhaowei Cai, Avinash Ravichandran, Paolo Favaro, Manchen Wang 等NeurIPS 2022 · 被引用 82 次
- Inserting Anybody in Diffusion Models via Celeb BasisGe Yuan, Xiaodong Cun, Yong Zhang, Maomao Li 等NeurIPS 2023 · 被引用 81 次
- Geodesic Multi-Modal Mixup for Robust Fine-TuningChangdae Oh, Junhyuk So, Hoyoon Byun, YongTaek Lim 等NeurIPS 2023 · 被引用 49 次
相关 Paper
- Adversarial Mixing Policy for Relaxing Locally Linear Constraints in MixupGuang Liu, Yuzhao Mao, Hailong Huang, Weiguo Gao 等EMNLP 2021 · 被引用 3 次
- Tailoring Mixup to Data for CalibrationQuentin Bouniot, Pavlo Mozharovskyi, Florence d'Alché-BucICLR 2025
- Global Mixup: Eliminating Ambiguity with ClusteringXiangjin Xie, Yangning Li, Wang Chen, Kai Ouyang 等AAAI 2023 · 被引用 5 次
- Interpolating Graph Pair to Regularize Graph ClassificationHongyu Guo, Yongyi MaoAAAI 2023 · 被引用 14 次
- Self-Evolution Learning for Mixup: Enhance Data Augmentation on Few-Shot Text Classification TasksHaoqi Zheng, Qihuang Zhong, Liang Ding, Zhiliang Tian 等EMNLP 2023 · 被引用 4 次
