MODALS: Modality-agnostic Automated Data Augmentation in the Latent Space
Tsz-Him Cheung, Dit-Yan Yeung
摘要
Data augmentation is an efficient way to expand a training dataset by creating additional artificial data. While data augmentation is found to be effective in improving the generalization capabilities of models for various machine learning tasks, the underlying augmentation methods are usually manually designed and carefully evaluated for each data modality separately, like image processing functions for image data and word-replacing rules for text data. In this work, we propose an automated data augmentation approach called MODALS (Modality-agnostic Automated Data Augmentation in the Latent Space) to augment data for any modality in a generic way. MODALS exploits automated data augmentation to fine-tune four universal data transformation operations in the latent space to adapt the transform to data of different modalities. Through comprehensive experiments, we demonstrate the effectiveness of MODALS on multiple datasets for text, tabular, time-series and image modalities.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper8
- Finding Order in Chaos: A Novel Data Augmentation Method for Time Series in Contrastive LearningBerken Utku Demirel, Christian HolzNeurIPS 2023 · 被引用 48 次
- Data Augmentation of Contrastive Learning is Estimating Positive-incentive NoiseHongyuan Zhang, Yanchen Xu, Sida Huang, Xuelong LiICML 2026 · 被引用 41 次
- A Feature-space Multimodal Data Augmentation Technique for Text-video RetrievalAlex Falcon, Giuseppe Serra, Oswald LanzACM MM 2022 · 被引用 25 次
- Towards Better Understanding and Better Generalization of Low-shot Classification in Histology Images with Contrastive LearningJiawei Yang, Hanbo Chen, Jiangpeng Yan, Xiaoyu Chen 等ICLR 2022 · 被引用 25 次
- Wiener Graph Deconvolutional Network Improves Graph Self-Supervised LearningJiashun Cheng, Man Li, Jia Li, Fugee TsungAAAI 2023 · 被引用 24 次
相关 Paper
- Pseudo-Non-Linear Data Augmentation: A Constrained Energy Minimization ViewpointPingbang Hu, Mahito SugiyamaICLR 2026
- AdaAug: Learning Class- and Instance-adaptive Data Augmentation PoliciesTsz-Him Cheung, Dit-Yan YeungICLR 2022 · 被引用 31 次
- AutoDA-Timeseries: Automated Data Augmentation for Time SeriesZijun Dou, Zhenhe Yao, Zhe Xie, Xidao Wen 等ICLR 2026
- Text AutoAugment: Learning Compositional Augmentation Policy for Text ClassificationShuhuai Ren, Jinchao Zhang, Lei Li, Xu Sun 等EMNLP 2021 · 被引用 22 次
- Learning Multimodal Data Augmentation in Feature SpaceZichang Liu, Zhiqiang Tang, Xingjian Shi, Aston Zhang 等ICLR 2023 · 被引用 9 次
