Randomized Quantization: A Generic Augmentation for Data Agnostic Self-supervised Learning
Huimin Wu, Chenyang Lei, Xiao Sun, Peng-Shuai Wang, Qifeng Chen, Kwang-Ting Cheng, Stephen Lin, Zhirong Wu
摘要
Self-supervised representation learning follows a paradigm of withholding some part of the data and tasking the network to predict it from the remaining part. Among many techniques, data augmentation lies at the core for creating the information gap. Towards this end, masking has emerged as a generic and powerful tool where content is withheld along the sequential dimension, e.g., spatial in images, temporal in audio, and syntactic in language. In this paper, we explore the orthogonal channel dimension for generic data augmentation by exploiting precision redundancy. The data for each channel is quantized through a non-uniform quantizer, with the quantized value sampled randomly within randomly sampled quantization bins. From another perspective, quantization is analogous to channel-wise masking, as it removes the information within each bin, but preserves the information across bins. Our approach significantly surpasses existing generic data augmentation methods, while showing on par performance against modalityspecific augmentations. We comprehensively evaluate our approach on vision, audio, 3D point clouds, as well as the DABS benchmark which is comprised of various data modalities. The code is available at https: //github.com/microsoft/random_quantize .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Binning as a Pretext Task: Improving Self-Supervised Learning in Tabular DomainsKyungeun Lee, Ye Seul Sim, Hye-Seung Cho, Moonjung Eo 等ICML 2024 · 被引用 17 次
- Data-Efficient Multimodal Fusion on a Single GPUNoël Vouitsis, Zhaoyan Liu, Satya Krishna Gorti, Valentin Villecroze 等CVPR 2024 · 被引用 6 次
- Improving Integrated Gradient-based Transferable Adversarial Examples by Refining the Integration PathYuchen Ren, Zhengyu Zhao, Chenhao Lin, Bo Yang 等AAAI 2025 · 被引用 4 次
- PASA: Progressive-Adaptive Spectral Augmentation for Automated Auscultation in Data-Scarce EnvironmentsYing Wang, Guoheng Huang, Xueyuan Gong, Xinxin Wang 等AAAI 2026
它引用的顶会 Paper26
- wav2vec 2.0: A Framework for Self-Supervised Learning of Speech RepresentationsAlexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, Michael AuliNeurIPS 2020 · 被引用 9,451 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh 等ICCV 2019 · 被引用 5,843 次
- BEiT: BERT Pre-Training of Image TransformersHangbo Bao, Li Dong, Songhao Piao, Furu WeiICLR 2022 · 被引用 3,632 次
相关 Paper
- Self-supervised Representation Learning from Random Data ProjectorsYi Sui, Tongzi Wu, Jesse C. Cresswell, Ga Wu 等ICLR 2024 · 被引用 17 次
- Frequency Selective Augmentation for Video Representation LearningJinhyung Kim, Taeoh Kim, Minho Shim, Dongyoon Han 等AAAI 2023 · 被引用 5 次
- PointCMP: Contrastive Mask Prediction for Self-supervised Learning on Point Cloud VideosZhiqiang Shen, Xiaoxiao Sheng, Longguang Wang, Yulan Guo 等CVPR 2023
- Self-supervised learning with random-projection quantizer for speech recognitionChung-Cheng Chiu, James Qin, Yu Zhang, Jiahui Yu 等ICML 2022 · 被引用 245 次
- Learning Spatio-temporal Representation by Channel Aliasing Video PerceptionYiqi Lin, Jinpeng Wang, Manlin Zhang, Andy J. MaACM MM 2021 · 被引用 2 次
