Deep Subspace Clustering with Data Augmentation
Mahdi Abavisani, Alireza Naghizadeh, Dimitris N. Metaxas, Vishal M. Patel
Abstract
The idea behind data augmentation techniques is based on the fact that slight changes in the percept do not change the brain cognition. In classification, neural networks use this fact by applying transformations to the inputs to learn to predict the same label. However, in deep subspace clustering (DSC), the ground-truth labels are not available, and as a result, one cannot easily use data augmentation techniques. We propose a technique to exploit the benefits of data augmentation in DSC algorithms. We learn representations that have consistent subspaces for slightly transformed inputs. In particular, we introduce a temporal ensembling component to the objective function of DSC algorithms to enable the DSC networks to maintain consistent subspaces for random transformations in the input data. In addition, we provide a simple yet effective unsupervised procedure to find efficient data augmentation policies. An augmentation policy is defined as an image processing transformation with a certain magnitude and probability of being applied to each image in each epoch. We search through the policies in a search space of the most common augmentation policies to find the best policy such that the DSC network yields the highest mean Silhouette coefficient in its clustering results on a target dataset. Our method achieves state-of-the-art performance on four standard subspace clustering datasets. The source code is available at: https://github.com/mahdiabavisani/DSCwithDA.git .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Perfectly Balanced: Improving Transfer and Robustness of Supervised Contrastive LearningMayee F. Chen, Daniel Y. Fu, Avanika Narayan, Michael Zhang et al.ICML 2022 · 58 citations
- Effective Node-Level Anomaly Detection in HPC Systems via Coarse-Grained Clustering and Fine-Grained Model SharingSibo Xia, Yongqian Sun, Xijie Pan, Yuan Yuan et al.SC 2025 · 3 citations
Builds on3
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 4,453 citations
- Random Erasing Data AugmentationZhun Zhong, Liang Zheng, Guoliang Kang, Shaozi Li et al.AAAI 2020 · 4,134 citations
- Stochastic Sparse Subspace ClusteringYing Chen, Chun-Guang Li, Chong YouCVPR 2020
Related papers
- Joint-Label Learning by Dual Augmentation for Time Series ClassificationQianli Ma, Zhenjing Zheng, Jiawei Zheng, Sen Li et al.AAAI 2021 · 17 citations
- CADDA: Class-wise Automatic Differentiable Data Augmentation for EEG SignalsCédric Rommel, Thomas Moreau, Joseph Paillard, Alexandre GramfortICLR 2022 · 52 citations
- Evolutionary Approach for AutoAugment Using the Thermodynamical Genetic AlgorithmAkira Terauchi, Naoki MoriAAAI 2021 · 9 citations
- Augmented Contrastive Clustering with Uncertainty-Aware Prototyping for Time Series Test Time AdaptationPeiliang Gong, Mohamed Ragab, Min Wu, Zhenghua Chen et al.KDD 2025 · 1 citation
- Multi-Scale Fusion Subspace Clustering Using Similarity ConstraintZhiyuan Dang, Cheng Deng, Xu Yang, Heng HuangCVPR 2020
