Hierarchically Robust Representation Learning
Qi Qian, Juhua Hu, Hao Li
Abstract
With the tremendous success of deep learning in visual tasks, the representations extracted from intermediate layers of learned models, that is, deep features, attract much attention of researchers. Previous empirical analysis shows that those features can contain appropriate semantic information. Therefore, with a model trained on a large-scale benchmark data set (e.g., ImageNet), the extracted features can work well on other tasks. In this work, we investigate this phenomenon and demonstrate that deep features can be suboptimal due to the fact that they are learned by minimizing the empirical risk. When the data distribution of the target task is different from that of the benchmark data set, the performance of deep features can degrade. Hence, we propose a hierarchically robust optimization method to learn more generic features. Considering the example-level and concept-level robustness simultaneously, we formulate the problem as a distributionally robust optimization problem with Wasserstein ambiguity set constraints, and an efficient algorithm with the conventional training pipeline is proposed. Experiments on benchmark data sets demonstrate the effectiveness of the robust deep representations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Improved Fine-Tuning by Better Leveraging Pre-Training DataZiquan Liu, Yi Xu, Yuanhong Xu, Qi Qian et al.NeurIPS 2022 · 69 citations
- Improved Visual Fine-tuning with Natural Language SupervisionJunyang Wang, Yuanhong Xu, Juhua Hu, Ming Yan et al.ICCV 2023 · 11 citations
- Weakly Supervised Representation Learning with Coarse LabelsYuanhong Xu, Qi Qian, Hao Li, Rong Jin et al.ICCV 2021 · 11 citations
Related papers
- Universal generalization guarantees for Wasserstein distributionally robust modelsTam Le, Jérôme MalickICLR 2025
- Bootstrap Your Uncertainty: Adaptive Robust Classification Driven by Optimal-TransportJiawei Huang, Minming Li, Hu DingNeurIPS 2025
- Exact Generalization Guarantees for (Regularized) Wasserstein Distributionally Robust ModelsWaïss Azizian, Franck Iutzeler, Jérôme MalickNeurIPS 2023 · 14 citations
- Examining and Combating Spurious Features under Distribution ShiftChunting Zhou, Xuezhe Ma, Paul Michel, Graham NeubigICML 2021 · 78 citations
- On Feature Learning in the Presence of Spurious CorrelationsPavel Izmailov, Polina Kirichenko, Nate Gruver, Andrew Gordon WilsonNeurIPS 2022 · 208 citations
