Invariant Representations through Adversarial Forgetting
Ayush Jaiswal, Daniel Moyer, Greg Ver Steeg, Wael AbdAlmageed, Premkumar Natarajan
摘要
We propose a novel approach to achieving invariance for deep neural networks in the form of inducing amnesia to unwanted factors of data through a new adversarial forgetting mechanism. We show that the forgetting mechanism serves as an information-bottleneck, which is manipulated by the adversarial training to learn invariance to unwanted factors. Empirical results show that the proposed framework achieves stateof-the-art performance at learning invariance in both nuisance and bias settings on a diverse collection of datasets and tasks. Related Work Recent work (Achille and Soatto 2018b; Alemi et al. 2016; Moyer et al. 2018) has modeled invariance in supervised DNNs through information bottleneck (Tishby, Pereira, and Bialek 1999), wherein representations minimize the mutual information I(x : z) while maximizing I(z : y). For nuisance variables (s ⊥ y), these methods bring about compression in the latent space, which removes information about s and indirectly minimizes I(z : s). Under optimality, the
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Controllable Guarantees for Fair Outcomes via Contrastive Information EstimationUmang Gupta, Aaron M. Ferber, Bistra Dilkina, Greg Ver SteegAAAI 2021 · 被引用 78 次
- Disentangled Information BottleneckZiqi Pan, Li Niu, Jianfu Zhang, Liqing ZhangAAAI 2021 · 被引用 55 次
- Fair Representations by CompressionXavier Gitiaux, Huzefa RangwalaAAAI 2021 · 被引用 19 次
- AGS: Affordable and Generalizable Substitute Training for Transferable Adversarial AttackRuikui Wang, Yuanfang Guo, Yunhong WangAAAI 2024 · 被引用 17 次
- Interpretable and Robust Behavior Abstraction via Environment-Disentangled Heterogeneous GraphZhibin Ni, Hai Wan, Xibin ZhaoAAAI 2026
相关 Paper
- Forget Sharpness: Perturbed Forgetting of Model Biases Within SAM DynamicsAnkit Vani, Frederick Tung, Gabriel L. Oliveira, Hossein Sharifi-NoghabiICML 2024
- Representation Unlearning: Forgetting through Information CompressionAntonio Almudévar, Alfonso OrtegaICML 2026 · 被引用 1 次
- Label-Agnostic Forgetting: A Supervision-Free Unlearning in Deep ModelsShaofei Shen, Chenhao Zhang, Yawen Zhao, Alina Bialkowski 等ICLR 2024 · 被引用 20 次
- Cauchy-Schwarz Divergence Information Bottleneck for RegressionShujian Yu, Xi Yu, Sigurd Løkse, Robert Jenssen 等ICLR 2024 · 被引用 16 次
- Towards Defending against Adversarial Examples via Attack-Invariant FeaturesDawei Zhou, Tongliang Liu, Bo Han, Nannan Wang 等ICML 2021 · 被引用 55 次
