Gradient-Guided Annealing for Domain Generalization
Aristotelis Ballas, Christos Diou
摘要
Figure 1. (left) Decision boundaries of a 4 th -degree polynomial logistic regression model with 2D input. In this example, feature x1 is class-specific and x2 is domain-specific, while color represents classes and shapes represent domains. The samples with solid red and green colors are included in the training data, whereas the fainted samples are part of the hidden held-out test set. As a result, domain shift is represented by a change in x2. Although the classifier should only infer based on x1, traditional gradient descent leads to overfitting (top-left). The proposed method, GGA (bottom-left), introduces an annealing process that depends on gradient agreement, leading to models that generalize well to new, unobserved target domains. (right) Schematics of the parameter updates of ERM (top-right) and GGA (bottom-right). Parameters updated via ERM are driven by gradient conflict, whereas GGA searches for a point where gradients align before continuing descending towards a minima.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Reasoning-Driven Multimodal LLM for Domain GeneralizationZhipeng Xu, Zilong Wang, Xinyang Jiang, Dongsheng Li 等ICLR 2026 · 被引用 11 次
- Noise-Aware Generalization: Robustness to In-Domain Noise and Out-of-Domain GeneralizationSiqi Wang, Aoming Liu, Bryan A. PlummerICLR 2026 · 被引用 3 次
- Rethinking Out-of-Distribution Detection and Generalization with Collective Behavior DynamicsZhenbin Wang, Lei Zhang, Wei Huang, Zhao Zhang 等NeurIPS 2025 · 被引用 2 次
- Anomaly-Related Residual Fields for Cross-domain Anomaly DetectionKewei Gao, Jiayi Xie, Zhengda Shen, Weijun Qin 等CVPR 2026
- HamiPose: Hamiltonian Optimization for Unsupervised Domain Adaptive Pose EstimationJiawen Li, Fei Jiang, Dandan Zhu, Aimin ZhouCVPR 2026
它引用的顶会 Paper21
- Gradient Surgery for Multi-Task LearningTianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine 等NeurIPS 2020 · 被引用 2,261 次
- Moment Matching for Multi-Source Domain AdaptationXingchao Peng, Qinxun Bai, Xide Xia, Zijun Huang 等ICCV 2019 · 被引用 2,239 次
- Sharpness-aware Minimization for Efficiently Improving GeneralizationPierre Foret, Ariel Kleiner, Hossein Mobahi, Behnam NeyshaburICLR 2021 · 被引用 1,861 次
- Distributionally Robust Neural NetworksShiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, Percy LiangICLR 2020 · 被引用 1,578 次
- In Search of Lost Domain GeneralizationIshaan Gulrajani, David Lopez-PazICLR 2021 · 被引用 1,416 次
相关 Paper
- Federated Unsupervised Domain Generalization Using Global and Local Alignment of GradientsFarhad Pourpanah, Mahdiyar Molahasani, Milad Soltany, Michael A. Greenspan 等AAAI 2025 · 被引用 10 次
- Domain Generalization via Gradient SurgeryLucas Mansilla, Rodrigo Echeveste, Diego H. Milone, Enzo FerranteICCV 2021 · 被引用 98 次
- Gradient Distribution Alignment Certificates Better Adversarial Domain AdaptationZhiqiang Gao, Shufei Zhang, Kaizhu Huang, Qiufeng Wang 等ICCV 2021 · 被引用 56 次
- One-Step Generalization Ratio Guided Optimization for Domain GeneralizationSumin Cho, Dongwon Kim, Kwangsu KimICML 2025
- Towards Understanding GD with Hard and Conjugate Pseudo-labels for Test-Time AdaptationJun-Kun Wang, Andre WibisonoICLR 2023 · 被引用 2 次
