Compositional Risk Minimization
Divyat Mahajan, Mohammad Pezeshki, Charles Arnal, Ioannis Mitliagkas, Kartik Ahuja, Pascal Vincent
摘要
Compositional generalization is a crucial step towards developing data-efficient intelligent machines that generalize in human-like ways. In this work, we tackle a challenging form of distribution shift, termed compositional shift, where some attribute combinations are completely absent at training but present in the test distribution. This shift tests the model's ability to generalize compositionally to novel attribute combinations in discriminative tasks. We model the data with flexible additive energy distributions, where each energy term represents an attribute, and derive a simple alternative to empirical risk minimization termed compositional risk minimization (CRM). We first train an additive energy classifier to predict the multiple attributes and then adjust this classifier to tackle compositional shifts. We provide an extensive theoretical analysis of CRM, where we show that our proposal extrapolates to special affine hulls of seen attribute combinations. Empirical evaluations on benchmark datasets confirms the improved robustness of CRM compared to other methods from the literature designed to tackle various forms of subpopulation shifts.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Product of Experts for Visual GenerationYunzhi Zhang, Carson Murtuza-Lanier, Zizhang Li, Yilun Du 等ICLR 2026 · 被引用 7 次
- Compositional Visual Planning via Inference-Time Diffusion ScalingYixin Zhang, Yunhao Luo, Utkarsh A. Mishra, Woo Chul Shin 等ICLR 2026 · 被引用 2 次
- Long-Text-to-Image Generation via Compositional Prompt DecompositionJen-Yuan Huang, Tong Lin, Yilun DuICLR 2026 · 被引用 1 次
- Energy-based Compositional Diffusion PlanningTao Sun, Utkarsh Mishra, Jiaxin Lu, Danfei Xu 等ICML 2026
- Compositional Scene Understanding through Inverse Generative ModelingYanbo Wang, Justin Dauwels, Yilun DuICML 2025
它引用的顶会 Paper30
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan 等ICLR 2020 · 被引用 1,496 次
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang 等ICML 2021 · 被引用 1,163 次
- Long-tail learning via logit adjustmentAditya Krishna Menon, Sadeep Jayasumana, Ankit Singh Rawat, Himanshu Jain 等ICLR 2021 · 被引用 937 次
- Balanced Meta-Softmax for Long-Tailed Visual RecognitionJiawei Ren, Cunjun Yu, Shunan Sheng, Xiao Ma 等NeurIPS 2020 · 被引用 861 次
- The Risks of Invariant Risk MinimizationElan Rosenfeld, Pradeep Kumar Ravikumar, Andrej RisteskiICLR 2021 · 被引用 356 次
相关 Paper
- Adaptive Risk Minimization: Learning to Adapt to Domain ShiftMarvin Zhang, Henrik Marklund, Nikita Dhawan, Abhishek Gupta 等NeurIPS 2021 · 被引用 284 次
- Bayesian Invariant Risk MinimizationYong Lin, Hanze Dong, Hao Wang, Tong ZhangCVPR 2022 · 被引用 48 次
- Heterogeneous Risk MinimizationJiashuo Liu, Zheyuan Hu, Peng Cui, Bo Li 等ICML 2021 · 被引用 170 次
- MAGANet: Achieving Combinatorial Generalization by Modeling a Group ActionGeonho Hwang, Jaewoong Choi, Hyunsoo Cho, Myungjoo KangICML 2023 · 被引用 4 次
- Diverse Prototypical Ensembles Improve Robustness to Subpopulation ShiftMinh Nguyen Nhat To, Paul F. R. Wilson, Viet Nguyen, Mohamed Harmanani 等ICML 2025
