Lune

ICLR2021Top-tier venue

Systematic generalisation with group invariant predictions

Faruk Ahmed, Yoshua Bengio, Harm van Seijen, Aaron C. Courville

2021Year
23Citations
48Top-tier citations

Abstract

We consider situations where the presence of dominant simpler correlations with the target variable in a training set can cause an SGD-trained neural network to be less reliant on more persistently-correlating complex features. When the non-persistent, simpler correlations correspond to non-semantic background factors, a neural network trained on this data can exhibit dramatic failure upon encountering systematic distributional shift, where the correlating background features are recombined with different objects. We perform an empirical study showing that group invariance methods across inferred partitionings of the training set can lead to significant improvements at such test-time situations. We suggest a new invariance penalty, showing with experiments on three synthetic datasets that it can perform better than alternatives. We find that even without assuming access to any systematic-shift validation sets, one can still find improvements over an ERM-trained reference model.

Ask about this paper

Ask your agent about it.

Lune has read the top-tier papers around this one, so every answer names the papers it rests on.

Questions to start from

Your agent calls

Lunesearch_papers

Ask in Lune

Free to start. No credit card required.

lune papers get b388f9d9-54a5-4cb7-8eb2-7dc93c358d17

Cited by top-tier papers48

Ask how each one uses it

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines