Out-of-Distribution Generalization via Risk Extrapolation (REx)

Krueger, David; Caballero, Ethan; Jacobsen, Joern-Henrik; Zhang, Amy; Binas, Jonathan; Zhang, Dinghuai; Priol, Remi Le; Courville, Aaron

Computer Science > Machine Learning

arXiv:2003.00688v5 (cs)

[Submitted on 2 Mar 2020 (v1), last revised 25 Feb 2021 (this version, v5)]

Title:Out-of-Distribution Generalization via Risk Extrapolation (REx)

Authors:David Krueger, Ethan Caballero, Joern-Henrik Jacobsen, Amy Zhang, Jonathan Binas, Dinghuai Zhang, Remi Le Priol, Aaron Courville

View PDF

Abstract:Distributional shift is one of the major obstacles when transferring machine learning prediction systems from the lab to the real world. To tackle this problem, we assume that variation across training domains is representative of the variation we might encounter at test time, but also that shifts at test time may be more extreme in magnitude. In particular, we show that reducing differences in risk across training domains can reduce a model's sensitivity to a wide range of extreme distributional shifts, including the challenging setting where the input contains both causal and anti-causal elements. We motivate this approach, Risk Extrapolation (REx), as a form of robust optimization over a perturbation set of extrapolated domains (MM-REx), and propose a penalty on the variance of training risks (V-REx) as a simpler variant. We prove that variants of REx can recover the causal mechanisms of the targets, while also providing some robustness to changes in the input distribution ("covariate shift"). By appropriately trading-off robustness to causally induced distributional shifts and covariate shift, REx is able to outperform alternative methods such as Invariant Risk Minimization in situations where these types of shift co-occur.

Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Neural and Evolutionary Computing (cs.NE); Machine Learning (stat.ML)
Cite as:	arXiv:2003.00688 [cs.LG]
	(or arXiv:2003.00688v5 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2003.00688

Submission history

From: David Krueger [view email]
[v1] Mon, 2 Mar 2020 06:29:50 UTC (1,131 KB)
[v2] Tue, 3 Mar 2020 04:15:23 UTC (1,131 KB)
[v3] Fri, 13 Mar 2020 22:57:37 UTC (1,131 KB)
[v4] Thu, 10 Dec 2020 21:46:28 UTC (1,131 KB)
[v5] Thu, 25 Feb 2021 17:53:07 UTC (1,488 KB)

Computer Science > Machine Learning

Title:Out-of-Distribution Generalization via Risk Extrapolation (REx)

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Out-of-Distribution Generalization via Risk Extrapolation (REx)

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators