Fetching the paper…
Reading the bibliography…
We introduce Invariant Risk Minimization (IRM), a learning paradigm to estimate invariant correlations across multiple training distributions.
Correlation and causation
Sewall Wright · 1921
Earlier work this paper cites.
The probability approach in econometrics
Trygve Haavelmo · 1944
Earlier work this paper cites.
Estimating causal effects of treatments in randomized and nonrandomized studies
Donald B. Rubin · 1974
Earlier work this paper cites.
Causal necessity: a pragmatic investigation of the necessity of laws
Brian Skyrms · 1980
Earlier work this paper cites.
Incompleteness, non locality and realism. a prolegomenon to the philosophy of quantum mechanics
Michael Redhead · 1987
Earlier work this paper cites.
Autonomy
John Aldrich · 1989
Earlier work this paper cites.
Principles of risk minimization for learning theory
Vladimir Vapnik · 1992
Earlier work this paper cites.
Comparison of classifier methods: a case study in handwritten digit recognition
Léon Bottou, Corinna Cortes, John S. Denker, Harris Drucker, Isabelle Guyon, Lawrence D. Jackel, Yann Le Cun, Urs A. Muller, Eduard Säckinger, Patrice Simard, and Vladimir Vapnik · 1994
Earlier work this paper cites.
NIST Special Database 19: Handprinted forms and characters database
Patrick J. Grother · 1995
Earlier work this paper cites.
Statistical Learning Theory
Vladimir N. Vapnik · 1998
Earlier work this paper cites.
Dimensions of scientific law
Sandra D. Mitchell · 2000
Earlier work this paper cites.
Two theorems on invariance and causality
Nancy Cartwright · 2003
Earlier work this paper cites.
Introduction to Smooth Manifolds
James M. Lee · 2003
Earlier work this paper cites.
Robust supervised learning
James Andrew Bagnell · 2005
Earlier work this paper cites.
Making things happen: A theory of causal explanation
James Woodward · 2005
Earlier work this paper cites.
Analysis of representations for domain adaptation
Shai Ben-David, John Blitzer, Koby Crammer, and Fernando Pereira · 2007
Earlier work this paper cites.
Robust optimization
Aharon Ben-Tal, Laurent El Ghaoui, and Arkadi Nemirovski · 2009
Earlier work this paper cites.
Causality: Models, Reasoning, and Inference
Judea Pearl · 2009
Earlier work this paper cites.
Unbiased look at dataset bias
Antonio Torralba and Alexei Efros · 2011
Earlier work this paper cites.
On causal and anticausal learning
Bernhard Schölkopf, Dominik Janzing, Jonas Peters, Eleni Sgouritsa, Kun Zhang, and Joris Mooij · 2012
Cited alongside, same era.
Invariant scattering convolution networks
Joan Bruna and Stephane Mallat · 2013
Cited alongside, same era.
Counterfactuals
David Lewis · 2013
Cited alongside, same era.
A simple method to determine if a music information retrieval system is a “horse”
Bob L. Sturm · 2014
Cited alongside, same era.
Intermittent process analysis with scattering moments
Joan Bruna, Stephane Mallat, Emmanuel Bacry, and Jean-François Muzy · 2015
Cited alongside, same era.
Maximin effects in inhomogeneous large-scale data
Nicolai Meinshausen and Peter Bühlmann · 2015
Cited alongside, same era.
The marginal value of adaptive gradient methods in machine learning
Ashia C. Wilson, Rebecca Roelofs, Mitchell Stern, Nati Srebro, and Benjamin Recht · 2017
Later among the works it cites.
Recognition in terra incognita
Sara Beery, Grant Van Horn, and Pietro Perona · 2018
Later among the works it cites.
Invariant causal prediction for nonlinear models
Christina Heinze-Deml, Jonas Peters, and Nicolai Meinshausen · 2018
Later among the works it cites.
Generalization in anti-causal learning
Niki Kilbertus, Giambattista Parascandolo, and Bernhard Schölkopf · 2018
Later among the works it cites.
Stable prediction across unknown environments
Kun Kuang, Peng Cui, Susan Athey, Ruoxuan Xiong, and Bo Li · 2018
Later among the works it cites.
Deep domain generalization via conditional invariant adversarial networks
Ya Li, Xinmei Tian, Mingming Gong, Yajing Liu, Tongliang Liu, Kun Zhang, and Dacheng Tao · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Statistics of robust optimization: A generalized empirical likelihood approach
John Duchi, Peter Glynn, and Hongseok Namkoong · 2016
Cited alongside, same era.
Domain-adversarial training of neural networks
Yaroslav Ganin, Evgeniya Ustinova, Hana Ajakan, Pascal Germain, Hugo Larochelle, François Laviolette, Mario March, and Victor Lempitsky · 2016
Cited alongside, same era.
Revisiting visual question answering baselines
Allan Jabri, Armand Joulin, and Laurens Van Der Maaten · 2016
Cited alongside, same era.
From dependence to causation
David Lopez-Paz · 2016
Cited alongside, same era.
Causal inference using invariant prediction: identification and confidence intervals
Jonas Peters, Peter Bühlmann, and Nicolai Meinshausen · 2016
Cited alongside, same era.
Understanding deep learning requires rethinking generalization
Chiyuan Zhang, Samy Bengio, Moritz Hardt, Benjamin Recht, and Oriol Vinyals · 2016
Cited alongside, same era.
Domain adaptation by using causal inference to predict invariant conditional distributions
Sara Magliacane, Thijs van Ommen, Tom Claassen, Stephan Bongers, Philip Versteeg, and Joris M Mooij · 2018
Later among the works it cites.
Deep learning: A critical appraisal
Gary Marcus · 2018
Later among the works it cites.
Causality from a distributional robustness point of view
Nicolai Meinshausen · 2018
Later among the works it cites.
Invariant models for causal transfer learning
Mateo Rojas-Carulla, Bernhard Schölkopf, Richard Turner, and Jonas Peters · 2018
Later among the works it cites.
Certifying some distributional robustness with principled adversarial training
Aman Sinha, Hongseok Namkoong, and John Duchi · 2018
Later among the works it cites.
Benign Overfitting in Linear Regression
Peter L. Bartlett, Philip M. Long, Gábor Lugosi, and Alexander Tsigler · 2019
Closest in time.
A meta-transfer objective for learning to disentangle causal mechanisms
Yoshua Bengio, Tristan Deleu, Nasim Rahaman, Rosemary Ke, Sébastien Lachapelle, Olexa Bilaniuk, Anirudh Goyal, and Christopher Pal · 2019
Closest in time.
Approximating CNNs with bag-of-local-features models works surprisingly well on imagenet
Wieland Brendel and Matthias Bethge · 2019
Closest in time.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Closest in time.
Imagenet-trained cnns are biased towards texture; increasing shape bias improves accuracy and robustness
Robert Geirhos, Patricia Rubisch, Claudio Michaelis, Matthias Bethge, Felix A. Wichmann, and Wieland Brendel · 2019
Closest in time.
Support and invertibility in domain-invariant representations
Fredrik D. Johansson, David A. Sontag, and Rajesh Ranganath · 2019
Closest in time.
Do we still need models or just more data and compute?, 2019
Max Welling · 2019
Closest in time.