Fetching the paper…
Reading the bibliography…
Conventional supervised learning methods, especially deep ones, are found to be sensitive to out-of-distribution (OOD) examples, largely because the learned representation mixes the semantic factor with the variation factor due to their domain-specific correlation, while only the semantic factor causes the output.
Disentangling factors of variation using few labels
F. Locatello, M. Tschannen, S. Bauer, G. Rätsch, B. Schölkopf, and O. Bachem · 1905
Earlier work this paper cites.
The identification of structural characteristics
T. C. Koopmans and O. Reiersol · 1950
Earlier work this paper cites.
Recognition-by-components: a theory of human image understanding
I. Biederman · 1987
Earlier work this paper cites.
Equivalence and synthesis of causal models
T. Verma and J. Pearl · 1991
Earlier work this paper cites.
Bayesian learning for neural networks
R. M. Neal · 1995
Earlier work this paper cites.
An introduction to variational methods for graphical models
M. I. Jordan, Z. Ghahramani, T. S. Jaakkola, and L. K. Saul · 1999
Earlier work this paper cites.
Causation, prediction, and search
P. Spirtes, C. N. Glymour, R. Scheines, and D. Heckerman · 2000
Earlier work this paper cites.
ICE-BeeM: Identifiable conditional energy-based deep models
I. Khemakhem, R. P. Monti, D. P. Kingma, and A. Hyvärinen · 2002
Earlier work this paper cites.
Ancestral graph Markov models
T. Richardson, P. Spirtes, et al · 2002
Earlier work this paper cites.
A new metric for probability distributions
D. M. Endres and J. E. Schindelin · 2003
Earlier work this paper cites.
Estimation of non-normalized statistical models by score matching
A. Hyvärinen · 2005
Earlier work this paper cites.
Designing data augmentation for simulating interventions
M. Ilse, J. M. Tomczak, and P. Forré · 2005
Earlier work this paper cites.
Pattern recognition and machine learning
C. M. Bishop · 2006
Earlier work this paper cites.
Elements of information theory
T. M. Cover and J. A. Thomas · 2006
Earlier work this paper cites.
Unsupervised learning of image manifolds by semidefinite programming
K. Q. Weinberger and L. K. Saul · 2006
Earlier work this paper cites.
Estimation of causal effects using linear non-gaussian causal models with hidden variables
P. O. Hoyer, S. Shimizu, A. J. Kerminen, and M. Palviainen · 2008
Earlier work this paper cites.
Supervised topic models
J. D. Mcauliffe and D. M. Blei · 2008
Earlier work this paper cites.
Graphical models, exponential families, and variational inference
M. J. Wainwright, M. I. Jordan, et al · 2008
Earlier work this paper cites.
Identifying confounders using additive noise models
D. Janzing, J. Peters, J. M. Mooij, and B. Schölkopf · 2009
Earlier work this paper cites.
Causality
J. Pearl · 2009
Earlier work this paper cites.
On the identifiability of the post-nonlinear causal model
K. Zhang and A. Hyvärinen · 2009
Earlier work this paper cites.
Domain adaptation via transfer component analysis
S. J. Pan, I. W. Tsang, J. T. Kwok, and Q. Yang · 2010
Earlier work this paper cites.
Detecting low-complexity unobserved causes
D. Janzing, E. Sgouritsa, O. Stegle, J. Peters, and B. Schölkopf · 2011
Earlier work this paper cites.
Probability and Measure
P. Billingsley · 2012
Earlier work this paper cites.
Machine learning: a probabilistic perspective
K. P. Murphy · 2012
Earlier work this paper cites.
On causal and anticausal learning
B. Schölkopf, D. Janzing, J. Peters, E. Sgouritsa, K. Zhang, and J. M. Mooij · 2012
Earlier work this paper cites.
Lecture 6.5-RMSprop: Divide the gradient by a running average of its recent magnitude
T. Tieleman and G. Hinton · 2012
Earlier work this paper cites.
Unsupervised domain adaptation by domain invariant projection
M. Baktashmotlagh, M. T. Harandi, B. C. Lovell, and M. Salzmann · 2013
Earlier work this paper cites.
Unbiased metric learning: On the utilization of multiple datasets and web images for softening bias
C. Fang, Y. Xu, and D. N. Rockmore · 2013
Earlier work this paper cites.
Domain generalization via invariant feature representation
K. Muandet, D. Balduzzi, and B. Schölkopf · 2013
Earlier work this paper cites.
Identifying finite mixtures of nonparametric product distributions and causal inference of confounders
E. Sgouritsa, D. Janzing, J. Peters, and B. Schölkopf · 2013
Earlier work this paper cites.
Domain adaptation under target and conditional shift
K. Zhang, B. Schölkopf, K. Muandet, and Z. Wang · 2013
Earlier work this paper cites.
https://www.imageclef.org/2014, 2014
The imageclef-da challenge 2014 · 2014
Earlier work this paper cites.
CAM: Causal additive models, high-dimensional order search and penalized regression
P. Bühlmann, J. Peters, J. Ernest, et al · 2014
Earlier work this paper cites.
Generative adversarial nets
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Earlier work this paper cites.
Auto-encoding variational Bayes
D. P. Kingma and M. Welling · 2014
Earlier work this paper cites.
Semi-supervised learning with deep generative models
D. P. Kingma, S. Mohamed, D. J. Rezende, and M. Welling · 2014
Earlier work this paper cites.
Causal discovery with continuous additive noise models
J. Peters, J. M. Mooij, D. Janzing, and B. Schölkopf · 2014
Earlier work this paper cites.
Introduction to nested Markov models
I. Shpitser, R. J. Evans, T. S. Richardson, and J. M. Robins · 2014
Earlier work this paper cites.
Dropout: a simple way to prevent neural networks from overfitting
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov · 2014
Earlier work this paper cites.
Intriguing properties of neural networks
C. Szegedy, W. Zaremba, I. Sutskever, J. Bruna, D. Erhan, I. Goodfellow, and R. Fergus · 2014
Cited alongside, same era.
Explaining and harnessing adversarial examples
I. J. Goodfellow, J. Shlens, and C. Szegedy · 2015
Cited alongside, same era.
How (not) to train your generative model: Scheduled sampling, likelihood, adversary?
F. Huszár · 2015
Cited alongside, same era.
Learning transferable features with deep adaptation networks
M. Long, Y. Cao, J. Wang, and M. Jordan · 2015
Cited alongside, same era.
InfoGAN: Interpretable representation learning by information maximizing generative adversarial nets
X. Chen, Y. Duan, R. Houthooft, J. Schulman, I. Sutskever, and P. Abbeel · 2016
Cited alongside, same era.
Testing the manifold hypothesis
Approximating CNNs with bag-of-local-features models works surprisingly well on ImageNet
W. Brendel and M. Bethge · 2019
Later among the works it cites.
Learning disentangled semantic representation for domain adaptation
R. Cai, Z. Li, P. Wei, J. Qiao, K. Zhang, and Z. Hao · 2019
Later among the works it cites.
Diagnosing and enhancing VAE models
B. Dai and D. Wipf · 2019
Later among the works it cites.
ImageNet-trained CNNs are biased towards texture; increasing shape bias improves accuracy and robustness
R. Geirhos, P. Rubisch, C. Michaelis, M. Bethge, F. A. Wichmann, and W. Brendel · 2019
Later among the works it cites.
Towards non-i.i.d. image classification: A dataset and baselines
Y. He, Z. Shen, and P. Cui · 2019
Later among the works it cites.
Conditional variance penalties and domain shift robustness
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. Fefferman, S. Mitter, and H. Narayanan · 2016
Cited alongside, same era.
Dropout as a Bayesian approximation: Representing model uncertainty in deep learning
Y. Gal and Z. Ghahramani · 2016
Cited alongside, same era.
Domain-adversarial training of neural networks
Y. Ganin, E. Ustinova, H. Ajakan, P. Germain, H. Larochelle, F. Laviolette, M. Marchand, and V. Lempitsky · 2016
Cited alongside, same era.
Domain adaptation with conditional transferable components
M. Gong, K. Zhang, T. Liu, D. Tao, C. Glymour, and B. Schölkopf · 2016
Cited alongside, same era.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
Adversarial examples in the physical world
A. Kurakin, I. Goodfellow, and S. Bengio · 2016
Cited alongside, same era.
Causal inference by using invariant prediction: identification and confidence intervals
J. Peters, P. Bühlmann, and N. Meinshausen · 2016
Cited alongside, same era.
C. Heinze-Deml and N. Meinshausen · 2019
Later among the works it cites.
Adversarial examples are not bugs, they are features
A. Ilyas, S. Santurkar, D. Tsipras, L. Engstrom, B. Tran, and A. Madry · 2019
Later among the works it cites.
Support and invertibility in domain-invariant representations
F. D. Johansson, D. Sontag, and R. Ranganath · 2019
Later among the works it cites.
Learning neural causal models from unknown interventions
N. R. Ke, O. Bilaniuk, A. Goyal, S. Bauer, H. Larochelle, C. Pal, and Y. Bengio · 2019
Later among the works it cites.
Leveraging directed causal discovery to detect latent common causes
C. M. Lee, C. Hart, J. G. Richens, and S. Johri · 2019
Later among the works it cites.
PyTorch: An imperative style, high-performance deep learning library
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, et al · 2019
Later among the works it cites.
Causality for machine learning
B. Schölkopf · 2019
Later among the works it cites.
The blessings of multiple causes
Y. Wang and D. M. Blei · 2019
Later among the works it cites.
Y. Yacoby, W. Pan, and F. Doshi-Velez · 2019
Later among the works it cites.
Towards accurate model selection in deep unsupervised domain adaptation
K. You, X. Wang, M. Long, and M. Jordan · 2019
Later among the works it cites.
Bridging theory and algorithm for domain adaptation
Y. Zhang, T. Liu, M. Long, and M. Jordan · 2019
Later among the works it cites.
On learning invariant representations for domain adaptation
H. Zhao, R. T. Des Combes, K. Zhang, and G. Gordon · 2019
Later among the works it cites.
A causal view of compositional zero-shot recognition
Y. Atzmon, F. Kreuk, U. Shalit, and G. Chechik · 2020
Closest in time.
A meta-transfer objective for learning to disentangle causal mechanisms
Y. Bengio, T. Deleu, N. Rahaman, N. R. Ke, S. Lachapelle, O. Bilaniuk, A. Goyal, and C. J. Pal · 2020
Closest in time.
Counterfactuals uncover the modular structure of deep generative models
M. Besserve, A. Mehrjou, R. Sun, and B. Schölkopf · 2020
Closest in time.
Causality matters in medical imaging
D. C. Castro, I. Walker, and B. Glocker · 2020
Closest in time.
Weakly supervised disentanglement by pairwise similarities
J. Chen and K. Batmanghelich · 2020
Closest in time.
Estimating generalization under distribution shifts via domain-invariant representations
C.-Y. Chuang, A. Torralba, and S. Jegelka · 2020
Closest in time.
Towards discriminability and diversity: Batch nuclear-norm maximization under label insufficient situations
S. Cui, S. Wang, J. Zhuo, L. Li, Q. Huang, and Q. Tian · 2020
Closest in time.
Underspecification presents challenges for credibility in modern machine learning
A. D’Amour, K. Heller, D. Moldovan, B. Adlam, B. Alipanahi, A. Beutel, C. Chen, J. Deaton, J. Eisenstein, M. D. Hoffman, et al · 2020
Closest in time.
In search of lost domain generalization
I. Gulrajani and D. Lopez-Paz · 2020
Closest in time.
Transfer-learning-library
J. Jiang, B. Fu, and M. Long · 2020
Closest in time.
Variational autoencoders and nonlinear ICA: A unifying framework
I. Khemakhem, D. P. Kingma, R. P. Monti, and A. Hyvärinen · 2020
Closest in time.
Out-of-distribution generalization via risk extrapolation (REx)
D. Krueger, E. Caballero, J.-H. Jacobsen, A. Zhang, J. Binas, R. L. Priol, and A. Courville · 2020
Closest in time.
Weakly-supervised disentanglement without compromises
F. Locatello, B. Poole, G. Rätsch, B. Schölkopf, O. Bachem, and M. Tschannen · 2020
Closest in time.
Learning to learn single domain generalization
F. Qiao, L. Zhao, and X. Peng · 2020
Closest in time.
Weakly supervised disentanglement with guarantees
R. Shu, Y. Chen, A. Kumar, S. Ermon, and B. Poole · 2020
Closest in time.
Few-shot domain adaptation by causal mechanism transfer
T. Teshima, I. Sato, and M. Sugiyama · 2020
Closest in time.
CausalVAE: Structured causal disentanglement in variational autoencoder
M. Yang, F. Liu, Z. Chen, X. Shen, J. Hao, and J. Wang · 2020
Closest in time.
A causal view on robustness of neural networks
C. Zhang, K. Zhang, and Y. Li · 2020
Closest in time.
On maximum likelihood training of score-based generative models
C. Durkan and Y. Song · 2021
Closest in time.
Representation learning via invariant causal mechanisms
J. Mitrovic, B. McWilliams, J. C. Walker, L. H. Buesing, and C. Blundell · 2021
Closest in time.
Toward causal representation learning
B. Schölkopf, F. Locatello, S. Bauer, N. R. Ke, N. Kalchbrenner, A. Goyal, and Y. Bengio · 2021
Closest in time.
Generalizing to unseen domains: A survey on domain generalization
J. Wang, C. Lan, C. Liu, Y. Ouyang, and T. Qin · 2021
Closest in time.
Towards a theoretical framework of out-of-distribution generalization
H. Ye, C. Xie, T. Cai, R. Li, Z. Li, and L. Wang · 2021
Closest in time.