Fetching the paper…
Reading the bibliography…
Gradient-based learning algorithms have an implicit simplicity bias which in effect can limit the diversity of predictors being sampled by the learning procedure.
Bagging predictors
Leo Breiman · 1996
Earlier work this paper cites.
Generalization error of ensemble estimators
Naonori Ueda and Ryohei Nakano · 1996
Earlier work this paper cites.
The MNIST database of handwritten digits
Yann Lecun and Corinna Cortes · 1998
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann Lecun, Léon Bottou, Yoshua Bengio, and Patrick Haffner · 1998
Earlier work this paper cites.
The information bottleneck method, 2000
Naftali Tishby, Fernando C. Pereira, and William Bialek · 2000
Earlier work this paper cites.
Gaussian Processes for Machine Learning (Adaptive Computation and Machine Learning)
Carl Edward Rasmussen and Christopher K. I. Williams · 2005
Earlier work this paper cites.
Robust Optimization , volume 28 of Princeton Series in Applied Mathematics
Aharon Ben-Tal, Laurent El Ghaoui, and Arkadi Nemirovski · 2009
Earlier work this paper cites.
Learning Multiple Layers of Features from Tiny Images
Alex Krizhevsky · 2009
Earlier work this paper cites.
Domain adaptation: Learning bounds and algorithms
Yishay Mansour, Mehryar Mohri, and Afshin Rostamizadeh · 2009
Earlier work this paper cites.
Domain adaptation in regression
Corinna Cortes and Mehryar Mohri · 2011
Earlier work this paper cites.
Ensemble Methods: Foundations and Algorithms
Zhi-Hua Zhou · 2012
Earlier work this paper cites.
Representation learning: A review and new perspectives
Yoshua Bengio, Aaron Courville, and Pascal Vincent · 2013
Earlier work this paper cites.
Intriguing properties of neural networks
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian J. Goodfellow, and Rob Fergus · 2014
Earlier work this paper cites.
Probabilistic backpropagation for scalable learning of bayesian neural networks
José Miguel Hernández-Lobato and Ryan P. Adams · 2015
Earlier work this paper cites.
Concrete problems in AI safety
Dario Amodei, Chris Olah, Jacob Steinhardt, Paul F. Christiano, John Schulman, and Dan Mané · 2016
Earlier work this paper cites.
Dropout as a bayesian approximation: Representing model uncertainty in deep learning
Yarin Gal and Zoubin Ghahramani · 2016
Earlier work this paper cites.
Domain-adversarial training of neural networks
Yaroslav Ganin, Evgeniya Ustinova, Hana Ajakan, Pascal Germain, Hugo Larochelle, François Laviolette, Mario Marchand, and Victor Lempitsky · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Examples are not enough, learn to criticize! criticism for interpretability
Been Kim, Oluwasanmi Koyejo, and Rajiv Khanna · 2016
Earlier work this paper cites.
Return of frustratingly easy domain adaptation
Baochen Sun, Jiashi Feng, and Kate Saenko · 2016
Earlier work this paper cites.
A closer look at memorization in deep networks
Devansh Arpit, Stanisław Jastrzundefinedbski, Nicolas Ballas, David Krueger, Emmanuel Bengio, Maxinder S. Kanwal, Tegan Maharaj, Asja Fischer, Aaron Courville, Yoshua Bengio, and Simon Lacoste-Julien · 2017
Cited alongside, same era.
Computing nonvacuous generalization bounds for deep (stochastic) neural networks with many more parameters than training data
Gintare Karolina Dziugaite and Daniel M. Roy · 2017
Cited alongside, same era.
Simple and Scalable Predictive Uncertainty Estimation Using Deep Ensembles
Balaji Lakshminarayanan, Alexander Pritzel, and Charles Blundell · 2017
Cited alongside, same era.
Conditional adversarial domain adaptation
Mingsheng Long, Zhangjie Cao, Jianmin Wang, and Michael I Jordan · 2017
Cited alongside, same era.
Asymmetric tri-training for unsupervised domain adaptation
Kuniaki Saito, Yoshitaka Ushiku, and Tatsuya Harada · 2017
Cited alongside, same era.
Out-of-distribution generalization with maximal invariant predictor
Masanori Koyama and Shoichiro Yamaguchi · 2020
Later among the works it cites.
Learning from failure: Training debiased classifier from biased classifier
Jun Hyun Nam, Hyuntak Cha, Sungsoo Ahn, Jaeho Lee, and Jinwoo Shin · 2020
Later among the works it cites.
Hidden stratification causes clinically meaningful failures in machine learning for medical imaging
Luke Oakden-Rayner, Jared Dunnmon, Gustavo Carneiro, and Christopher Ré · 2020
Later among the works it cites.
Ensembles of locally independent prediction models
Andrew Slavin Ross, Weiwei Pan, Leo A. Celi, and Finale Doshi-Velez · 2020
Later among the works it cites.
Distributionally robust neural networks
Shiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, and Percy Liang · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Deep hashing network for unsupervised domain adaptation
Hemanth Venkateswara, Jose Eusebio, Shayok Chakraborty, and Sethuraman Panchanathan · 2017
Cited alongside, same era.
Fashion-MNIST: a novel image dataset for benchmarking machine learning algorithms
Han Xiao, Kashif Rasul, and Roland Vollgraf · 2017
Cited alongside, same era.
From detection of individual metastases to classification of lymph node status at the patient level: the camelyon17 challenge
Peter Bandi, Oscar Geessink, Quirine Manson, Marcory Van Dijk, Maschenka Balkenhol, Meyke Hermsen, Babak Ehteshami Bejnordi, Byungjae Lee, Kyunghyun Paeng, Aoxiao Zhong, et al · 2018
Cited alongside, same era.
Recognition in terra incognita
Sara Beery, Grant Van Horn, and Pietro Perona · 2018
Cited alongside, same era.
Implicit bias of gradient descent on linear convolutional networks
Suriya Gunasekar, Jason D. Lee, Daniel Soudry, and Nati Srebro · 2018
Cited alongside, same era.
Maximum classifier discrepancy for unsupervised domain adaptation
Kuniaki Saito, Kohei Watanabe, Yoshitaka Ushiku, and Tatsuya Harada · 2018
Cited alongside, same era.
Martin Arjovsky, Léon Bottou, Ishaan Gulrajani, and David Lopez-Paz · 2019
Cited alongside, same era.
The pitfalls of simplicity bias in neural networks
Harshay Shah, Kaustav Tamuly, Aditi Raghunathan, Prateek Jain, and Praneeth Netrapalli · 2020
Later among the works it cites.
Diverse ensembles improve calibration
Asa Cooper Stickland and Iain Murray · 2020
Later among the works it cites.
Uncertainty estimation using a single deep deterministic neural network
Joost van Amersfoort, Lewis Smith, Yee Whye Teh, and Yarin Gal · 2020
Later among the works it cites.
WILDS: A benchmark of in-the-wild distribution shifts
Pang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie, Marvin Zhang, Akshay Balsubramani, Weihua Hu, Michihiro Yasunaga, Richard Lanas Phillips, Irena Gao, Tony Lee, Etienne David, Ian Stavness, Wei Guo, Berton A. Earnshaw, Imran S. Haque, Sara Beery, Jure Leskovec, Anshul Kundaje, Emma Pierson, Sergey Levine, Chelsea Finn, and Percy Liang · 2021
Later among the works it cites.
Out-of-distribution generalization via risk extrapolation (rex)
David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang, Jonathan Binas, Dinghuai Zhang, Rémi Le Priol, and Aaron C. Courville · 2021
Later among the works it cites.
Gradient starvation: A learning proclivity in neural networks
Mohammad Pezeshki, Sékou-Oumar Kaba, Yoshua Bengio, Aaron C. Courville, Doina Precup, and Guillaume Lajoie · 2021
Later among the works it cites.
DICE: diversity in deep ensembles via conditional redundancy adversarial estimation
Alexandre Ramé and Matthieu Cord · 2021
Later among the works it cites.
DIBS: diversity inducing information bottleneck in model ensembles
Samarth Sinha, Homanga Bharadhwaj, Anirudh Goyal, Hugo Larochelle, Animesh Garg, and Florian Shkurti · 2021
Later among the works it cites.
A survey on gender bias in natural language processing
Karolina Stanczak and Isabelle Augenstein · 2021
Later among the works it cites.
Damien Teney, Ehsan Abbasnejad, Simon Lucey, and Anton van den Hengel · 2021
Later among the works it cites.
Can subnetwork structure be the key to out-of-distribution generalization?
Dinghuai Zhang, Kartik Ahuja, Yilun Xu, Yisen Wang, and Aaron C. Courville · 2021
Later among the works it cites.
Novel disease detection using ensembles with regularized disagreement
Alexandru Ţifrea, Eric Stavarache, and Fanny Yang · 2021
Later among the works it cites.
Diversify and disambiguate: Learning from underspecified data
Yoonho Lee, Huaxiu Yao, and Chelsea Finn · 2022
Closest in time.
Extending the WILDS benchmark for unsupervised adaptation
Shiori Sagawa, Pang Wei Koh, Tony Lee, Irena Gao, Sang Michael Xie, Kendrick Shen, Ananya Kumar, Weihua Hu, Michihiro Yasunaga, Henrik Marklund, Sara Beery, Etienne David, Ian Stavness, Wei Guo, Jure Leskovec, Kate Saenko, Tatsunori Hashimoto, Sergey Levine, Chelsea Finn, and Percy Liang · 2022
Closest in time.