Fetching the paper…
Reading the bibliography…
Subpopulation shift exists widely in many real-world applications, which refers to the training and test distributions that contain the same subpopulation groups but with different subpopulation proportions.
J. Denker and Y. LeCun, “Transforming neural-net output levels to probability distributions,” in NeurIPS , 1990
1990
Earlier work this paper cites.
D. J. MacKay, “A practical bayesian framework for backpropagation networks,” Neural computation , vol. 4, no. 3, pp. 448–472, 1992
1992
Earlier work this paper cites.
D. A. Nix and A. S. Weigend, “Estimating the mean and variance of the target probability distribution,” in ICNN , 1994
1994
Earlier work this paper cites.
H. Shimodaira, “Improving predictive inference under covariate shift by weighting the log-likelihood function,” Journal of statistical planning and inference , vol. 90, no. 2, pp. 227–244, 2000
2000
Earlier work this paper cites.
N. Japkowicz, “The class imbalance problem: Significance and strategies,” in IJCAI , 2000
2000
Earlier work this paper cites.
O. Chapelle, J. Weston, L. Bottou, and V. Vapnik, “Vicinal risk minimization,” Advances in neural information processing systems , vol. 13, 2000
2000
Earlier work this paper cites.
P. L. Bartlett and S. Mendelson, “Rademacher and gaussian complexities: Risk bounds and structural results,” Journal of Machine Learning Research , vol. 3, no. Nov, pp. 463–482, 2002
2002
Earlier work this paper cites.
Q. V. Le, A. J. Smola, and S. Canu, “Heteroscedastic gaussian process regression,” in ICML , 2005
2005
Earlier work this paper cites.
J. Huang, A. Gretton, K. Borgwardt, B. Schölkopf, and A. Smola, “Correcting sample selection bias by unlabeled data,” in NeurIPS , 2006
2006
Earlier work this paper cites.
S. Bickel, M. Brückner, and T. Scheffer, “Discriminative learning for differing training and test distributions,” in ICML , 2007
2007
Earlier work this paper cites.
W. Liu and S. Chawla, “Class confidence weighted knn algorithms for imbalanced data sets,” in Pacific-Asia conference on knowledge discovery and data mining . Springer, 2011, pp. 345–356
2011
Earlier work this paper cites.
R. M. Neal, Bayesian learning for neural networks . Springer Science & Business Media, 2012, vol. 118
2012
Earlier work this paper cites.
Z. Hu and L. J. Hong, “Kullback-leibler divergence constrained distributionally robust optimization,” Available at Optimization Online , pp. 1695–1724, 2013
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
J. Wen, C.-N. Yu, and R. Greiner, “Robust learning under uncertain test distributions: Relating covariate shift to model misspecification,” in ICML , 2014
2014
Earlier work this paper cites.
Z. Liu, P. Luo, X. Wang, and X. Tang, “Deep learning face attributes in the wild,” in ICCV , 2015
2015
Earlier work this paper cites.
S. Barocas and A. D. Selbst, “Big data’s disparate impact,” Calif. L. Rev. , vol. 104, p. 671, 2016
2016
Earlier work this paper cites.
H. Namkoong and J. C. Duchi, “Stochastic gradient methods for distributionally robust optimization with f-divergences,” in NeurIPS , 2016
2016
Earlier work this paper cites.
B. Sun and K. Saenko, “Deep coral: Correlation alignment for deep domain adaptation,” in European conference on computer vision . Springer, 2016, pp. 443–450
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in CVPR , 2016
2016
Earlier work this paper cites.
Y. Gal and Z. Ghahramani, “Dropout as a bayesian approximation: Representing model uncertainty in deep learning,” in ICML , 2016
2016
Earlier work this paper cites.
T.-Y. Lin, P. Goyal, R. Girshick, K. He, and P. Dollár, “Focal loss for dense object detection,” in ICCV , 2017
2017
Earlier work this paper cites.
A. Kendall and Y. Gal, “What uncertainties do we need in bayesian deep learning for computer vision?” in NeurIPS , 2017
2017
Earlier work this paper cites.
B. Lakshminarayanan, A. Pritzel, and C. Blundell, “Simple and scalable predictive uncertainty estimation using deep ensembles,” in NeurIPS , 2017
2017
Earlier work this paper cites.
G. Huang, Y. Li, G. Pleiss, Z. Liu, J. E. Hopcroft, and K. Q. Weinberger, “Snapshot ensembles: Train 1, get m for free,” in ICLR , 2017
2017
Earlier work this paper cites.
G. Kalweit and J. Boedecker, “Uncertainty-driven imagination for continuous deep reinforcement learning,” in Conference on Robot Learning . PMLR, 2017, pp. 195–206
2017
Earlier work this paper cites.
T. Hashimoto, M. Srivastava, H. Namkoong, and P. Liang, “Fairness without demographics in repeated loss minimization,” in ICML , 2018
2018
Earlier work this paper cites.
W. Hu, G. Niu, I. Sato, and M. Sugiyama, “Does distributionally robust supervised learning give robust classifiers?” in ICML , 2018
2018
Earlier work this paper cites.
H. Zhang, M. Cisse, Y. N. Dauphin, and D. Lopez-Paz, “mixup: Beyond empirical risk minimization,” in ICLR , 2018
2018
Earlier work this paper cites.
A. Kendall, Y. Gal, and R. Cipolla, “Multi-task learning using uncertainty to weigh losses for scene geometry and semantics,” in CVPR , 2018
2018
Cited alongside, same era.
D. Mahajan, R. Girshick, V. Ramanathan, K. He, M. Paluri, Y. Li, A. Bharambe, and L. van der Maaten, “Exploring the limits of weakly supervised pretraining,” in ECCV , 2018
2018
Cited alongside, same era.
H. Ritter, A. Botev, and D. Barber, “A scalable laplace approximation for neural networks,” in 6th International Conference on Learning Representations, ICLR 2018-Conference Track Proceedings , vol. 6. International Conference on Representation Learning, 2018
2018
Cited alongside, same era.
P. Bandi, O. Geessink, Q. Manson, M. Van Dijk, M. Balkenhol, M. Hermsen, B. E. Bejnordi, B. Lee, K. Paeng, A. Zhong et al. , “From detection of individual metastases to classification of lymph node status at the patient level: the camelyon17 challenge,” IEEE transactions on medical imaging , vol. 38, no. 2, pp. 550–560, 2018
2018
J. Moon, J. Kim, Y. Shin, and S. Hwang, “Confidence-aware learning for deep neural networks,” in ICML , 2020
2020
Later among the works it cites.
R. Zhai, C. Dan, A. Suggala, J. Z. Kolter, and P. Ravikumar, “Boosted cvar classification,” in NeurIPS , 2021
2021
Later among the works it cites.
P. Michel, T. Hashimoto, and G. Neubig, “Modeling the second player in distributionally robust optimization,” in ICLR , 2021
2021
Later among the works it cites.
R. Zhai, C. Dan, Z. Kolter, and P. Ravikumar, “Doro: Distributional and outlier robust optimization,” in ICML , 2021
2021
Later among the works it cites.
D. Xu, Y. Ye, and C. Ruan, “Understanding the role of importance weighting for deep learning,” in ICLR , 2021
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
G. Christie, N. Fendley, J. Wilson, and R. Mukherjee, “Functional map of the world,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 6172–6180
2018
Cited alongside, same era.
2019
Cited alongside, same era.
Y. Cui, M. Jia, T.-Y. Lin, Y. Song, and S. Belongie, “Class-balanced loss based on effective number of samples,” in CVPR , 2019
2019
Cited alongside, same era.
K. Cao, C. Wei, A. Gaidon, N. Arechiga, and T. Ma, “Learning imbalanced datasets with label-distribution-aware margin loss,” in NeurIPS , 2019
2019
Cited alongside, same era.
J. Shu, Q. Xie, L. Yi, Q. Zhao, S. Zhou, Z. Xu, and D. Meng, “Meta-weight-net: Learning an explicit mapping for sample weighting,” in NeurIPS , 2019
2019
Cited alongside, same era.
J. Byrd and Z. Lipton, “What is the effect of importance weighting in deep learning?” in ICML , 2019
2019
Cited alongside, same era.
S. Yun, D. Han, S. J. Oh, S. Chun, J. Choe, and Y. Yoo, “Cutmix: Regularization strategy to train strong classifiers with localizable features,” in CVPR , 2019
2019
Cited alongside, same era.
V. Verma, A. Lamb, C. Beckham, A. Najafi, I. Mitliagkas, D. Lopez-Paz, and Y. Bengio, “Manifold mixup: Better representations by interpolating hidden states,” in ICML , 2019
2019
Cited alongside, same era.
E. Z. Liu, B. Haghgoo, A. S. Chen, A. Raghunathan, P. W. Koh, S. Sagawa, P. Liang, and C. Finn, “Just train twice: Improving group robustness without training group information,” in ICML , 2021
2021
Later among the works it cites.
L. Zhang, Z. Deng, K. Kawaguchi, A. Ghorbani, and J. Zou, “How does mixup help with robustness and generalization?” in ICLR , 2021
2021
Later among the works it cites.
M. Hong, J. Choi, and G. Kim, “Stylemix: Separating content and style for enhanced data augmentation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 14 862–14 870
2021
Later among the works it cites.
M. Havasi, R. Jenatton, S. Fort, J. Z. Liu, J. Snoek, B. Lakshminarayanan, A. M. Dai, and D. Tran, “Training independent subnetworks for robust prediction,” in ICLR , 2021
2021
Later among the works it cites.
H. Ma, Z. Han, C. Zhang, H. Fu, J. T. Zhou, and Q. Hu, “Trustworthy multimodal regression with mixture of normal-inverse gamma distributions,” NeurIPS , 2021
2021
Later among the works it cites.
Y. Geng, Z. Han, C. Zhang, and Q. Hu, “Uncertainty-aware multi-view representation learning,” in AAAI , 2021
2021
Later among the works it cites.
D. Deng, L. Wu, and B. E. Shi, “Iterative distillation for better uncertainty estimates in multitask emotion recognition,” in ICCV , 2021
2021
Later among the works it cites.
K. Li, A. Gupta, A. Reddy, V. H. Pong, A. Zhou, J. Yu, and S. Levine, “Mural: Meta-learning uncertainty-aware rewards for outcome-driven reinforcement learning,” in ICML , 2021
2021
Later among the works it cites.
R. Arora, P. Bartlett, P. Mianjy, and N. Srebro, “Dropout: Explicit forms and capacity control,” in ICML , 2021
2021
Later among the works it cites.
P. W. Koh, S. Sagawa, H. Marklund, S. M. Xie, M. Zhang, A. Balsubramani, W. Hu, M. Yasunaga, R. L. Phillips, I. Gao et al. , “Wilds: A benchmark of in-the-wild distribution shifts,” in ICML , 2021
2021
Later among the works it cites.
K. Ahuja, E. Caballero, D. Zhang, J.-C. Gagnon-Audet, Y. Bengio, I. Mitliagkas, and I. Rish, “Invariance principle meets information bottleneck for out-of-distribution generalization,” in NeurIPS , 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
D. Krueger, E. Caballero, J.-H. Jacobsen, A. Zhang, J. Binas, D. Zhang, R. Le Priol, and A. Courville, “Out-of-distribution generalization via risk extrapolation (rex),” in ICML , 2021
2021
Later among the works it cites.
Microsoft, “Neural Network Intelligence,” 1 2021. [Online]. Available: https://github.com/microsoft/nni
2021
Later among the works it cites.
D. Mincu and S. Roy, “Developing robust benchmarks for driving forward ai innovation in healthcare,” Nature Machine Intelligence , vol. 4, no. 11, pp. 916–921, 2022
2022
Later among the works it cites.
P. Michel, T. Hashimoto, and G. Neubig, “Distributionally robust models with parametric likelihood ratios,” in ICLR , 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
Z. Han, Z. Liang, F. Yang, L. Liu, L. Li, Y. Bian, P. Zhao, B. Wu, C. Zhang, and J. Yao, “UMIX: Improving importance weighting for subpopulation shift via uncertainty-aware mixup,” in Advances in Neural Information Processing Systems , 2022
2022
Later among the works it cites.
L. Carratino, M. Ciss, R. Jenatton, and J.-P. Vert, “On mixup regularization,” Journal of Machine Learning Research , vol. 23, no. 325, pp. 1–31, 2022
2022
Later among the works it cites.
H. Yao, Y. Wang, S. Li, L. Zhang, W. Liang, J. Zou, and C. Finn, “Improving out-of-distribution robustness via selective augmentation,” in ICML , 2022
2022
Later among the works it cites.
Z. Han, C. Zhang, H. Fu, and J. T. Zhou, “Trusted multi-view classification with dynamic evidential fusion,” IEEE TPAMI , 2022
2022
Later among the works it cites.
V. Piratla, P. Netrapalli, and S. Sarawagi, “Focus on the common good: Group distributional robustness follows,” in ICLR , 2022
2022
Later among the works it cites.
A. Kumar, T. Ma, P. Liang, and A. Raghunathan, “Calibrated ensembles can mitigate accuracy tradeoffs under distribution shift,” in The 38th Conference on Uncertainty in Artificial Intelligence , 2022
2022
Later among the works it cites.