Fetching the paper…
Reading the bibliography…
Miscalibration in deep learning refers to there is a discrepancy between the predicted confidence and performance.
G. W. Brier, “Verification of forecasts expressed in terms of probability,” Monthly weather review , vol. 78, no. 1, pp. 1–3, 1950
1950
Earlier work this paper cites.
J. Denker and Y. LeCun, “Transforming neural-net output levels to probability distributions,” in Advances in Neural Information Processing Systems , 1990
1990
Earlier work this paper cites.
D. J. MacKay, “A practical bayesian framework for backpropagation networks,” Neural Computation , vol. 4, no. 3, pp. 448–472, 1992
1992
Earlier work this paper cites.
P. J. Huber, “Robust estimation of a location parameter,” in Breakthroughs in statistics: Methodology and distribution . Springer, 1992, pp. 492–518
1992
Earlier work this paper cites.
A. Krizhevsky, G. Hinton et al. , “Learning multiple layers of features from tiny images,” 2009
2009
Earlier work this paper cites.
R. M. Neal, Bayesian learning for neural networks . Springer Science & Business Media, 2012, vol. 118
2012
Earlier work this paper cites.
L. Bossard, M. Guillaumin, and L. Van Gool, “Food-101–mining discriminative components with random forests,” in Computer Vision–ECCV 2014 . Springer, pp. 446–461
2014
Earlier work this paper cites.
Y. LeCun, Y. Bengio, and G. Hinton, “Deep learning,” Nature , vol. 521, no. 7553, pp. 436–444, 2015
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Earlier work this paper cites.
S. Zagoruyko and N. Komodakis, “Wide residual networks,” in BMVC , 2016
2016
Earlier work this paper cites.
C. Guo, G. Pleiss, Y. Sun, and K. Q. Weinberger, “On calibration of modern neural networks,” in International Conference on Machine Learning . PMLR, 2017, pp. 1321–1330
2017
Earlier work this paper cites.
A. Esteva, B. Kuprel, R. A. Novoa, J. Ko, S. M. Swetter, H. M. Blau, and S. Thrun, “Dermatologist-level classification of skin cancer with deep neural networks,” Nature , vol. 542, no. 7639, pp. 115–118, 2017
2017
Earlier work this paper cites.
G. Pereyra, G. Tucker, J. Chorowski, Ł. Kaiser, and G. Hinton, “Regularizing neural networks by penalizing confident output distributions,” in International Conference on Learning Representations , 2017
2017
Earlier work this paper cites.
B. Lakshminarayanan, A. Pritzel, and C. Blundell, “Simple and scalable predictive uncertainty estimation using deep ensembles,” in Advances in Neural Information Processing Systems , 2017
2017
Earlier work this paper cites.
A. Kendall and Y. Gal, “What uncertainties do we need in bayesian deep learning for computer vision?” in Advances in Neural Information Processing Systems , 2017
2017
Earlier work this paper cites.
Y. Geifman and R. El-Yaniv, “Selective classification for deep neural networks,” Advances in Neural Information Processing Systems , vol. 30, 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
M. Toneva, A. Sordoni, R. T. des Combes, A. Trischler, Y. Bengio, and G. J. Gordon, “An empirical study of example forgetting during deep neural network learning,” in International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
A. Kumar, S. Sarawagi, and U. Jain, “Trainable calibration measures for neural networks from kernel mean embeddings,” in International Conference on Machine Learning, ICML 2018 , J. G. Dy and A. Krause, Eds
2018
Cited alongside, same era.
A. Malinin and M. Gales, “Predictive uncertainty estimation via prior networks,” Advances in Neural Information Processing Systems , vol. 31, 2018
2018
Cited alongside, same era.
M. Sensoy, L. Kaplan, and M. Kandemir, “Evidential deep learning to quantify classification uncertainty,” Advances in Neural Information Processing Systems , vol. 31, 2018
2018
Cited alongside, same era.
D.-B. Wang, L. Feng, and M.-L. Zhang, “Rethinking calibration of deep neural networks: Do not be afraid of overconfidence,” Advances in Neural Information Processing Systems , vol. 34, pp. 11 809–11 820, 2021
2021
Later among the works it cites.
X. Du, Z. Wang, M. Cai, and Y. Li, “Vos: Learning what you don’t know by virtual outlier synthesis,” in International Conference on Learning Representations , 2021
2021
Later among the works it cites.
J. Chen, Y. Li, X. Wu, Y. Liang, and S. Jha, “Atom: Robustifying out-of-distribution detection using outlier mining,” in Machine Learning and Knowledge Discovery in Databases. , 2021, pp. 430–445
2021
Later among the works it cites.
Y. Bai, S. Mei, H. Wang, and C. Xiong, “Don’t just blame over-parametrization for over-confidence: Theoretical analysis of calibration in binary classification,” in International Conference on Machine Learning . PMLR, 2021, pp. 566–576
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
K. Lee, H. Lee, K. Lee, and J. Shin, “Training confidence-calibrated classifiers for detecting out-of-distribution samples,” in International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
D. Hendrycks, M. Mazeika, and T. Dietterich, “Deep anomaly detection with outlier exposure,” in International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
P. Bandi, O. Geessink, Q. Manson, M. Van Dijk, M. Balkenhol, M. Hermsen, B. E. Bejnordi, B. Lee, K. Paeng, A. Zhong et al. , “From detection of individual metastases to classification of lymph node status at the patient level: the camelyon17 challenge,” IEEE transactions on medical imaging , vol. 38, no. 2, pp. 550–560, 2018
2018
Cited alongside, same era.
Y. Geifman, G. Uziel, and R. El-Yaniv, “Bias-reduced uncertainty estimation for deep neural classifiers,” in International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
R. Müller, S. Kornblith, and G. E. Hinton, “When does label smoothing help?” Advances in Neural Information Processing Systems , vol. 32, 2019
2019
Cited alongside, same era.
A. C. Lorena, L. P. Garcia, J. Lehmann, M. C. Souto, and T. K. Ho, “How complex is your classification problem? a survey on measuring classification complexity,” ACM Computing Surveys (CSUR) , 2019
2019
Cited alongside, same era.
S. Yun, D. Han, S. J. Oh, S. Chun, J. Choe, and Y. Yoo, “Cutmix: Regularization strategy to train strong classifiers with localizable features,” in Proceedings of the IEEE/CVF international conference on computer vision , 2019, pp. 6023–6032
2019
Cited alongside, same era.
J. Liu, J. Paisley, M.-A. Kioumourtzoglou, and B. Coull, “Accurate uncertainty estimation and decomposition in ensemble learning,” Advances in Neural Information Processing Systems , vol. 32, 2019
2019
Cited alongside, same era.
P. W. Koh, S. Sagawa, H. Marklund, S. M. Xie, M. Zhang, A. Balsubramani, W. Hu, M. Yasunaga, R. L. Phillips, I. Gao et al. , “Wilds: A benchmark of in-the-wild distribution shifts,” in International Conference on Machine Learning . PMLR, 2021, pp. 5637–5664
2021
Later among the works it cites.
A. Ghosh, T. Schaaf, and M. Gormley, “Adafocal: Calibration-aware adaptive focal loss,” Advances in Neural Information Processing Systems , vol. 35, pp. 1583–1595, 2022
2022
Later among the works it cites.
H. Wei, R. Xie, H. Cheng, L. Feng, B. An, and Y. Li, “Mitigating neural network overconfidence with logit normalization,” in International Conference on Machine Learning . PMLR, 2022, pp. 23 631–23 644
2022
Later among the works it cites.
N. Seedat, J. Crabbé, I. Bica, and M. van der Schaar, “Data-iq: Characterizing subgroups with heterogeneous outcomes in tabular data,” Advances in Neural Information Processing Systems , 2022
2022
Later among the works it cites.
V. Vasudevan, B. Caine, R. Gontijo Lopes, S. Fridovich-Keil, and R. Roelofs, “When does dough become a bagel? analyzing the remaining mistakes on imagenet,” Advances in Neural Information Processing Systems , vol. 35, pp. 6720–6734, 2022
2022
Later among the works it cites.
H. Yao, Y. Wang, S. Li, L. Zhang, W. Liang, J. Zou, and C. Finn, “Improving out-of-distribution robustness via selective augmentation,” in International Conference on Machine Learning . PMLR, 2022, pp. 25 407–25 437
2022
Later among the works it cites.
Z. Han, Z. Liang, F. Yang, L. Liu, L. Li, Y. Bian, P. Zhao, B. Wu, C. Zhang, and J. Yao, “Umix: Improving importance weighting for subpopulation shift via uncertainty-aware mixup,” Advances in Neural Information Processing Systems , vol. 35, pp. 37 704–37 718, 2022
2022
Later among the works it cites.
Y. Yu, S. Bates, Y. Ma, and M. Jordan, “Robust calibration with multi-domain temperature scaling,” Advances in Neural Information Processing Systems , vol. 35, pp. 27 510–27 523, 2022
2022
Later among the works it cites.
J. Katz-Samuels, J. B. Nakhleh, R. Nowak, and Y. Li, “Training ood detectors in their natural habitats,” in International Conference on Machine Learning . PMLR, 2022, pp. 10 848–10 865
2022
Later among the works it cites.
Y. Ming, Y. Fan, and Y. Li, “Poem: Out-of-distribution detection with posterior sampling,” in International Conference on Machine Learning . PMLR, 2022, pp. 15 650–15 665
2022
Later among the works it cites.
L. Tao, M. Dong, and C. Xu, “Dual focal loss for calibration,” in International Conference on Machine Learning , 2023
2023
Later among the works it cites.
Y. Yang, H. Zhang, D. Katabi, and M. Ghassemi, “Change is hard: A closer look at subpopulation shift,” in International Conference on Machine Learning , 2023
2023
Later among the works it cites.
H. Wei, H. Zhuang, R. Xie, L. Feng, G. Niu, B. An, and Y. Li, “Mitigating memorization of noisy labels by clipping the model prediction,” in International Conference on Machine Learning , 2023
2023
Later among the works it cites.