Fetching the paper…
Reading the bibliography…
Reliable confidence estimation for the predictions is important in many safety-critical applications.
Brier, G.W.: Verification of forecasts expressed in terms of probability. Monthly Weather Review pp. 1–3 (1950)
1950
Earlier work this paper cites.
Brier, G.W., et al.: Verification of forecasts expressed in terms of probability. Monthly weather review 78
1950
Earlier work this paper cites.
Deng, J., Dong, W., Socher, R., Li, L.J., Li, K., Fei-Fei, L.: Imagenet: A large-scale hierarchical image database. In: CVPR. pp. 248–255 (2009)
2009
Earlier work this paper cites.
Krizhevsky, A., Hinton, G., et al.: Learning multiple layers of features from tiny images. Tech. rep., Citeseer (2009)
2009
Earlier work this paper cites.
Leidner, D., Borst, C., Dietrich, A., Beetz, M., Albu-Schäffer, A.: Classifying compliant manipulation tasks for automated planning in robotics. In: 2015 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS). pp. 1769–1776 (2015)
2015
Earlier work this paper cites.
Naeini, M.P., Cooper, G.F., Hauskrecht, M.: Obtaining well calibrated probabilities using bayesian binning. In: AAAI. pp. 2901–2907 (2015)
2015
Earlier work this paper cites.
Simonyan, K., Zisserman, A.: Very deep convolutional networks for large-scale image recognition. In: ICLR (2015)
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
Gal, Y., Ghahramani, Z.: Dropout as a bayesian approximation: Representing model uncertainty in deep learning. In: ICML. pp. 1050–1059 (2016)
2016
Earlier work this paper cites.
He, K., Zhang, X., Ren, S., Sun, J.: Deep residual learning for image recognition. In: CVPR. pp. 770–778 (2016)
2016
Earlier work this paper cites.
He, K., Zhang, X., Ren, S., Sun, J.: Identity mappings in deep residual networks. In: ECCV. pp. 630–645 (2016)
2016
Earlier work this paper cites.
Miotto, R., Li, L., Kidd, B.A., Dudley, J.T.: Deep patient: an unsupervised representation to predict the future of patients from the electronic health records. Scientific reports p. 26094 (2016)
2016
Earlier work this paper cites.
Zagoruyko, S., Komodakis, N.: Wide residual networks. In: BMVC (2016)
2016
Earlier work this paper cites.
Esteva, A., Kuprel, B., Novoa, R.A., Ko, J., Swetter, S.M., Blau, H.M., Thrun, S.: Dermatologist-level classification of skin cancer with deep neural networks. nature 542
2017
Earlier work this paper cites.
Geifman, Y., El-Yaniv, R.: Selective classification for deep neural networks. In: NeurIPS. pp. 4878–4887 (2017)
2017
Earlier work this paper cites.
Guo, C., Pleiss, G., Sun, Y., Weinberger, K.Q.: On calibration of modern neural networks. In: ICML. pp. 1321–1330 (2017)
2017
Earlier work this paper cites.
Hendrycks, D., Gimpel, K.: A baseline for detecting misclassified and out-of-distribution examples in neural networks. In: ICLR (2017)
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
Huang, G., Liu, Z., van der Maaten, L., Weinberger, K.Q.: Densely connected convolutional networks. In: CVPR. pp. 2261–2269 (2017)
2017
Earlier work this paper cites.
Kendall, A., Gal, Y.: What uncertainties do we need in bayesian deep learning for computer vision? In: NeurIPS. pp. 5574–5584 (2017)
2017
Earlier work this paper cites.
Kull, M., de Menezes e Silva Filho, T., Flach, P.A.: Beta calibration: a well-founded and easily implemented improvement on logistic calibration for binary classifiers. In: AISTATS. pp. 623–631 (2017)
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
Achille, A., Soatto, S.: Emergence of invariance and disentanglement in deep representations. J. Mach. Learn. Res. 19
2018
Earlier work this paper cites.
Izmailov, P., Wilson, A., Podoprikhin, D., Vetrov, D., Garipov, T.: Averaging weights leads to wider optima and better generalization. In: UAI. pp. 876–885 (2018)
2018
Earlier work this paper cites.
Jiang, H., Kim, B., Gupta, M.R.: To trust or not to trust a classifier. In: NeurIPS (2018)
2018
Earlier work this paper cites.
Lee, K., Lee, K., Lee, H., Shin, J.: A simple unified framework for detecting out-of-distribution samples and adversarial attacks. In: NeurIPS. pp. 7167–7177 (2018)
2018
Earlier work this paper cites.
Liang, S., Li, Y., Srikant, R.: Enhancing the reliability of out-of-distribution image detection in neural networks. In: ICLR (2018)
2018
Earlier work this paper cites.
Zhang, H., Cisse, M., Dauphin, Y.N., Lopez-Paz, D.: Mixup: Beyond empirical risk minimization. In: ICLR (2018)
2018
Cited alongside, same era.
Chaudhari, P., Choromanska, A., Soatto, S., LeCun, Y., Baldassi, C., Borgs, C., Chayes, J., Sagun, L., Zecchina, R.: Entropy-sgd: Biasing gradient descent into wide valleys. Journal of Statistical Mechanics: Theory and Experiment 2019
2019
Cited alongside, same era.
Corbière, C., Thome, N., Bar-Hen, A., Cord, M., Pérez, P.: Addressing failure prediction by learning model confidence. In: NeurIPS. pp. 2898–2909 (2019)
2019
Cited alongside, same era.
Geifman, Y., Uziel, G., El-Yaniv, R.: Bias-reduced uncertainty estimation for deep neural classifiers. In: ICLR (2019)
2019
Cited alongside, same era.
Hendrycks, D., Dietterich, T.G.: Benchmarking neural network robustness to common corruptions and perturbations. In: ICLR (2019)
Mukhoti, J., Kulharia, V., Sanyal, A., Golodetz, S., Torr, P.H.S., Dokania, P.K.: Calibrating deep neural networks using focal loss. In: NeurIPS (2020)
2020
Later among the works it cites.
Patel, K., Beluch, W.H., Yang, B., Pfeiffer, M., Zhang, D.: Multi-class uncertainty calibration via mutual information maximization-based binning. In: ICLR (2020)
2020
Later among the works it cites.
Rahimi, A., Shaban, A., Cheng, C., Hartley, R., Boots, B.: Intra order-preserving functions for calibration of multi-class neural networks. In: NeurIPS (2020)
2020
Later among the works it cites.
Rice, L., Wong, E., Kolter, Z.: Overfitting in adversarially robust deep learning. In: ICML. pp. 8093–8104 (2020)
2020
Later among the works it cites.
Shehzad, M.N., Bashir, Q., Farooq, U., Ahmed, G., Raza, M., Kumar, P.M., Khalid, M.: Threshold temperature scaling: Heuristic to address temperature and power issues in mpsocs. Microprocess. Microsystems p. 103124 (2020)
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
Hendrycks, D., Mazeika, M., Dietterich, T.G.: Deep anomaly detection with outlier exposure. In: ICLR (2019)
2019
Cited alongside, same era.
Kull, M., Perelló-Nieto, M., Kängsepp, M., de Menezes e Silva Filho, T., Song, H., Flach, P.A.: Beyond temperature scaling: Obtaining well-calibrated multi-class probabilities with dirichlet calibration. In: NeurIPS. pp. 12295–12305 (2019)
2019
Cited alongside, same era.
Kumar, A., Liang, P.S., Ma, T.: Verified uncertainty calibration. NeurIPS (2019)
2019
Cited alongside, same era.
Maddox, W.J., Izmailov, P., Garipov, T., Vetrov, D.P., Wilson, A.G.: A simple baseline for bayesian uncertainty in deep learning. NeurIPS 32
2019
Cited alongside, same era.
Mozafari, A.S., Gomes, H.S., Leão, W., Gagné, C.: Unsupervised temperature scaling: An unsupervised post-processing calibration method of deep networks. arXiv: Computer Vision and Pattern Recognition (2019)
2019
Cited alongside, same era.
Müller, R., Kornblith, S., Hinton, G.: When does label smoothing help? In: NeurIPS. pp. 4696–4705 (2019)
2019
Cited alongside, same era.
Nixon, J., Dusenberry, M.W., Zhang, L., Jerfel, G., Tran, D.: Measuring calibration in deep learning. In: CVPR Workshops. vol. 2 (2019)
2019
Cited alongside, same era.
2020
Later among the works it cites.
Shen, Z., Liu, Z., Xu, D., Chen, Z., Cheng, K.T., Savvides, M.: Is label smoothing truly incompatible with knowledge distillation: An empirical study. In: ICLR (2020)
2020
Later among the works it cites.
Wen, Y., Jerfel, G., Muller, R., Dusenberry, M.W., Snoek, J., Lakshminarayanan, B., Tran, D.: Combining ensembles and data augmentation can harm your calibration. In: ICLR (2020)
2020
Later among the works it cites.
Wu, D., Xia, S., Wang, Y.: Adversarial weight perturbation helps robust generalization. In: NeurIPS (2020)
2020
Later among the works it cites.
Xing, C., Arik, S.Ö., Zhang, Z., Pfister, T.: Distance-based learning from errors for confidence calibration. In: ICLR (2020)
2020
Later among the works it cites.
Yun, S., Park, J., Lee, K., Shin, J.: Regularizing class-wise predictions via self-knowledge distillation. In: CVPR. pp. 13873–13882 (2020)
2020
Later among the works it cites.
Cha, J., Chun, S., Lee, K., Cho, H.C., Park, S., Lee, Y., Park, S.: Swad: Domain generalization by seeking flat minima. In: NeurIPS (2021)
2021
Later among the works it cites.
Chen, T., Zhang, Z., Liu, S., Chang, S., Wang, Z.: Robust overfitting may be mitigated by properly learned smoothening. In: ICLR (2021)
2021
Later among the works it cites.
Corbière, C., Thome, N., Saporta, A., Vu, T.H., Cord, M., Perez, P.: Confidence estimation via auxiliary models. IEEE Transactions on Pattern Analysis and Machine Intelligence (2021)
2021
Later among the works it cites.
Luo, Y., Wong, Y., Kankanhalli, M.S., Zhao, Q.: Learning to predict trustworthiness with steep slope loss. NeurIPS (2021)
2021
Later among the works it cites.
Minderer, M., Djolonga, J., Romijnders, R., Hubis, F., Zhai, X., Houlsby, N., Tran, D., Lucic, M.: Revisiting the calibration of modern neural networks. NeurIPS (2021)
2021
Later among the works it cites.
Pittorino, F., Lucibello, C., Feinauer, C., Perugini, G., Baldassi, C., Demyanenko, E., Zecchina, R.: Entropic gradient descent algorithms and wide flat minima. Journal of Statistical Mechanics: Theory and Experiment (12), 124015 (2021)
2021
Later among the works it cites.
Tolstikhin, I.O., Houlsby, N., Kolesnikov, A., Beyer, L., Zhai, X., Unterthiner, T., Yung, J., Steiner, A., Keysers, D., Uszkoreit, J., et al.: Mlp-mixer: An all-mlp architecture for vision. NeurIPS (2021)
2021
Later among the works it cites.
Wang, D., Feng, L., Zhang, M.: Rethinking calibration of deep neural networks: Do not be afraid of overconfidence. In: NeurIPS (2021)
2021
Later among the works it cites.
Zhang, W., Vaidya, I.: Mixup training leads to reduced overfitting and improved calibration for the transformer architecture. CoRR (2021)
2021
Later among the works it cites.
Zhong, Z., Cui, J., Liu, S., Jia, J.: Improving calibration for long-tailed recognition. In: CVPR. pp. 16489–16498 (2021)
2021
Later among the works it cites.
Hebbalaguppe, R., Prakash, J., Madan, N., Arora, C.: A stitch in time saves nine: A train-time regularizing loss for improved neural network calibration. In: CVPR. pp. 16081–16090 (June 2022)
2022
Later among the works it cites.
Liu, B., Ben Ayed, I., Galdran, A., Dolz, J.: The devil is in the margin: Margin-based label smoothing for network calibration. In: CVPR. pp. 80–88 (June 2022)
2022
Later among the works it cites.
Murphy, K.P.: Probabilistic Machine Learning: An introduction. MIT Press (2022), probml.ai
2022
Later among the works it cites.
Trockman, A., Kolter, J.Z.: Patches are all you need? arXiv preprint arXiv:2201.09792 (2022)
2022
Later among the works it cites.
Zhang, L., Deng, Z., Kawaguchi, K., Zou, J.: When and how mixup improves calibration. In: ICML. pp. 26135–26160 (2022)
2022
Later among the works it cites.
Zhu, F., Zhang, X.Y., Wang, R.Q., Liu, C.L.: Learning by seeing more classes. IEEE Transactions on Pattern Analysis and Machine Intelligence (2022)
2022
Later among the works it cites.