Fetching the paper…
Reading the bibliography…
Leveraging the models' outputs, specifically the logits, is a common approach to estimating the test accuracy of a pre-trained neural network on out-of-distribution (OOD) samples without requiring access to the corresponding ground truth labels.
Benchmarking neural network robustness to common corruptions and perturbations
Hendrycks, D. and Dietterich, T. (2019) · 1903
Earlier work this paper cites.
Variabilità e mutabilità: contributo allo studio delle distribuzioni e delle relazioni statistiche. [Fasc. I.]
Gini, C. (1912) · 1912
Earlier work this paper cites.
Characterizing the decision boundary of deep neural networks
Karimi, H., Derr, T., and Tang, J. (2019) · 1912
Earlier work this paper cites.
Rank correlation methods
Kendall, M. G. (1948) · 1948
Earlier work this paper cites.
Possible generalization of Boltzmann-Gibbs statistics
Tsallis, C. (1988) · 1988
Earlier work this paper cites.
A note on a general definition of the coefficient of determination
Nagelkerke, N. J. et al. (1991) · 1991
Earlier work this paper cites.
Python reference manual
Van Rossum, G. and Drake Jr, F. L. (1995) · 1995
Earlier work this paper cites.
Statistical Learning Theory
Vapnik, V. N. (1998) · 1998
Earlier work this paper cites.
Understanding the decision boundary of deep neural networks: An empirical study
Mickisch, D., Assion, F., Greßner, F., Günther, W., and Motta, M. (2020) · 2002
Earlier work this paper cites.
Nonextensive Entropy: Interdisciplinary Applications
Gell-Mann, M. and Tsallis, C. (2004) · 2004
Earlier work this paper cites.
Co-validation: Using model disagreement on unlabeled data to validate classification algorithms
Madani, O., Pennock, D., and Flake, G. (2004) · 2004
Earlier work this paper cites.
Semi-supervised classification by low density separation
Chapelle, O. and Zien, A. (2005) · 2005
Earlier work this paper cites.
On bayesian bounds
Banerjee, A. (2006) · 2006
Earlier work this paper cites.
Pattern Recognition and Machine Learning
Bishop, C. M. (2006) · 2006
Earlier work this paper cites.
Estimating generalization under distribution shifts via domain-invariant representations
Chuang, C.-Y., Torralba, A., and Jegelka, S. (2020) · 2007
Earlier work this paper cites.
Matplotlib: A 2d graphics environment
Hunter, J. D. (2007) · 2007
Earlier work this paper cites.
The Tsallis entropy of natural information
Sneddon, R. (2007) · 2007
Earlier work this paper cites.
Breeds: Benchmarks for subpopulation shift
Santurkar, S., Tsipras, D., and Madry, A. (2020) · 2008
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database
Deng, J., Dong, W., Socher, R., Li, L.-J., Li, K., and Fei-Fei, L. (2009) · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, A. and Hinton, G. (2009) · 2009
Earlier work this paper cites.
Dataset Shift in Machine Learning
Quionero-Candela, J., Sugiyama, M., Schwaighofer, A., and Lawrence, N. D. (2009) · 2009
Earlier work this paper cites.
Unsupervised supervised learning i: Estimating classification and regression errors without labels
Donmez, P., Lebanon, G., and Balasubramanian, K. (2010) · 2010
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Dosovitskiy, A. (2020) · 2010
Earlier work this paper cites.
Pac-bayesian analysis of co-clustering and beyond
Seldin, Y. and Tishby, N. (2010) · 2010
Earlier work this paper cites.
Pseudo-Label : The Simple and Efficient Semi-Supervised Learning Method for Deep Neural Networks
Lee, D.-H. (2013) · 2013
Earlier work this paper cites.
Identifying and attacking the saddle point problem in high-dimensional non-convex optimization
Dauphin, Y. N., Pascanu, R., Gulcehre, C., Cho, K., Ganguli, S., and Bengio, Y. (2014) · 2014
Earlier work this paper cites.
On the number of linear regions of deep neural networks
Montufar, G. F., Pascanu, R., Cho, K., and Bengio, Y. (2014) · 2014
Earlier work this paper cites.
Orbit regularization
Negrinho, R. and Martins, A. (2014) · 2014
Earlier work this paper cites.
The loss surfaces of multilayer networks
Choromanska, A., Henaff, M., Mathieu, M., Arous, G. B., and LeCun, Y. (2015) · 2015
Earlier work this paper cites.
Tiny ImageNet visual recognition challenge
Le, Y. and Yang, X. (2015) · 2015
Earlier work this paper cites.
An exploration of softmax alternatives belonging to the spherical loss family
de Brébisson, A. and Vincent, P. (2016) · 2016
Earlier work this paper cites.
Topology and geometry of half-rectified network optimization
Freeman, C. D. and Bruna, J. (2016) · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J. (2016) · 2016
Earlier work this paper cites.
A baseline for detecting misclassified and out-of-distribution examples in neural networks
Hendrycks, D. and Gimpel, K. (2016) · 2016
Earlier work this paper cites.
SGDR: Stochastic gradient descent with warm restarts
Loshchilov, I. and Hutter, F. (2016) · 2016
Earlier work this paper cites.
Estimating accuracy from unlabeled data: A bayesian approach
Platanios, E. A., Dubey, A., and Mitchell, T. (2016) · 2016
Cited alongside, same era.
Exponential expressivity in deep neural networks through transient chaos
Poole, B., Lahiri, S., Raghu, M., Sohl-Dickstein, J., and Ganguli, S. (2016) · 2016
Cited alongside, same era.
Improved deep metric learning with multi-class n-pair loss objective
Sohn, K. (2016) · 2016
Cited alongside, same era.
Wide residual networks
Zagoruyko, S. and Komodakis, N. (2016) · 2016
Cited alongside, same era.
Noisy softmax: Improving the generalization ability of dcnn via postponing the early softmax saturation
Chen, B., Deng, W., and Du, J. (2017) · 2017
Cited alongside, same era.
Sharp minima can generalize for deep nets
Dinh, L., Pascanu, R., Bengio, S., and Bengio, Y. (2017) · 2017
Cited alongside, same era.
Multi-class data description for out-of-distribution detection
Lee, D., Yu, S., and Yu, H. (2020) · 2020
Later among the works it cites.
Learning to optimize domain specific normalization for domain generalization
Seo, S., Suh, Y., Kim, D., Kim, G., Han, J., and Han, B. (2020) · 2020
Later among the works it cites.
Fixmatch: simplifying semi-supervised learning with consistency and confidence
Sohn, K., Berthelot, D., Li, C.-L., Zhang, Z., Carlini, N., Cubuk, E. D., Kurakin, A., Zhang, H., and Raffel, C. (2020) · 2020
Later among the works it cites.
Detecting errors and estimating accuracy on unlabeled data with self-training ensembles
Chen, J., Liu, F., Avci, B., Wu, X., Liang, Y., and Jha, S. (2021) · 2021
Later among the works it cites.
What does rotation prediction tell us about classifier accuracy under varying testing environments?
Deng, W., Gould, S., and Zheng, L. (2021) · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Deeper, broader and artier domain generalization
Li, D., Yang, Y., Song, Y.-Z., and Hospedales, T. M. (2017) · 2017
Cited alongside, same era.
Sphereface: Deep hypersphere embedding for face recognition
Liu, W., Wen, Y., Yu, Z., Li, M., Raj, B., and Song, L. (2017) · 2017
Cited alongside, same era.
Tsallis regularized optimal transport and ecological inference
Muzellec, B., Nock, R., Patrini, G., and Nielsen, F. (2017) · 2017
Cited alongside, same era.
Estimating accuracy from unlabeled data: A probabilistic logic approach
Platanios, E., Poon, H., Mitchell, T. M., and Horvitz, E. J. (2017) · 2017
Cited alongside, same era.
L2-constrained softmax loss for discriminative face verification
Ranjan, R., Castillo, C. D., and Chellappa, R. (2017) · 2017
Cited alongside, same era.
Deep hashing network for unsupervised domain adaptation
Venkateswara, H., Eusebio, J., Chakraborty, S., and Panchanathan, S. (2017) · 2017
Cited alongside, same era.
Are labels always necessary for classifier accuracy evaluation?
Deng, W. and Zheng, L. (2021) · 2021
Later among the works it cites.
Adversarially adaptive normalization for single domain generalization
Fan, X., Wang, Q., Ke, J., Yang, F., Gong, B., and Zhou, M. (2021) · 2021
Later among the works it cites.
Predicting with confidence on unseen distributions
Guillory, D., Shankar, V., Ebrahimi, S., Darrell, T., and Schmidt, L. (2021) · 2021
Later among the works it cites.
Assessing generalization of SGD via disagreement
Jiang, Y., Nagarajan, V., Baek, C., and Kolter, J. Z. (2021) · 2021
Later among the works it cites.
Wilds: A benchmark of in-the-wild distribution shifts
Koh, P. W., Sagawa, S., Marklund, H., Xie, S. M., Zhang, M., Balsubramani, A., Hu, W., Yasunaga, M., Phillips, R. L., Gao, I., et al. (2021) · 2021
Later among the works it cites.
On interaction between augmentations and corruptions in natural corruption robustness
Mintun, E., Kirillov, A., and Xie, S. (2021) · 2021
Later among the works it cites.
Exploring covariate and concept shift for out-of-distribution detection
Tian, J., Hsu, Y.-C., Shen, Y., Jin, H., and Kira, Z. (2021) · 2021
Later among the works it cites.
Tent: Fully test-time adaptation by entropy minimization
Wang, D., Shelhamer, E., Liu, S., Olshausen, B., and Darrell, T. (2021) · 2021
Later among the works it cites.
Deep learning generalization and the convex hull of training sets
Yousefzadeh, R. (2021) · 2021
Later among the works it cites.
Amini, M.-R., Feofanov, V., Pauletto, L., Hadjadj, L., Devijver, E., and Maximov, Y. (2022) · 2022
Later among the works it cites.
Contrastive test-time adaptation
Chen, D., Wang, D., Darrell, T., and Ebrahimi, S. (2022) · 2022
Later among the works it cites.
Leveraging unlabeled data to predict out-of-distribution performance
Garg, S., Balakrishnan, S., Lipton, Z. C., Neyshabur, B., and Sedghi, H. (2022) · 2022
Later among the works it cites.
X-risk analysis for ai research
Hendrycks, D. and Mazeika, M. (2022) · 2022
Later among the works it cites.
A new adversarial domain generalization network based on class boundary feature detection for bearing fault diagnosis
Li, J., Shen, C., Kong, L., Wang, D., Xia, M., and Zhu, Z. (2022) · 2022
Later among the works it cites.
A convnet for the 2020s
Liu, Z., Mao, H., Wu, C.-Y., Feichtenhofer, C., Darrell, T., and Xie, S. (2022) · 2022
Later among the works it cites.
If your data distribution shifts, use self-learning
Rusak, E., Schneider, S., Pachitariu, G., Eck, L., Gehler, P. V., Bringmann, O., Brendel, W., and Bethge, M. (2022) · 2022
Later among the works it cites.
Predicting out-of-distribution error with the projection norm
Yu, Y., Yang, Z., Wei, A., Ma, Y., and Steinhardt, J. (2022) · 2022
Later among the works it cites.
Deng, W., Suh, Y., Gould, S., and Zheng, L. (2023) · 2023
Later among the works it cites.
Random matrix analysis to balance between supervised and unsupervised learning under the low density separation assumption
Feofanov, V., Tiomoko, M., and Virmaux, A. (2023) · 2023
Later among the works it cites.
Characterizing out-of-distribution error via optimal transport
Lu, Y., Qin, Y., Zhai, R., Shen, A., Chen, K., Wang, Z., Kolouri, S., Stepputtis, S., Campbell, J., and Sycara, K. (2023) · 2023
Later among the works it cites.
A bag-of-prototypes representation for dataset-level applications
Tu, W., Deng, W., Gedeon, T., and Zheng, L. (2023) · 2023
Later among the works it cites.
On the importance of feature separability in predicting out-of-distribution error
Xie, R., Wei, H., Cao, Y., Feng, L., and An, B. (2023) · 2023
Later among the works it cites.
Multi-class probabilistic bounds for majority vote classifiers with partially labeled data
Feofanov, V., Devijver, E., and Amini, M.-R. (2024) · 2024
Closest in time.
Characterizing out-of-distribution error via optimal transport
Lu, Y., Qin, Y., Zhai, R., Shen, A., Chen, K., Wang, Z., Kolouri, S., Stepputtis, S., Campbell, J., and Sycara, K. (2024) · 2024
Closest in time.
Leveraging ensemble diversity for robust self-training in the presence of sample selection bias
Odonnat, A., Feofanov, V., and Redko, I. (2024) · 2024
Closest in time.
Energy-based automated model evaluation
Peng, R., Zou, H., Wang, H., Zeng, Y., Huang, Z., and Zhao, J. (2024) · 2024
Closest in time.
Inflation based on the Tsallis entropy
Teimoori, Z., Rezazadeh, K., and Rostami, A. (2024) · 2024
Closest in time.
softmax is not enough (for sharp out-of-distribution)
Veličković, P., Perivolaropoulos, C., Barbero, F., and Pascanu, R. (2024) · 2024
Closest in time.
Characterising gradients for unsupervised accuracy estimation under distribution shift
Xie, R., Odonnat, A., Feofanov, V., Redko, I., Zhang, J., and An, B. (2024) · 2024
Closest in time.
Large language models as markov chains
Zekri, O., Odonnat, A., Benechehab, A., Bleistein, L., Boullé, N., and Redko, I. (2024) · 2024
Closest in time.