Fetching the paper…
Reading the bibliography…
As a main field of artificial intelligence, natural language processing (NLP) has achieved remarkable success via deep neural networks.
S. Zhang, C. Gong, and E. Choi, “Knowing more about questions can help: Improving calibration in question answering,” in ACL-IJCNLP , 2021, pp. 1958–1970
1970
Earlier work this paper cites.
L. VENUTI, “The translator’s invisibility,” Criticism , pp. 179–212, 1986
1986
Earlier work this paper cites.
R. M. Neal, “Bayesian training of backpropagation networks by the hybrid monte carlo method,” Citeseer, Tech. Rep., 1992
1992
Earlier work this paper cites.
G. E. Hinton and D. Van Camp, “Keeping the neural networks simple by minimizing the description length of the weights,” in 6COLT93 , 1993, pp. 5–13
1993
Earlier work this paper cites.
S. Lei, X. Zhang, J. He, F. Chen, and C.-T. Lu, “Uncertainty-aware cross-lingual transfer with pseudo partial labels,” in NAACL-Findings , 2022, pp. 1987–1997
1997
Earlier work this paper cites.
N. Varshney, S. Mishra, and C. Baral, “Investigating selective prediction approaches across several tasks in IID, OOD, and adversarial settings,” in ACL-Findings , 2022, pp. 1995–2002
2002
Earlier work this paper cites.
A. Stent, M. Marge, and M. Singhai, “Evaluating evaluation methods for generation in the presence of variation.” in CICLing , 2005, pp. 341–351
2005
Earlier work this paper cites.
C. Corley and R. Mihalcea, “Measuring the semantic similarity of texts,” in Proceedings of the ACL Workshop on Empirical Modeling of Semantic Equivalence and Entailment . Ann Arbor, Michigan: Association for Computational Linguistics, Jun. 2005, pp. 13–18. [Online]. Available: https://aclanthology.org/W05-1203
2005
Earlier work this paper cites.
C. K. Williams and C. E. Rasmussen, Gaussian processes for machine learning , 2006
2006
Earlier work this paper cites.
J. Quinonero-Candela, C. E. Rasmussen, F. Sinz, O. Bousquet, and B. Scholkopf, “Evaluating predictive uncertainty challenge,” Lecture Notes in Computer Science , pp. 1–27, 2006
2006
Earlier work this paper cites.
R. Levy, “A noisy-channel model of human sentence comprehension under uncertain input,” in EMNLP , 2008, pp. 234–243
2008
Earlier work this paper cites.
J. Zhu, H. Wang, T. Yao, and B. K. Tsou, “Active learning with sampling by uncertainty and density for word sense disambiguation and text classification,” in COLING , 2008, pp. 1137–1144
2008
Earlier work this paper cites.
R. El-Yaniv and Y. Wiener, “On the foundations of noise-free selective classification,” J. Mach. Learn. Res. , p. 1605–1641, 2010
2010
Earlier work this paper cites.
2011
Earlier work this paper cites.
V. Dragos, “An ontological analysis of uncertainty in soft data,” in FUSION , 2013, pp. 1566–1573
2013
Earlier work this paper cites.
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean, “Distributed representations of words and phrases and their compositionality,” NeurIPS , 2013
2013
Earlier work this paper cites.
K. Shah, T. Conn, and L. Specia, “An investigation on the effectiveness of features for translation quality estimation,” in MTSummit , 2013
2013
Earlier work this paper cites.
G.-P. Bonneau, H.-C. Hege, C. R. Johnson, M. M. Oliveira, K. Potter, P. Rheingans, and T. Schultz, “Overview and state-of-the-art of uncertainty visualization,” Scientific Visualization: Uncertainty, Multifield, Biomedical, and Scalable Visualization , pp. 3–27, 2014
2014
Earlier work this paper cites.
O. Levy and Y. Goldberg, “Linguistic regularities in sparse and explicit word representations,” in CoNLL , 2014, pp. 171–180
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
D. Beck, T. Cohn, and L. Specia, “Joint emotion analysis via multi-task gaussian processes,” in EMNLP , 2014, pp. 1798–1803
2014
Earlier work this paper cites.
I. J. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” stat , p. 10, 2014
2014
Earlier work this paper cites.
V. Kuleshov and P. S. Liang, “Calibrated structured prediction,” NeurIPS , 2015
2015
Earlier work this paper cites.
C. Blundell, J. Cornebise, K. Kavukcuoglu, and D. Wierstra, “Weight uncertainty in neural network,” in ICML , 2015, pp. 1613–1622
2015
Earlier work this paper cites.
D. P. Kingma, T. Salimans, and M. Welling, “Variational dropout and the local reparameterization trick,” NeurIPS , 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
S. Jean, O. Firat, K. Cho, R. Memisevic, and Y. Bengio, “Montreal neural machine translation systems for WMT’15,” in Proceedings of the Tenth Workshop on Statistical Machine Translation , 2015, pp. 134–140
2015
Earlier work this paper cites.
I. Partalas, C. Lopez, N. Derbas, and R. Kalitvianski, “Learning to search for recognizing named entities in Twitter,” in WNUT , 2016, pp. 171–177
2016
Earlier work this paper cites.
C. Louizos and M. Welling, “Structured and efficient variational deep learning with matrix gaussian posteriors,” in ICML , 2016, pp. 1708–1716
2016
Earlier work this paper cites.
Y. Gal and Z. Ghahramani, “Dropout as a bayesian approximation: Representing model uncertainty in deep learning,” in ICML , 2016, pp. 1050–1059
2016
Earlier work this paper cites.
D. Beck, L. Specia, and T. Cohn, “Exploring prediction uncertainty in machine translation quality estimation,” in SIGNLL , 2016, pp. 208–218
2016
Earlier work this paper cites.
C. Guo, G. Pleiss, Y. Sun, and K. Q. Weinberger, “On calibration of modern neural networks,” in ICML , 2017, pp. 1321–1330
2017
Earlier work this paper cites.
B. Lakshminarayanan, A. Pritzel, and C. Blundell, “Simple and scalable predictive uncertainty estimation using deep ensembles,” NeurIPS , 2017
2017
Earlier work this paper cites.
M. John, “Uncertainty in visual text analysis in the context of the digital humanities,” 2017
2017
Earlier work this paper cites.
F. Zhai, S. Potdar, B. Xiang, and B. Zhou, “Neural models for sequence chunking,” in AAAI , 2017
2017
Earlier work this paper cites.
S. Ruder and B. Plank, “Learning to select data for transfer learning with Bayesian optimization,” in EMNLP , Sep. 2017, pp. 372–382
2017
Earlier work this paper cites.
T.-Y. Lin, P. Goyal, R. Girshick, K. He, and P. Dollár, “Focal loss for dense object detection,” in Proceedings of the IEEE international conference on computer vision , 2017, pp. 2980–2988
2017
Earlier work this paper cites.
M. Assel, D. D. Sjoberg, and A. J. Vickers, “The brier score does not evaluate the clinical utility of diagnostic tests or prediction models,” Diagnostic and prognostic research , pp. 1–7, 2017
2017
Earlier work this paper cites.
D. Hendrycks and K. Gimpel, “A baseline for detecting misclassified and out-of-distribution examples in neural networks,” in ICLR , 2017
2017
Earlier work this paper cites.
S. Ryu, S. Kim, J. Choi, H. Yu, and G. G. Lee, “Neural sentence embedding using only in-domain sentences for out-of-domain sentence detection in dialog systems,” Pattern Recognition Letters , pp. 26–32, 2017
2017
Earlier work this paper cites.
P. Koehn and R. Knowles, “Six challenges for neural machine translation,” in Proceedings of the First Workshop on Neural Machine Translation , 2017, pp. 28–39
2017
Earlier work this paper cites.
A. Siddhant and Z. C. Lipton, “Deep Bayesian active learning for natural language processing: Results of a large-scale empirical study,” in EMNLP , 2018, pp. 2904–2909
2018
Earlier work this paper cites.
A. Malinin and M. Gales, “Predictive uncertainty estimation via prior networks,” NeurIPS , 2018
2018
Earlier work this paper cites.
A. Radford, K. Narasimhan, T. Salimans, I. Sutskever et al. , “Improving language understanding by generative pre-training,” 2018
2018
Earlier work this paper cites.
M. Ott, M. Auli, D. Grangier, and M. Ranzato, “Analyzing uncertainty in neural machine translation,” in ICML , 2018, pp. 3956–3965
2018
Earlier work this paper cites.
T. Liu, X. Zhang, W. Zhou, and W. Jia, “Neural relation extraction via inner-sentence noise reduction and transfer learning,” in EMNLP , 2018, pp. 2195–2204
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
E. M. Bender and B. Friedman, “Data statements for natural language processing: Toward mitigating system bias and enabling better science,” TACL , pp. 587–604, 2018
2018
Earlier work this paper cites.
R. Panchendrarajan and A. Amaresan, “Bidirectional lstm-crf for named entity recognition,” in Proceedings of the 32nd Pacific Asia conference on language, information and computation , 2018
2018
Earlier work this paper cites.
K. Murray and D. Chiang, “Correcting length bias in neural machine translation,” in Proceedings of the Third Conference on Machine Translation: Research Papers , 2018, pp. 212–223
2018
Earlier work this paper cites.
V. Kuleshov, N. Fenner, and S. Ermon, “Accurate uncertainties for deep learning using calibrated regression,” in International conference on machine learning . PMLR, 2018, pp. 2796–2804
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
M. Sensoy, L. Kaplan, and M. Kandemir, “Evidential deep learning to quantify classification uncertainty,” NeurIPS , 2018
2018
Cited alongside, same era.
S. Ryu, S. Koo, H. Yu, and G. G. Lee, “Out-of-domain detection based on generative adversarial network,” in EMNLP , 2018, pp. 714–718
2018
Cited alongside, same era.
K. Lee, H. Lee, K. Lee, and J. Shin, “Training confidence-calibrated classifiers for detecting out-of-distribution samples,” in ICLR , 2018
2018
Cited alongside, same era.
Y. Xiao and W. Y. Wang, “Quantifying uncertainties in natural language processing tasks,” in AAAI , 2019, pp. 7322–7329
2019
Cited alongside, same era.
F. Guzmán, P.-J. Chen, M. Ott, J. Pino, G. Lample, P. Koehn, V. Chaudhary, and M. Ranzato, “The FLORES evaluation datasets for low-resource machine translation: Nepali–English and Sinhala–English,” in EMNLP-IJCNLP , Hong Kong, China, Nov. 2019, pp. 6098–6111. [Online]. Available: https://aclanthology.org/D19-1632
E. Hüllermeier and W. Waegeman, “Aleatoric and epistemic uncertainty in machine learning: An introduction to concepts and methods,” Machine Learning , pp. 457–506, 2021
2021
Later among the works it cites.
U. Arora, W. Huang, and H. He, “Types of out-of-distribution texts and how to detect them,” in EMNLP , 2021, pp. 10 687–10 701
2021
Later among the works it cites.
J. Xin, R. Tang, Y. Yu, and J. Lin, “The art of abstention: Selective prediction and error regularization for natural language processing,” in ACL-IJCNLP , 2021, pp. 1040–1051
2021
Later among the works it cites.
M. Li, M. Li, K. Xiong, and J. Lin, “Multi-task dense retrieval via model uncertainty fusion for open-domain question answering,” in EMNLP-Findings , 2021, pp. 274–287
2021
Later among the works it cites.
W. Zhou, F. Liu, and M. Chen, “Contrastive out-of-distribution detection for pretrained transformers,” in EMNLP , 2021, pp. 1100–1111
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
B. Poole, S. Ozair, A. Van Den Oord, A. Alemi, and G. Tucker, “On variational bounds of mutual information,” in ICML , 2019, pp. 5171–5180
2019
Cited alongside, same era.
Y. Ovadia, E. Fertig, J. Ren, Z. Nado, D. Sculley, S. Nowozin, J. Dillon, B. Lakshminarayanan, and J. Snoek, “Can you trust your model’s uncertainty? evaluating predictive uncertainty under dataset shift,” NeurIPS , 2019
2019
Cited alongside, same era.
2019
Cited alongside, same era.
A. Chaudhary, J. Xie, Z. Sheikh, G. Neubig, and J. Carbonell, “A little annotation does a lot of good: A study in bootstrapping low-resource named entity recognizers,” in EMNLP-IJCNLP , 2019, pp. 5164–5174
2019
Cited alongside, same era.
X. Zhang, F. Chen, C.-T. Lu, and N. Ramakrishnan, “Mitigating uncertainty in document classification,” in NAACL , 2019, pp. 3126–3136
2019
Cited alongside, same era.
D. Hendrycks, M. Mazeika, and T. Dietterich, “Deep anomaly detection with outlier exposure,” in ICLR , 2019
2019
Cited alongside, same era.
Q. Zhang, A. Lipani, S. Liang, and E. Yilmaz, “Reply-aided detection of misinformation via bayesian deep learning,” in WWW , 2019, p. 2333–2343
2019
Cited alongside, same era.
2021
Later among the works it cites.
J. Xin, R. Tang, Y. Yu, and J. Lin, “The art of abstention: Selective prediction and error regularization for natural language processing,” in ACL-IJCNLP , 2021, pp. 1040–1051
2021
Later among the works it cites.
K. Margatina, G. Vernikos, L. Barrault, and N. Aletras, “Active learning by acquiring contrastive examples,” in EMNLP , 2021, pp. 650–663
2021
Later among the works it cites.
A. Shelmanov, D. Puzyrev, L. Kupriyanova, D. Belyakov, D. Larionov, N. Khromov, O. Kozlova, E. Artemova, D. V. Dylov, and A. Panchenko, “Active learning for sequence tagging with deep pre-trained models and Bayesian uncertainty estimates,” in EACL , 2021, pp. 1698–1712
2021
Later among the works it cites.
Y. Shen, Y.-C. Hsu, A. Ray, and H. Jin, “Enhancing the generalization for intent classification and out-of-domain detection in slu,” in COLING , 2021, pp. 2443–2453
2021
Later among the works it cites.
Y. Hu and L. Khan, “Uncertainty-aware reliable text classification,” in SIGKDD , 2021, p. 628–636
2021
Later among the works it cites.
M. Wu, Y. Li, M. Zhang, L. Li, G. Haffari, and Q. Liu, “Uncertainty-aware balancing for multilingual and multi-domain neural machine translation training,” in EMNLP , 2021, pp. 7291–7305
2021
Later among the works it cites.
S. Garg and A. Moschitti, “Will this question be answered? question filtering via answer model distillation for efficient question answering,” in EMNLP , 2021, pp. 7329–7346
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
J. Xin, R. Tang, Y. Yu, and J. Lin, “BERxiT: Early exiting for BERT with better fine-tuning and extension to regression,” in EACL , 2021, pp. 91–104
2021
Later among the works it cites.
T. Schuster, A. Fisch, T. Jaakkola, and R. Barzilay, “Consistent accelerated inference via confident adaptive transformers,” in EMNLP , 2021, pp. 4962–4979
2021
Later among the works it cites.
X. Chen, M. Boratko, M. Chen, S. S. Dasgupta, X. L. Li, and A. McCallum, “Probabilistic box embeddings for uncertain knowledge graph reasoning,” in NAACL , 2021, pp. 882–893
2021
Later among the works it cites.
T. Glushkova, C. Zerva, R. Rei, and A. F. Martins, “Uncertainty-aware machine translation evaluation,” in EMNLP-Findings , 2021, pp. 3920–3938
2021
Later among the works it cites.
Z. Jiang, J. Araki, H. Ding, and G. Neubig, “How can we know when language models know? on the calibration of language models for question answering,” TACL , pp. 962–977, 2021
2021
Later among the works it cites.
T.-X. Sun, X.-Y. Liu, X.-P. Qiu, and X.-J. Huang, “Paradigm shift in natural language processing,” Machine Intelligence Research , vol. 19, no. 3, pp. 169–183, 2022
2022
Later among the works it cites.
J. Kivimäki et al. , “Uncertainty estimation with calibrated confidence scores,” 2022
2022
Later among the works it cites.
J. Pei, C. Wang, and G. Szarvas, “Transformer uncertainty estimation with hierarchical stochastic attention,” in AAAI , 2022, pp. 11 147–11 155
2022
Later among the works it cites.
C. Zerva, T. Glushkova, R. Rei, and A. F. Martins, “Disentangling uncertainty in machine translation evaluation,” in EMNLP , 2022, pp. 8622–8641
2022
Later among the works it cites.
A. Delaforge, J. Azé, S. Bringay, C. Mollevi, A. Sallaberry, and M. Servajean, “Ebbe-text: Explaining neural networks by exploring text classification decision boundaries,” IEEE Transactions on Visualization and Computer Graphics , 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
B. Qin, L. Wang, B. Hui, B. Li, X. Wei, B. Li, F. Huang, L. Si, M. Yang, and Y. Li, “SUN: Exploring intrinsic uncertainties in text-to-SQL parsers,” in COLING , 2022, pp. 5298–5308
2022
Later among the works it cites.
T. Schuster, A. Fisch, J. Gupta, M. Dehghani, D. Bahri, V. Tran, Y. Tay, and D. Metzler, “Confident adaptive language modeling,” NeurIPS , pp. 17 456–17 472, 2022
2022
Later among the works it cites.
W. Fedus, B. Zoph, and N. Shazeer, “Switch transformers: Scaling to trillion parameter models with simple and efficient sparsity,” The Journal of Machine Learning Research , vol. 23, no. 1, pp. 5232–5270, 2022
2022
Later among the works it cites.
A. Pandey, S. Daw, N. Unnam, and V. Pudi, “Multilinguals at SemEval-2022 task 11: Complex NER in semantically ambiguous settings for low resource languages,” in SemEval , 2022, pp. 1469–1476
2022
Later among the works it cites.
J. Li, A. Sun, J. Han, and C. Li, “A survey on deep learning for named entity recognition,” IEEE Trans. on Knowl. and Data Eng. , p. 50–70, 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
A. Malinin and M. Gales, “Uncertainty estimation in autoregressive structured prediction,” in ICLR , 2022
2022
Later among the works it cites.
Y. Wang, D. Beck, T. Baldwin, and K. Verspoor, “Uncertainty estimation and reduction of pre-trained models for text regression,” Transactions of the Association for Computational Linguistics , vol. 10, pp. 680–696, 2022. [Online]. Available: https://aclanthology.org/2022.tacl-1.39
2022
Later among the works it cites.
E. Zhu and J. Li, “Boundary smoothing for named entity recognition,” in ACL , 2022, pp. 7096–7108
2022
Later among the works it cites.
Y. Xiao, P. P. Liang, U. Bhatt, W. Neiswanger, R. Salakhutdinov, and L.-P. Morency, “Uncertainty quantification with pre-trained language models: A large-scale empirical analysis,” in EMNLP Findings , Dec. 2022, pp. 7273–7284
2022
Later among the works it cites.
V. Raina and M. Gales, “Answer uncertainty and unanswerability in multiple-choice machine reading comprehension,” in ACL-Findings , 2022, pp. 1020–1034
2022
Later among the works it cites.
J. S. Andersen and W. Maalej, “Efficient, uncertainty-based moderation of neural networks text classifiers,” in ACL-Findings , 2022, pp. 1536–1546
2022
Later among the works it cites.
A. Vazhentsev, G. Kuzmin, A. Shelmanov, A. Tsvigun, E. Tsymbalov, K. Fedyanin, M. Panov, A. Panchenko, G. Gusev, M. Burtsev, M. Avetisian, and L. Zhukov, “Uncertainty estimation of transformer predictions for misclassification detection,” in ACL , 2022, pp. 8237–8252
2022
Later among the works it cites.
Y. Yu, L. Kong, J. Zhang, R. Zhang, and C. Zhang, “AcTune: Uncertainty-based active self-training for active fine-tuning of pretrained language models,” in NAACL , 2022, pp. 1422–1436
2022
Later among the works it cites.
M. Liu, Z. Tu, T. Zhang, T. Su, X. Xu, and Z. Wang, “Ltp: a new active learning strategy for crf-based named entity recognition,” Neural Processing Letters , pp. 2433–2454, 2022
2022
Later among the works it cites.
A. Gidiotis and G. Tsoumakas, “Should we trust this summary? Bayesian abstractive summarization to the rescue,” in ACL-Findings , 2022, pp. 4119–4131
2022
Later among the works it cites.
N. Varshney, S. Mishra, and C. Baral, “Towards improving selective prediction ability of nlp systems,” in RepL4NLP , 2022, pp. 221–226
2022
Later among the works it cites.
L. Wei, D. Hu, W. Zhou, and S. Hu, “Uncertainty-aware propagation structure reconstruction for fake news detection,” in COLING , 2022, pp. 2759–2768
2022
Later among the works it cites.
Y. Li, L. Liu, and S. Shi, “Rethinking negative sampling for handling missing entity annotations,” in ACL , 2022, pp. 7188–7197
2022
Later among the works it cites.
S. Lin, J. Hilton, and O. Evans, “Teaching models to express their uncertainty in words,” 2022
2022
Later among the works it cites.
L. Kuhn, Y. Gal, and S. Farquhar, “Semantic uncertainty: Linguistic invariances for uncertainty estimation in natural language generation,” in ICLR , 2023
2023
Closest in time.
Y. Yu, H. Sajjad, and J. Xu, “Learning uncertainty for unknown domains with zero-target-assumption,” in ICLR , 2023
2023
Closest in time.
2023
Closest in time.
Z. Lin, D. Phan, P. Pasupat, J. Z. Liu, and J. Shang, “On compositional uncertainty quantification for seq2seq graph parsing,” in ICLR , 2023
2023
Closest in time.
2023
Closest in time.
D. Hendrycks, “Natural selection favors ais over humans,” 2023
2023
Closest in time.