Fetching the paper…
Reading the bibliography…
Recent advances of powerful Language Models have allowed Natural Language Generation (NLG) to emerge as an important technology that can not only perform traditional tasks like summarisation or translation, but also serve as a natural language interface to a variety of applications.
Calibration of encoder decoder models for neural machine translation
Aviral Kumar and Sunita Sarawagi. 2019 · 1903
Earlier work this paper cites.
Quality of uncertainty quantification for bayesian neural network inference
Jiayu Yao, Weiwei Pan, Soumya Ghosh, and Finale Doshi-Velez. 2019 · 1906
Earlier work this paper cites.
Stop measuring calibration when humans disagree
Joris Baan, Wilker Aziz, Barbara Plank, and Raquel Fernandez. 2022 · 1915
Earlier work this paper cites.
Risk, uncertainty and profit , volume 31
Frank Hyneman Knight. 1921 · 1921
Earlier work this paper cites.
The meaning of meaning: A study of the influence of thought and of the science of symbolism
Charles Kay Ogden and Ivor Armstrong Richards. 1923 · 1923
Earlier work this paper cites.
The foundations of mathematics and other logical essays
Frank Plumpton Ramsey. 1931 · 1931
Earlier work this paper cites.
Theory of games and economic behavior, 2nd rev
John Von Neumann and Oskar Morgenstern. 1947 · 1947
Earlier work this paper cites.
Computing Machinery and Intelligence
Alan Turing. 1950 · 1950
Earlier work this paper cites.
Statistical decision functions
Abraham Wald. 1951 · 1951
Earlier work this paper cites.
Modality as referential multiplicity
Jaakko Hintikka. 1957 · 1957
Earlier work this paper cites.
Foundations of the Theory of Probability , 2 edition
Andrey N. Kolmogorov. 1960 · 1960
Earlier work this paper cites.
Modality and quantification
Jaakko Hintikka. 1961 · 1961
Earlier work this paper cites.
Eliza—a computer program for the study of natural language communication between man and machine
Joseph Weizenbaum. 1966 · 1966
Earlier work this paper cites.
The foundations of statistics
Leonard J Savage. 1972 · 1972
Earlier work this paper cites.
Theory of probability: A critical introductory treatment , volume 6
Bruno De Finetti. 1974 · 1974
Earlier work this paper cites.
The emergence of probability: A philosophical study of early ideas about probability, induction and statistical inference
Ian Hacking. 1975 · 1975
Earlier work this paper cites.
A Mathematical Theory of Evidence
Glenn Shafer. 1976 · 1976
Earlier work this paper cites.
Selective-lama: Selective prediction for confidence-aware evaluation of language models
Hiyori Yoshikawa and Naoaki Okazaki. 2023 · 1983
Earlier work this paper cites.
An Introduction to Possibilistic and Fuzzy Logics
D. Dubois and H. Prade. 1990 · 1990
Earlier work this paper cites.
Uncertainty: A Guide to Dealing with Uncertainty in Quantitative Risk and Policy Analysis
Millett Granger Morgan and Max Henrion. 1990 · 1990
Earlier work this paper cites.
Rank-based systems: A simple approach to belief revision, belief update, and reasoning about evidence and actions
Moisés Goldszmidt and Judea Pearl. 1992 · 1992
Earlier work this paper cites.
The fisher, neyman-pearson theories of testing hypotheses: one theory or two?
Erich L Lehmann. 1993 · 1993
Earlier work this paper cites.
Speaking: From intention to articulation
Willem JM Levelt. 1993 · 1993
Earlier work this paper cites.
Bayesian theory , volume 405
José M Bernardo and Adrian FM Smith. 1994 · 1994
Earlier work this paper cites.
A sequential algorithm for training text classifiers: Corrigendum and additional data
David D Lewis. 1995 · 1995
Earlier work this paper cites.
Plausibility measures and default reasoning
Nir Friedman and Joseph Y. Halpern. 1996 · 1996
Earlier work this paper cites.
Theory of Point Estimation , second edition
Erich L. Lehmann and George Casella. 1998 · 1998
Earlier work this paper cites.
Psycholinguistics
Thomas Scovel. 1998 · 1998
Earlier work this paper cites.
An introduction to variational methods for graphical models
Michael I Jordan, Zoubin Ghahramani, Tommi S Jaakkola, and Lawrence K Saul. 1999 · 1999
Earlier work this paper cites.
Monte Carlo statistical methods , volume 2
Christian P Robert, George Casella, and George Casella. 1999 · 1999
Earlier work this paper cites.
Effect of ambiguity and lexical availability on syntactic and lexical production
Victor S Ferreira and Gary S Dell. 2000 · 2000
Earlier work this paper cites.
Building Natural Language Generation Systems
Ehud Reiter and Robert Dale. 2000 · 2000
Earlier work this paper cites.
Inductive confidence machines for regression
Harris Papadopoulos, Kostas Proedrou, Volodya Vovk, and Alex Gammerman. 2002 · 2002
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Bayesian Data Analysis , 2nd ed. edition
Andrew Gelman, John B. Carlin, Hal S. Stern, and Donald B. Rubin. 2004 · 2004
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
METEOR: An automatic metric for MT evaluation with improved correlation with human judgments
Satanjeev Banerjee and Alon Lavie. 2005 · 2005
Earlier work this paper cites.
Testing statistical hypotheses , third edition
E. L. Lehmann and Joseph P. Romano. 2005 · 2005
Earlier work this paper cites.
Algorithmic learning in a random world , volume 29
Vladimir Vovk, Alexander Gammerman, and Glenn Shafer. 2005 · 2005
Earlier work this paper cites.
Evaluation of text generation: A survey
Asli Celikyilmaz, Elizabeth Clark, and Jianfeng Gao. 2020 · 2006
Earlier work this paper cites.
Gaussian processes for machine learning , volume 1
Carl Edward Rasmussen, Christopher KI Williams, et al. 2006 · 2006
Earlier work this paper cites.
Wat zei je? detecting out-of-distribution translations with variational transformers
Tim Z Xiao, Aidan N Gomez, and Yarin Gal. 2020 · 2006
Earlier work this paper cites.
Aleatory or epistemic? does it matter?
Armen Der Kiureghian and Ove Ditlevsen. 2009 · 2009
Earlier work this paper cites.
Formal representations of uncertainty
Didier Dubois and Henri Prade. 2009 · 2009
Earlier work this paper cites.
On the foundations of noise-free selective classification
Ran El-Yaniv et al. 2010 · 2010
Earlier work this paper cites.
A review of uncertainty quantification in deep learning: Techniques, applications and challenges
Moloud Abdar, Farhad Pourpanah, Sadiq Hussain, Dana Rezazadegan, Li Liu, Mohammad Ghavamzadeh, Paul W. Fieguth, Xiaochun Cao, Abbas Khosravi, U. Rajendra Acharya, Vladimir Makarenkov, and Saeid Nahavandi. 2020 · 2011
Earlier work this paper cites.
Bayesian active learning for classification and preference learning
Neil Houlsby, Ferenc Huszár, Zoubin Ghahramani, and Máté Lengyel. 2011 · 2011
Earlier work this paper cites.
Active learning and crowdsourcing for machine translation in low resource scenarios
Vamshi Ambati. 2012 · 2012
Earlier work this paper cites.
The era of cognitive systems: An inside look at ibm watson and how it works
Rob High. 2012 · 2012
Earlier work this paper cites.
A unifying view on dataset shift in classification
Jose G Moreno-Torres, Troy Raeder, Rocío Alaiz-Rodríguez, Nitesh V Chawla, and Francisco Herrera. 2012 · 2012
Earlier work this paper cites.
Understanding uncertainty
Dennis V Lindley. 2013 · 2013
Earlier work this paper cites.
Truth is a lie: Crowd truth and the seven myths of human annotation
Lora Aroyo and Chris Welty. 2015 · 2015
Earlier work this paper cites.
Weight uncertainty in neural network
Charles Blundell, Julien Cornebise, Koray Kavukcuoglu, and Daan Wierstra. 2015 · 2015
Earlier work this paper cites.
Obtaining well calibrated probabilities using bayesian binning
Mahdi Pakdaman Naeini, Gregory Cooper, and Milos Hauskrecht. 2015 · 2015
Earlier work this paper cites.
Generating sentences from a continuous space
Samuel R. Bowman, Luke Vilnis, Oriol Vinyals, Andrew Dai, Rafal Jozefowicz, and Samy Bengio. 2016 · 2016
Earlier work this paper cites.
A theoretically grounded application of dropout in recurrent neural networks
Yarin Gal and Zoubin Ghahramani. 2016 · 2016
Earlier work this paper cites.
A diversity-promoting objective function for neural conversation models
Jiwei Li, Michel Galley, Chris Brockett, Jianfeng Gao, and Bill Dolan. 2016 · 2016
Earlier work this paper cites.
Controlling politeness in neural machine translation via side constraints
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016a · 2016
Earlier work this paper cites.
Minimum risk training for neural machine translation
Shiqi Shen, Yong Cheng, Zhongjun He, Wei He, Hua Wu, Maosong Sun, and Yang Liu. 2016 · 2016
Earlier work this paper cites.
Variational neural machine translation
Biao Zhang, Deyi Xiong, Jinsong Su, Hong Duan, and Min Zhang. 2016 · 2016
Earlier work this paper cites.
Towards neural machine translation with latent tree attention
James Bradbury and Richard Socher. 2017 · 2017
Earlier work this paper cites.
Deep reinforcement learning from human preferences
Paul F Christiano, Jan Leike, Tom Brown, Miljan Martic, Shane Legg, and Dario Amodei. 2017 · 2017
Earlier work this paper cites.
Bayesian recurrent neural networks
Meire Fortunato, Charles Blundell, and Oriol Vinyals. 2017 · 2017
Earlier work this paper cites.
Scalable Bayesian learning of recurrent neural networks for language modeling
Zhe Gan, Chunyuan Li, Changyou Chen, Yunchen Pu, Qinliang Su, and Lawrence Carin. 2017 · 2017
Cited alongside, same era.
On calibration of modern neural networks
Chuan Guo, Geoff Pleiss, Yu Sun, and Kilian Q Weinberger. 2017 · 2017
Cited alongside, same era.
Reasoning about uncertainty
Joseph Y Halpern. 2017 · 2017
Cited alongside, same era.
Exploiting cross-sentence context for neural machine translation
Longyue Wang, Zhaopeng Tu, Andy Way, and Qun Liu. 2017 · 2017
Cited alongside, same era.
Latent alignment and variational attention
Yuntian Deng, Yoon Kim, Justin Chiu, Demi Guo, and Alexander Rush. 2018 · 2018
Cited alongside, same era.
Hierarchical neural story generation
Angela Fan, Mike Lewis, and Yann Dauphin. 2018 · 2018
Cited alongside, same era.
A survey of natural language generation
Chenhe Dong, Yinghui Li, Haifan Gong, Miaoxin Chen, Junxin Li, Ying Shen, and Min Yang. 2022 · 2022
Later among the works it cites.
Neural natural language generation: A survey on multilinguality, multimodality, controllability and learning
Erkut Erdem, Menekse Kuyu, Semih Yagcioglu, Anette Frank, Letitia Parcalabescu, Barbara Plank, Andrii Babii, Oleksii Turuta, Aykut Erdem, Iacer Calixto, Elena Lloret, Elena-Simona Apostol, Ciprian-Octavian Truică, Branislava Šandrih, Sanda Martinčić-Ipšić, Gábor Berend, Albert Gatt, and Grăzina Korvel. 2022 · 2022
Later among the works it cites.
Quality-aware decoding for neural machine translation
Patrick Fernandes, António Farinhas, Ricardo Rei, José G. C. de Souza, Perez Ogayo, Graham Neubig, and Andre Martins. 2022 · 2022
Later among the works it cites.
Should we trust this summary? Bayesian abstractive summarization to the rescue
Alexios Gidiotis and Grigorios Tsoumakas. 2022 · 2022
Later among the works it cites.
Good-enough language production
Adele E Goldberg and Fernanda Ferreira. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Survey of the state of the art in natural language generation: Core tasks, applications and evaluation
Albert Gatt and Emiel Krahmer. 2018 · 2018
Cited alongside, same era.
A deep reinforced model for abstractive summarization
Romain Paulus, Caiming Xiong, and Richard Socher. 2018 · 2018
Cited alongside, same era.
Deep Bayesian active learning for natural language processing: Results of a large-scale empirical study
Aditya Siddhant and Zachary C. Lipton. 2018 · 2018
Cited alongside, same era.
Improving the transformer translation model with document-level context
Jiacheng Zhang, Huanbo Luan, Maosong Sun, Feifei Zhai, Jingfang Xu, Min Zhang, and Yang Liu. 2018a · 2018
Cited alongside, same era.
Generating informative and diverse conversational responses via adversarial information maximization
Yizhe Zhang, Michel Galley, Jianfeng Gao, Zhe Gan, Xiujun Li, Chris Brockett, and Bill Dolan. 2018c · 2018
Cited alongside, same era.
Improving tree-lstm with tree attention
Mahtab Ahmed, Muhammad Rifayat Samee, and Robert E Mercer. 2019 · 2019
Cited alongside, same era.
Truncation sampling as language model desmoothing
John Hewitt, Christopher Manning, and Percy Liang. 2022 · 2022
Later among the works it cites.
State-of-the-art generalisation research in nlp: a taxonomy and review
Dieuwke Hupkes, Mario Giulianelli, Verna Dankers, Mikel Artetxe, Yanai Elazar, Tiago Pimentel, Christos Christodoulopoulos, Karim Lasri, Naomi Saphra, Arabella Sinclair, et al. 2022 · 2022
Later among the works it cites.
Survey of hallucination in natural language generation
Ziwei Ji, Nayeon Lee, Rita Frieske, Tiezheng Yu, Dan Su, Yan Xu, Etsuko Ishii, Yejin Bang, Andrea Madotto, and Pascale Fung. 2022 · 2022
Later among the works it cites.
Investigating Reasons for Disagreement in Natural Language Inference
Nan-Jiang Jiang and Marie-Catherine de Marneffe. 2022 · 2022
Later among the works it cites.
Language models (mostly) know what they know
Saurav Kadavath, Tom Conerly, Amanda Askell, Tom Henighan, Dawn Drain, Ethan Perez, Nicholas Schiefer, Zac Hatfield Dodds, Nova DasSarma, Eli Tran-Johnson, et al. 2022 · 2022
Later among the works it cites.
Adaptive label smoothing with self-knowledge in natural language generation
Dongkyu Lee, Ka Chun Cheung, and Nevin Zhang. 2022 · 2022
Later among the works it cites.
Teaching models to express their uncertainty in words
Stephanie Lin, Jacob Hilton, and Owain Evans. 2022 · 2022
Later among the works it cites.
Reducing conversational agents’ overconfidence through linguistic calibration
Sabrina J Mielke, Arthur Szlam, Emily Dinan, and Y-Lan Boureau. 2022 · 2022
Later among the works it cites.
Cross-task generalization via natural language crowdsourcing instructions
Swaroop Mishra, Daniel Khashabi, Chitta Baral, and Hannaneh Hajishirzi. 2022 · 2022
Later among the works it cites.
Introducing chatgpt
OpenAI. 2022 · 2022
Later among the works it cites.
Fine-tuning language models via epistemic neural networks
Ian Osband, Seyed Mohammad Asghari, Benjamin Van Roy, Nat McAleese, John Aslanides, and Geoffrey Irving. 2022 · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Gray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul Christiano, Jan Leike, and Ryan Lowe. 2022 · 2022
Later among the works it cites.
The “problem” of human label variation: On ground truth in data, modeling and evaluation
Barbara Plank. 2022 · 2022
Later among the works it cites.
Mutual information alleviates hallucinations in abstractive summarization
Liam van der Poel, Ryan Cotterell, and Clara Meister. 2022 · 2022
Later among the works it cites.
Bayesformer: Transformer with uncertainty estimation
Karthik Abinav Sankararaman, Sinong Wang, and Han Fang. 2022 · 2022
Later among the works it cites.
Jam or cream first? modeling ambiguity in neural machine translation with SCONES
Felix Stahlberg and Shankar Kumar. 2022 · 2022
Later among the works it cites.
Rethinking document-level neural machine translation
Zewei Sun, Mingxuan Wang, Hao Zhou, Chengqi Zhao, Shujian Huang, Jiajun Chen, and Lei Li. 2022 · 2022
Later among the works it cites.
Uniform versus uncertainty sampling: When being active is less efficient than staying passive
Alexandru Tifrea, Jacob Clarysse, and Fanny Yang. 2022 · 2022
Later among the works it cites.
Exploring predictive uncertainty and calibration in NLP: A study on the impact of method & data scarcity
Dennis Ulmer, Jes Frellsen, and Christian Hardmeier. 2022 · 2022
Later among the works it cites.
Benchmarking scalable predictive uncertainty in text classification
Jordy Van Landeghem, Matthew Blaschko, Bertrand Anckaert, and Marie-Francine Moens. 2022 · 2022
Later among the works it cites.
Finetuned language models are zero-shot learners
Jason Wei, Maarten Bosma, Vincent Zhao, Kelvin Guu, Adams Wei Yu, Brian Lester, Nan Du, Andrew M. Dai, and Quoc V Le. 2022 · 2022
Later among the works it cites.
An explanation of in-context learning as implicit bayesian inference
Sang Michael Xie, Aditi Raghunathan, Percy Liang, and Tengyu Ma. 2022 · 2022
Later among the works it cites.
Disentangling uncertainty in machine translation evaluation
Chrysoula Zerva, Taisiya Glushkova, Ricardo Rei, and André F. T. Martins. 2022 · 2022
Later among the works it cites.
Sparks of artificial general intelligence: Early experiments with gpt-4
Sébastien Bubeck, Varun Chandrasekaran, Ronen Eldan, Johannes Gehrke, Eric Horvitz, Ece Kamar, Peter Lee, Yin Tat Lee, Yuanzhi Li, Scott Lundberg, et al. 2023 · 2023
Closest in time.
A close look into the calibration of pre-trained language models
Yangyi Chen, Lifan Yuan, Ganqu Cui, Zhiyuan Liu, and Heng Ji. 2023 · 2023
Closest in time.
Should you marginalize over possible tokenizations?
Nadezhda Chirkova, Germán Kruszewski, Jos Rozen, and Marc Dymetman. 2023 · 2023
Closest in time.
Selectively answering ambiguous questions
Jeremy R Cole, Michael JQ Zhang, Daniel Gillick, Julian Martin Eisenschlos, Bhuwan Dhingra, and Jacob Eisenstein. 2023 · 2023
Closest in time.
Markus Freitag, Behrooz Ghorbani, and Patrick Fernandes. 2023 · 2023
Closest in time.
Pal: Program-aided language models
Luyu Gao, Aman Madaan, Shuyan Zhou, Uri Alon, Pengfei Liu, Yiming Yang, Jamie Callan, and Graham Neubig. 2023 · 2023
Closest in time.
Repairing the cracked foundation: A survey of obstacles in evaluation practices for generated text
Sebastian Gehrmann, Elizabeth Clark, and Thibault Sellam. 2023 · 2023
Closest in time.
Evaluating machine translation quality with conformal predictive distributions
Patrizio Giovannotti. 2023 · 2023
Closest in time.
Mario Giulianelli, Joris Baan, Wilker Aziz, Raquel Fernández, and Barbara Plank. 2023 · 2023
Closest in time.
An important next step on our ai journey
Google. 2023 · 2023
Closest in time.
Sources of uncertainty in machine learning–a statisticians’ view
Cornelia Gruber, Patrick Oliver Schenk, Malte Schierholz, Frauke Kreuter, and Göran Kauermann. 2023 · 2023
Closest in time.
Looking for a needle in a haystack: A comprehensive study of hallucinations in neural machine translation
Nuno M. Guerreiro, Elena Voita, and André Martins. 2023 · 2023
Closest in time.
A theory of emergent in-context learning as implicit structure induction
Michael Hahn and Navin Goyal. 2023 · 2023
Closest in time.
On the blind spots of model-based evaluation metrics for text generation
Tianxing He, Jingyu Zhang, Tianle Wang, Sachin Kumar, Kyunghyun Cho, James Glass, and Yulia Tsvetkov. 2023 · 2023
Closest in time.
Uncertainty in natural language processing: Sources, quantification, and applications
Mengting Hu, Zhen Zhang, Shiwan Zhao, Minlie Huang, and Bingzhe Wu. 2023 · 2023
Closest in time.
Ru-sure? uncertainty-aware code suggestions by maximizing utility across random user intents
Daniel D Johnson, Daniel Tarlow, and Christian Walder. 2023 · 2023
Closest in time.
Semantic uncertainty: Linguistic invariances for uncertainty estimation in natural language generation
Lorenz Kuhn, Yarin Gal, and Sebastian Farquhar. 2023 · 2023
Closest in time.
Ai transparency in the age of llms: A human-centered research roadmap
Q Vera Liao and Jennifer Wortman Vaughan. 2023 · 2023
Closest in time.
Generating with confidence: Uncertainty quantification for black-box large language models
Zhen Lin, Shubhendu Trivedi, and Jimeng Sun. 2023 · 2023
Closest in time.
Possible Worlds
Christopher Menzel. 2023 · 2023
Closest in time.
Don’t blame the annotator: Bias already starts in the annotation instructions
Mihir Parmar, Swaroop Mishra, Mor Geva, and Chitta Baral. 2023 · 2023
Closest in time.
On the usefulness of embeddings, clusters and strings for text generation evaluation
Tiago Pimentel, Clara Meister, and Ryan Cotterell. 2023 · 2023
Closest in time.
Victor Quach, Adam Fisch, Tal Schuster, Adam Yala, Jae Ho Sohn, Tommi S Jaakkola, and Regina Barzilay. 2023 · 2023
Closest in time.
Conformal nucleus sampling
Shauli Ravfogel, Yoav Goldberg, and Jacob Goldberger. 2023 · 2023
Closest in time.
Why don’t you do it right? analysing annotators’ disagreement in subjective tasks
Marta Sandri, Elisa Leonardelli, Sara Tonelli, and Elisabetta Jezek. 2023 · 2023
Closest in time.
Toolformer: Language models can teach themselves to use tools
Timo Schick, Jane Dwivedi-Yu, Roberto Dessì, Roberta Raileanu, Maria Lomeli, Luke Zettlemoyer, Nicola Cancedda, and Thomas Scialom. 2023 · 2023
Closest in time.
Prompting GPT-3 to be reliable
Chenglei Si, Zhe Gan, Zhengyuan Yang, Shuohang Wang, Jianfeng Wang, Jordan Lee Boyd-Graber, and Lijuan Wang. 2023 · 2023
Closest in time.
Zero and few-shot semantic parsing with ambiguous inputs
Elias Stengel-Eskin, Kyle Rawlins, and Benjamin Van Durme. 2023 · 2023
Closest in time.
Data feedback loops: Model-driven amplification of dataset biases
Rohan Taori and Tatsunori Hashimoto. 2023 · 2023
Closest in time.
Veniamin Veselovsky, Manoel Horta Ribeiro, and Robert West. 2023 · 2023
Closest in time.
React: Synergizing reasoning and acting in language models
Shunyu Yao, Jeffrey Zhao, Dian Yu, Nan Du, Izhak Shafran, Karthik R. Narasimhan, and Yuan Cao. 2023 · 2023
Closest in time.
Polina Zablotskaia, Du Phan, Joshua Maynez, Shashi Narayan, Jie Ren, and Jeremiah Liu. 2023 · 2023
Closest in time.
Conformalizing machine translation evaluation
Chrysoula Zerva and André FT Martins. 2023 · 2023
Closest in time.
Calibrating sequence likelihood improves conditional language generation
Yao Zhao, Misha Khalman, Rishabh Joshi, Shashi Narayan, Mohammad Saleh, and Peter J. Liu. 2023 · 2023
Closest in time.