Fetching the paper…
Reading the bibliography…
The increased deployment of LMs for real-world tasks involving knowledge and facts makes it important to understand model epistemology: what LMs think they know, and how their attitudes toward that knowledge are affected by language use in their inputs.
FACT , pages 143–173. De Gruyter Mouton, Berlin, Boston
Paul Kiparsky and Carol Kiparsky. 1970 · 1970
Earlier work this paper cites.
Some observations on factivity
Lauri Karttunen. 1971 · 1971
Earlier work this paper cites.
Hedges: A study in meaning criteria and the logic of fuzzy concepts
George Lakoff. 1975 · 1975
Earlier work this paper cites.
On hedging in physician-physician discourse
E. Prince, C. Bosk, and J. Frader. 1982 · 1982
Earlier work this paper cites.
Politeness: Some universals in language usage , volume 4
Penelope Brown and Stephen C Levinson. 1987 · 1987
Earlier work this paper cites.
Boosting, hedging and the negotiation of academic knowledge
Ken Hyland. 1998 · 1998
Earlier work this paper cites.
Evidentiality
Alexandra Y Aikhenvald. 2004 · 2004
Earlier work this paper cites.
Stance and engagement: A model of interaction in academic discourse
Ken Hyland. 2005 · 2005
Earlier work this paper cites.
Factbank: a corpus annotated with event factuality
Roser Saurí and James Pustejovsky. 2009 · 2009
Earlier work this paper cites.
What projects and why
Mandy Simons, Judith Tonhauser, David Beaver, and Craige Roberts. 2010 · 2010
Earlier work this paper cites.
Did it happen? the pragmatic complexity of veridicality assessment
Marie-Catherine de Marneffe, Christopher D. Manning, and Christopher Potts. 2012 · 2012
Earlier work this paper cites.
Disciplinary discourses: Writer stance in research articles
Ken Hyland. 2014 · 2014
Earlier work this paper cites.
Hedging and speaker commitment
Anna Prokofieva and Julia Hirschberg. 2014 · 2014
Earlier work this paper cites.
The role of explanations on trust and reliance in clinical decision support systems
Adrian Bussone, Simone Stumpf, and Dympna O’Sullivan. 2015 · 2015
Earlier work this paper cites.
TriviaQA: A large scale distantly supervised challenge dataset for reading comprehension
Mandar Joshi, Eunsol Choi, Daniel Weld, and Luke Zettlemoyer. 2017 · 2017
Earlier work this paper cites.
Integrating deep linguistic features in factuality prediction over unified datasets
Gabriel Stanovsky, Judith Eckle-Kohler, Yevgeniy Puzikov, Ido Dagan, and Iryna Gurevych. 2017 · 2017
Earlier work this paper cites.
Neural models of factuality
Rachel Rudinger, Aaron Steven White, and Benjamin Van Durme. 2018 · 2018
Earlier work this paper cites.
The commitmentbank: Investigating projection in naturally occurring discourse
Marie-Catherine De Marneffe, Mandy Simons, and Judith Tonhauser. 2019 · 2019
Cited alongside, same era.
Definitely, maybe: A new experimental paradigm for investigating the pragmatics of evidential devices across languages
Judith Degen, Andreas Trotzke, Gregory Scontras, Eva Wittenberg, and Noah D Goodman. 2019 · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al. 2019 · 2019
Cited alongside, same era.
Calibration of pre-trained transformers
Shrey Desai and Greg Durrett. 2020 · 2020
Cited alongside, same era.
The pile: An 800gb dataset of diverse text for language modeling
Leo Gao, Stella Biderman, Sid Black, Laurence Golding, Travis Hoppe, Charles Foster, Jason Phang, Horace He, Anish Thite, Noa Nabeshima, et al. 2020 · 2020
Cited alongside, same era.
Uncertainty estimation for language reward models
Adam Gleave and Geoffrey Irving. 2022 · 2022
Later among the works it cites.
Demystifying prompts in language models via perplexity estimation
Hila Gonen, Srini Iyer, Terra Blevins, Noah A Smith, and Luke Zettlemoyer. 2022 · 2022
Later among the works it cites.
Language models (mostly) know what they know
Saurav Kadavath, Tom Conerly, Amanda Askell, Tom Henighan, Dawn Drain, Ethan Perez, Nicholas Schiefer, Zac Hatfield Dodds, Nova DasSarma, Eli Tran-Johnson, et al. 2022 · 2022
Later among the works it cites.
PERSONACHATGEN: Generating personalized dialogues using GPT-3
Young-Jun Lee, Chae-Gyun Lim, Yunsu Choi, Ji-Hui Lm, and Ho-Jin Choi. 2022 · 2022
Later among the works it cites.
Teaching models to express their uncertainty in words
Stephanie C. Lin, Jacob Hilton, and Owain Evans. 2022 · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Selective question answering under domain shift
Amita Kamath, Robin Jia, and Percy Liang. 2020 · 2020
Cited alongside, same era.
Calibrated language model fine-tuning for in- and out-of-distribution data
Lingkai Kong, Haoming Jiang, Yuchen Zhuang, Jie Lyu, Tuo Zhao, and Chao Zhang. 2020 · 2020
Cited alongside, same era.
Automatically neutralizing subjective bias in text
Reid Pryzant, Richard Diehl Martinez, Nathan Dass, Sadao Kurohashi, Dan Jurafsky, and Diyi Yang. 2020 · 2020
Cited alongside, same era.
On the inference calibration of neural machine translation
Shuo Wang, Zhaopeng Tu, Shuming Shi, and Yang Liu. 2020 · 2020
Cited alongside, same era.
How machine-learning recommendations influence clinician treatment selections: the example of antidepressant selection
Maia Jacobs, Melanie F Pradier, Thomas H McCoy Jr, Roy H Perlis, Finale Doshi-Velez, and Krzysztof Z Gajos. 2021 · 2021
Cited alongside, same era.
He thinks he knows better than the doctors: BERT for event factuality fails on pragmatics
Nanjiang Jiang and Marie-Catherine de Marneffe. 2021 · 2021
Cited alongside, same era.
How can we know when language models know? on the calibration of language models for question answering
Zhengbao Jiang, Jun Araki, Haibo Ding, and Graham Neubig. 2021 · 2021
Cited alongside, same era.
Later among the works it cites.
Fantastically ordered prompts and where to find them: Overcoming few-shot prompt order sensitivity
Yao Lu, Max Bartolo, Alastair Moore, Sebastian Riedel, and Pontus Stenetorp. 2022 · 2022
Later among the works it cites.
Reducing conversational agents’ overconfidence through linguistic calibration
Sabrina J Mielke, Arthur Szlam, Emily Dinan, and Y-Lan Boureau. 2022 · 2022
Later among the works it cites.
Social simulacra: Creating populated prototypes for social computing systems
Joon Sung Park, Lindsay Popowski, Carrie Cai, Meredith Ringel Morris, Percy Liang, and Michael S Bernstein. 2022 · 2022
Later among the works it cites.
“You might think about slightly revising the title”: Identifying hedges in peer-tutoring interactions
Yann Raphalen, Chloé Clavel, and Justine Cassell. 2022 · 2022
Later among the works it cites.
Quantifying uncertainty in foundation models via ensembles
Meiqi Sun, Wilson Yan, Pieter Abbeel, and Igor Mordatch. 2022 · 2022
Later among the works it cites.
Prompt-and-rerank: A method for zero-shot and few-shot arbitrary textual style transfer with small language models
Mirac Suzgun, Luke Melas-Kyriazi, and Dan Jurafsky. 2022 · 2022
Later among the works it cites.
Semantic uncertainty: Linguistic invariances for uncertainty estimation in natural language generation
Lorenz Kuhn, Yarin Gal, and Sebastian Farquhar. 2023 · 2023
Closest in time.
Holistic evaluation of language models
Percy Liang, Rishi Bommasani, Tony Lee, Dimitris Tsipras, Dilara Soylu, Michihiro Yasunaga, Yian Zhang, Deepak Narayanan, Yuhuai Wu, Ananya Kumar, Benjamin Newman, Binhang Yuan, Bobby Yan, Ce Zhang, Christian Cosgrove, Christopher D. Manning, Christopher R’e, Diana Acosta-Navas, Drew A. Hudson, E. Zelikman, Esin Durmus, Faisal Ladhak, Frieda Rong, Hongyu Ren, Huaxiu Yao, Jue Wang, Keshav Santhanam, Laurel J. Orr, Lucia Zheng, Mert Yuksekgonul, Mirac Suzgun, Nathan S. Kim, Neel Guha, Niladri S. Chatterji, Omar Khattab, Peter Henderson, Qian Huang, Ryan Chi, Sang Michael Xie, Shibani Santurkar, Surya Ganguli, Tatsunori Hashimoto, Thomas F. Icard, Tianyi Zhang, Vishrav Chaudhary, William Wang, Xuechen Li, Yifan Mai, Yuhui Zhang, and Yuta Koreeda. 2023 · 2023
Closest in time.
Calibrated Interpretation: Confidence Estimation in Semantic Parsing
Elias Stengel-Eskin and Benjamin Van Durme. 2023 · 2023
Closest in time.
"according to …" prompting language models improves quoting from pre-training data
Orion Weller, Marc Marone, Nathaniel Weir, Dawn J. Lawrie, Daniel Khashabi, and Benjamin Van Durme. 2023 · 2023
Closest in time.
Calibrating structured output predictors for natural language processing
Abhyuday Jagannatha and Hong Yu. 2020 · 2092
Closest in time.