Fetching the paper…
Reading the bibliography…
Despite the impressive performance of large language models (LLMs) across various benchmarks, their ability to address ambiguously specified problems--frequent in real-world interactions--remains underexplored.
Answer-based Adversarial Training for Generating Clarification Questions, 2019
Sudha Rao and Hal Daumé III · 1904
Earlier work this paper cites.
Studies on the Semantics of Questions and the Pragmatics of Answers
Jeroen Groenendijk · 1984
Earlier work this paper cites.
Ai and legal reasoning
Edwina Rissland · 1988
Earlier work this paper cites.
Bayesian Methods for Adaptive Models
David John Cameron Mackay · 1992
Earlier work this paper cites.
Requirements engineering: social and technical issues
Marina Jirotka and Joseph A. Goguen (eds.) · 1994
Earlier work this paper cites.
Probabilistic robotics
Sebastian Thrun · 2002
Earlier work this paper cites.
Pathways to “evidence-informed” policy and practice: A framework for action
Shelley Bowen and Anthony B Zwi · 2005
Earlier work this paper cites.
Autotutor: An intelligent tutoring system with mixed-initiative dialogue
Arthur C. Graesser, Phoebe Chipman, Brian C. Haynes, and Andrew Olney · 2005
Earlier work this paper cites.
Assessing uncertainty in urban simulations using bayesian melding
Hana Ševčíková, Adrian E. Raftery, and Paul A. Waddell · 2006
Earlier work this paper cites.
Bayesian active learning for classification and preference learning
Neil Houlsby, Ferenc Huszár, Zoubin Ghahramani, and Máté Lengyel · 2011
Earlier work this paper cites.
Help helps, but only so much: Research on help seeking with intelligent tutoring systems
Vincent Aleven, Ido Roll, Bruce M. McLaren, and Kenneth R. Koedinger · 2016
Earlier work this paper cites.
Epistemic (meta) cognition: Ways of thinking about knowledge and knowing
Sarit Barzilai and Anat Zohar · 2016
Earlier work this paper cites.
A survey of motion planning and control techniques for self-driving urban vehicles
Brian Paden, Michal Čáp, Sze Zheng Yong, Dmitry Yershov, and Emilio Frazzoli · 2016
Earlier work this paper cites.
Design and Analysis of Experiments
Douglas C. Montgomery · 2017
Earlier work this paper cites.
Advances in financial machine learning
Marcos Lopez De Prado · 2018
Earlier work this paper cites.
Learning to Ask Good Questions: Ranking Clarification Questions using Neural Expected Value of Perfect Information
Sudha Rao and Hal Daumé III · 2018
Earlier work this paper cites.
Guidelines for human-ai interaction
Saleema Amershi, Daniel Weld, Mihaela Vorvoreanu, Adam Fourney, Besmira Nushi, Phillip Collisson, et al · 2019
Earlier work this paper cites.
Deep Medicine: How Artificial Intelligence Can Make Healthcare Human Again
Eric J. Topol · 2019
Earlier work this paper cites.
AmbigQA: Answering Ambiguous Open-domain Questions
Sewon Min, Julian Michael, Hannaneh Hajishirzi, and Luke Zettlemoyer · 2020
Earlier work this paper cites.
Evaluating Large Language Models Trained on Code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, Alex Ray, Raul Puri, Gretchen Krueger, Michael Petrov, Heidy Khlaaf, Girish Sastry, Pamela Mishkin, Brooke Chan, Scott Gray, Nick Ryder, Mikhail Pavlov, Alethea Power, Lukasz Kaiser, Mohammad Bavarian, Clemens Winter, Philippe Tillet, Felipe Petroski Such, Dave Cummings, Matthias Plappert, Fotios Chantzis, Elizabeth Barnes, Ariel Herbert-Voss, William Hebgen Guss, Alex Nichol, Alex Paino, Nikolas Tezak, Jie Tang, Igor Babuschkin, Suchir Balaji, Shantanu Jain, William Saunders, Christopher Hesse, Andrew N. Carr, Jan Leike, Josh Achiam, Vedant Misra, Evan Morikawa, Alec Radford, Matthew Knight, Miles Brundage, Mira Murati, Katie Mayer, Peter Welinder, Bob McGrew, Dario Amodei, Sam McCandlish, Ilya Sutskever, and Wojciech Zaremba · 2021
Cited alongside, same era.
Wordcraft: A human-AI collaborative editor for story writing
Andy Coenen, Luke Davis, Daphne Ippolito, Emily Reif, and Ann Yuan · 2021
Cited alongside, same era.
Measuring Coding Challenge Competence With APPS
Dan Hendrycks, Steven Basart, Saurav Kadavath, Mantas Mazeika, Akul Arora, Ethan Guo, Collin Burns, Samir Puranik, Horace He, Dawn Song, and Jacob Steinhardt · 2021
Cited alongside, same era.
Bayesian Preference Elicitation with Keyphrase-Item Coembeddings for Interactive Recommendation
Hojin Yang, Scott Sanner, Ga Wu, and Jin Peng Zhou · 2021
Active Learning Principles for In-Context Learning with Large Language Models
Katerina Margatina, Timo Schick, Nikolaos Aletras, and Jane Dwivedi-Yu · 2023
Later among the works it cites.
Asking Clarifying Questions using Language Models and Probabilistic Reasoning
Top Piriyakulkij, Volodymyr Kuleshov, and Kevin Ellis · 2023
Later among the works it cites.
Direct Preference Optimization: Your Language Model is Secretly a Reward Model
Rafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D Manning, Stefano Ermon, and Chelsea Finn · 2023
Later among the works it cites.
Modern Bayesian Experimental Design, 2023
Tom Rainforth, Adam Foster, Desi R. Ivanova, and Freddie Bickford Smith · 2023
Later among the works it cites.
A survey of hallucination in large foundation models
Vipula Rawte, Amit Sheth, and Amitava Das · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Training a helpful and harmless assistant with reinforcement learning from human feedback
Yuntao Bai, Andy Jones, Kamal Ndousse, Amanda Askell, Anna Chen, Nova DasSarma, Dawn Drain, Stanislav Fort, Deep Ganguli, Tom Henighan, and others · 2022
Cited alongside, same era.
Selection-inference: Exploiting large language models for interpretable logical reasoning
Antonia Creswell, Murray Shanahan, and Irina Higgins · 2022
Cited alongside, same era.
Assistance with large language models
Dmitrii Krasheninnikov, Egor Krasheninnikov, and David Krueger · 2022
Cited alongside, same era.
Clam: Selective clarification for ambiguous questions with large language models
Lorenz Kuhn, Yarin Gal, and Sebastian Farquhar · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul F Christiano, Jan Leike, and Ryan Lowe · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, and others · 2022
Cited alongside, same era.
Data-driven ai emergency planning in process industry
Fengli Zhang, Qianzhe Qiao, Jinjiang Wang, and Pinpin Liu · 2022
Cited alongside, same era.
Active Example Selection for In-Context Learning
Yiming Zhang, Shi Feng, and Chenhao Tan · 2022
Cited alongside, same era.
Why Did the Chicken Cross the Road? Rephrasing and Analyzing Ambiguous Questions in VQA
Elias Stengel-Eskin, Jimena Guallar-Blasco, Yi Zhou, and Benjamin Van Durme · 2023
Later among the works it cites.
Task Ambiguity in Humans and Language Models
Alex Tamkin, Kunal Handa, Avash Shrestha, and Noah Goodman · 2023
Later among the works it cites.
Self-Consistency Improves Chain of Thought Reasoning in Language Models
Xuezhi Wang, Jason Wei, Dale Schuurmans, Quoc V. Le, Ed H. Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou · 2023
Later among the works it cites.
Tree of Thoughts: Deliberate Problem Solving with Large Language Models, December 2023
Shunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran, Thomas L. Griffiths, Yuan Cao, and Karthik Narasimhan · 2023
Later among the works it cites.
STar-GATE: Teaching language models to ask clarifying questions
Chinmaya Andukuri, Jan-Philipp Fränken, Tobias Gerstenberg, and Noah Goodman · 2024
Later among the works it cites.
Bayesian Optimization with LLM-Based Acquisition Functions for Natural Language Preference Elicitation, 2024
David Eric Austin, Anton Korikov, Armin Toroghi, and Scott Sanner · 2024
Later among the works it cites.
Active Prompting with Chain-of-Thought for Large Language Models, 2024
Shizhe Diao, Pengcheng Wang, L. I. N. Yong, Rui Pan, Xiang Liu, and Tong Zhang · 2024
Later among the works it cites.
Bayesian Preference Elicitation with Language Models, 2024
Kunal Handa, Yarin Gal, Ellie Pavlick, Noah Goodman, Jacob Andreas, Alex Tamkin, and Belinda Z. Li · 2024
Later among the works it cites.
Uncertainty of thoughts: Uncertainty-aware planning enhances information seeking in large language models, 2024
Zhiyuan Hu, Chumin Liu, Xidong Feng, Yilun Zhao, See-Kiong Ng, Anh Tuan Luu, Junxian He, Pang Wei Koh, and Bryan Hooi · 2024
Later among the works it cites.
Is your code generated by chatgpt really correct? rigorous evaluation of large language models for code generation
Jiawei Liu, Chunqiu Steven Xia, Yuyao Wang, and Lingming Zhang · 2024
Later among the works it cites.
Probabilistic risk assessment in cybersecurity: Bayesian methods for quantifying and mitigating cyber risks
Fulsundar Amita Purushottam, Ajay Kumar, Vikas Haribhau Satonkar, Shweta Gaikwad, Stefi Diliprao Sonawane, and Bhushan shirwadkar · 2024
Later among the works it cites.
Mathematical discoveries from program search with large language models
Bernardino Romera-Paredes, Mohammadamin Barekatain, Alexander Novikov, Matej Balog, M Pawan Kumar, Emilien Dupont, Francisco JR Ruiz, Jordan S Ellenberg, Pengming Wang, Omar Fawzi, and others · 2024
Later among the works it cites.
Towards Automated Knowledge Integration From Human-Interpretable Representations
Kasia Kobalczyk and Mihaela van der Schaar · 2025
Closest in time.