Fetching the paper…
Reading the bibliography…
Large language models (LLMs) excel at answering questions but remain passive learners-absorbing static data without the ability to question and refine knowledge.
Taxonomy of educational objectives. handbook i: Cognitive domain
Benjamin Samuel Bloom and David R. Krathwohl. 1966 · 1966
Earlier work this paper cites.
Mind in society: Development of higher psychological processes
Lev Semenovich Vygotsky and Michael Cole. 1978 · 1978
Earlier work this paper cites.
The 2 sigma problem: The search for methods of group instruction as effective as one-to-one tutoring
Benjamin S. Bloom. 1984 · 1984
Earlier work this paper cites.
A theory of questions and question asking
Ashwin Ram. 1991 · 1991
Earlier work this paper cites.
Towards scalable dataset construction: An active learning approach
Brendan Collins, Jia Deng, Kai Li, and Li Fei-Fei. 2008 · 2008
Earlier work this paper cites.
doc2dial: A goal-oriented document-grounded dialogue dataset
Song Feng, Hui Wan, R. Chulaka Gunasekara, Siva Sankalp Patel, Sachindra Joshi, and Luis A. Lastras. 2020 · 2011
Earlier work this paper cites.
The relative effectiveness of human tutoring, intelligent tutoring systems, and other tutoring systems
Kurt VanLehn. 2011 · 2011
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick. 2014 · 2014
Earlier work this paper cites.
Distilling the knowledge in a neural network
Geoffrey Hinton. 2015 · 2015
Earlier work this paper cites.
Joint concept learning and semantic parsing from natural language explanations
Shashank Srivastava, Igor Labutov, and Tom Mitchell. 2017 · 2017
Earlier work this paper cites.
Training classifiers with natural language explanations
Braden Hancock, Paroma Varma, Stephanie Wang, Martin Bringmann, Percy Liang, and Christopher Ré. 2018 · 2018
Cited alongside, same era.
Learning to ask good questions: Ranking clarification questions using neural expected value of perfect information
Sudha Rao and Hal Daumé III. 2018 · 2018
Cited alongside, same era.
Learning to ask for conversational machine learning
Shashank Srivastava, Igor Labutov, and Tom Mitchell. 2019 · 2019
Cited alongside, same era.
Conversational learning
Forough Arabshahi, Kathryn Mazaitis, Toby Jia-Jun Li, Brad A Myers, and Tom Mitchell. 2020 · 2020
Cited alongside, same era.
How much knowledge can you pack into the parameters of a language model?
Adam Roberts, Colin Raffel, and Noam Shazeer. 2020 · 2020
Cited alongside, same era.
Multidoc2dial: Modeling dialogues grounded in multiple documents
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al. 2022 · 2022
Later among the works it cites.
Large language models are reasoning teachers
Namgyu Ho, Laura Schmid, and Se-Young Yun. 2023 · 2023
Later among the works it cites.
Eliciting human preferences with language models
Belinda Z Li, Alex Tamkin, Noah Goodman, and Jacob Andreas. 2023 · 2023
Later among the works it cites.
Self-refine: Iterative refinement with self-feedback
Aman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan, Luyu Gao, Sarah Wiegreffe, Uri Alon, Nouha Dziri, Shrimai Prabhumoye, Yiming Yang, et al. 2023 · 2023
Later among the works it cites.
Direct preference optimization: Your language model is secretly a reward model
Rafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D Manning, Stefano Ermon, and Chelsea Finn. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Song Feng, Siva Sankalp Patel, Hui Wan, and Sachindra Joshi. 2021 · 2021
Cited alongside, same era.
Can language models learn from explanations in context?
Andrew Lampinen, Ishita Dasgupta, Stephanie Chan, Kory Mathewson, Mh Tessler, Antonia Creswell, James McClelland, Jane Wang, and Felix Hill. 2022 · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al. 2022 · 2022
Cited alongside, same era.
CLUES: A benchmark for learning classifiers using natural language explanations
Rakesh R. Menon, Sayan Ghosh, and Shashank Srivastava. 2022 · 2022
Cited alongside, same era.
Active learning helps pretrained models learn the intended task
Alex Tamkin, Dat Pham Nguyen, Salil Deshpande, Jesse Mu, and Noah Goodman. 2022 · 2022
Cited alongside, same era.
Let the llms talk: Simulating human-to-human conversational qa via zero-shot llm-to-llm interactions
Zahra Abbasiantaeb, Yifei Yuan, Evangelos Kanoulas, and Mohammad Aliannejadi. 2024 · 2024
Closest in time.
Teaching large language models to self-debug
Xinyun Chen, Maxwell Lin, Nathanael Schärli, and Denny Zhou. 2024 · 2024
Closest in time.
Bayesian preference elicitation with language models
Kunal Handa, Yarin Gal, Ellie Pavlick, Noah Goodman, Jacob Andreas, Alex Tamkin, and Belinda Z Li. 2024 · 2024
Closest in time.
Mediq: Question-asking LLMs and a benchmark for reliable interactive clinical reasoning
Shuyue Stella Li, Vidhisha Balachandran, Shangbin Feng, Jonathan S. Ilgen, Emma Pierson, Pang Wei Koh, and Yulia Tsvetkov. 2024 · 2024
Closest in time.
Toward in-context teaching: Adapting examples to students’ misconceptions
Alexis Ross and Jacob Andreas. 2024 · 2024
Closest in time.