Fetching the paper…
Reading the bibliography…
Evaluating the quality of automatically generated question items has been a long standing challenge.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 1901
Earlier work this paper cites.
Coefficient alpha and the internal structure of tests
Lee J Cronbach. 1951 · 1951
Earlier work this paper cites.
Bloom’s taxonomy of educational objectives
TAXONOMY MADE EASY BLOOM’S. 1965 · 1965
Earlier work this paper cites.
The 2 sigma problem: The search for methods of group instruction as effective as one-to-one tutoring
Benjamin S Bloom. 1984 · 1984
Earlier work this paper cites.
Comparison of 1-, 2-, and 3-parameter IRT models
Deborah Harris. 1989 · 1989
Earlier work this paper cites.
The role of deliberate practice in the acquisition of expert performance
K Anders Ericsson, Ralf T Krampe, and Clemens Tesch-Römer. 1993 · 1993
Earlier work this paper cites.
Intelligent tutoring goes to school in the big city
Kenneth R Koedinger, John R Anderson, William H Hadley, and Mary A Mark. 1997 · 1997
Earlier work this paper cites.
Peer instruction: Ten years of experience and results
Catherine H Crouch and Eric Mazur. 2001 · 2001
Earlier work this paper cites.
Descriptive and explanatory item response models
Mark Wilson and Paul De Boeck. 2004 · 2004
Earlier work this paper cites.
The influence of experience and deliberate practice on the development of superior expert performance
K Anders Ericsson et al · 2006
Earlier work this paper cites.
Designing interactions . Vol. 17
Bill Moggridge and Bill Atkinson. 2007 · 2007
Earlier work this paper cites.
Review of recent systems for automatic assessment of programming assignments. In Proceedings of the 10th Koli calling international conference on computing education research . 86–93
Petri Ihantola, Tuukka Ahoniemi, Ville Karavirta, and Otto Seppälä. 2010 · 2010
Earlier work this paper cites.
The Knowledge-Learning-Instruction framework: Bridging the science-practice chasm to enhance robust student learning
Kenneth R Koedinger, Albert T Corbett, and Charles Perfetti. 2012 · 2012
Earlier work this paper cites.
The ICAP framework: Linking cognitive engagement to active learning outcomes
Michelene TH Chi and Ruth Wylie. 2014 · 2014
Earlier work this paper cites.
Automatic selection of informative sentences: The sentences that can generate multiple choice questions
Mukta Majumder and Sujan Kumar Saha. 2014 · 2014
Earlier work this paper cites.
Assessing item difficulty and discrimination indices of teacher-developed multiple-choice tests. In Assessment for Learning Within and Beyond the Classroom: Taylor’s 8th Teaching and Learning Conference 2015 Proceedings . Springer, 417–426
Ahmad Zamri Khairani and Hasni Shamsuddin. 2016 · 2015
Earlier work this paper cites.
Learning is not a spectator sport: Doing is better than watching for learning from a MOOC. In Proceedings of the second (2015) ACM conference on learning@ scale . ACM, 111–120
Kenneth R Koedinger, Jihee Kim, Julianna Zhuxin Jia, Elizabeth A McLaughlin, and Norman L Bier. 2015a · 2015
Earlier work this paper cites.
Learning is not a spectator sport: Doing is better than watching for learning from a MOOC. In Proceedings of the second (2015) ACM conference on learning@ scale . ACM, 111–120
Kenneth R Koedinger, Jihee Kim, Julianna Zhuxin Jia, Elizabeth A McLaughlin, and Norman L Bier. 2015b · 2015
Earlier work this paper cites.
A System for Generating Multiple Choice Questions: With a Novel Approach for Sentence Selection. In NLP-TEA@ACL/IJCNLP
Mukta Majumder and Sujan Kumar Saha. 2015 · 2015
Earlier work this paper cites.
Ontology-based multiple choice question generation
Tahani Alsubait, Bijan Parsia, and Ulrike Sattler. 2016 · 2016
Cited alongside, same era.
AXIS: Generating Explanations at Scale with Learnersourcing and Machine Learning. In Proceedings of the Third (2016) ACM Conference on Learning @ Scale (Edinburgh, Scotland, UK) (L@S ’16) . ACM, New York, NY, USA, 379–388
Joseph Jay Williams, Juho Kim, Anna Rafferty, Samuel Maldonado, Krzysztof Z. Gajos, Walter S. Lasecki, and Neil Heffernan. 2016 · 2016
Cited alongside, same era.
Multiple choice question generation utilizing an ontology. In Proceedings of the 12th Workshop on Innovative Use of NLP for Building Educational Applications . 303–312
Katherine Stasaski and Marti A Hearst. 2017 · 2017
Cited alongside, same era.
Evaluation of academic performance based on learning analytics and ontology: a systematic mapping study. In 2018 IEEE Frontiers in Education Conference (FIE) . IEEE, 1–5
Laecio A Costa, Laís N Salvador, and Ricardo R Amorim. 2018 · 2018
Cited alongside, same era.
ReadingQuizMaker: A Human-NLP Collaborative System that Supports Instructors to Design High-Quality Reading Quiz Questions. In Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems . 1–18
Xinyi Lu, Simin Fan, Jessica Houghton, Lu Wang, and Xu Wang. 2023 · 2023
Later among the works it cites.
GPTeach: Interactive TA Training with GPT Based Students
Julia M Markel, Steven G Opferman, James A Landay, and Chris Piech. 2023 · 2023
Later among the works it cites.
Crowdsourcing the Evaluation of Multiple-Choice Questions Using Item-Writing Flaws and Bloom’s Taxonomy
Steven Moore, Huy A Nguyen, Ellen Fang, and John Stamper. 2023b · 2023
Later among the works it cites.
Prompting ai art: An investigation into the creative skill of prompt engineering
Jonas Oppenlaender, Rhema Linder, and Johanna Silvennoinen. 2023 · 2023
Later among the works it cites.
Human-like problem-solving abilities in large language models using ChatGPT
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Predicting question quality using recurrent neural networks. In Artificial Intelligence in Education: 19th International Conference, AIED 2018, London, UK, June 27–30, 2018, Proceedings, Part I 19 . Springer, 491–502
Stefan Ruseti, Mihai Dascalu, Amy M Johnson, Renu Balyan, Kristopher J Kopp, Danielle S McNamara, Scott A Crossley, and Stefan Trausan-Matu. 2018 · 2018
Cited alongside, same era.
QG-Net: A Data-Driven Question Generation Model for Educational Content. In Proceedings of the Fifth Annual ACM Conference on Learning at Scale (London, United Kingdom) (L@S ’18) . Association for Computing Machinery, New York, NY, USA, Article 7, 10 pages
Zichao Wang, Andrew S. Lan, Weili Nie, Andrew E. Waters, Phillip J. Grimaldi, and Richard G. Baraniuk. 2018 · 2018
Cited alongside, same era.
Measuring actual learning versus feeling of learning in response to being actively engaged in the classroom
Louis Deslauriers, Logan S McCarty, Kelly Miller, Kristina Callaghan, and Greg Kestin. 2019 · 2019
Cited alongside, same era.
UpGrade: Sourcing Student Open-Ended Solutions to Create Scalable Learning Opportunities. In Proceedings of the Sixth (2019) ACM Conference on Learning @ Scale (Chicago, IL, USA) (L@S ’19) . ACM, New York, NY, USA, Article 17, 10 pages
Xu Wang, Srinivasa Teja Talluri, Carolyn Rose, and Kenneth Koedinger. 2019 · 2019
Cited alongside, same era.
A Systematic Review of Automatic Question Generation for Educational Purposes
Ghader Kurdi, Jared Leo, Bijan Parsia, Uli Sattler, and Salam Al-Emari. 2020 · 2020
Cited alongside, same era.
QMaps: Engaging Students in Voluntary Question Generation and Linking (CHI ’20) . Association for Computing Machinery, New York, NY, USA, 1–14
Iman Yeckehzaare, Tirdad Barghi, and Paul Resnick. 2020 · 2020
Cited alongside, same era.
Automatic question generation and answer assessment: a survey
Bidyut Das, Mukta Majumder, Santanu Phadikar, and Arif Ahmed Sekh. 2021 · 2021
Cited alongside, same era.
What’s In It for the Learners? Evidence from a Randomized Field Experiment on Learnersourcing Questions in a MOOC. In Proceedings of the Eighth ACM Conference on Learning@ Scale . 221–233
Anjali Singh, Christopher Brooks, Yiwen Lin, and Warren Li. 2021 · 2021
Cited alongside, same era.
Graziella Orrù, Andrea Piarulli, Ciro Conversano, and Angelo Gemignani. 2023 · 2023
Later among the works it cites.
Generative agents: Interactive simulacra of human behavior. In Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology . 1–22
Joon Sung Park, Joseph O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S Bernstein. 2023 · 2023
Later among the works it cites.
Rehearsal: Simulating conflict to teach conflict resolution
Omar Shaikh, Valentino Chai, Michele J Gelfand, Diyi Yang, and Michael S Bernstein. 2023 · 2023
Later among the works it cites.
Large Language Models Can Be Easily Distracted by Irrelevant Context. In Proceedings of the 40th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 202) , Andreas Krause, Emma Brunskill, Kyunghyun Cho, Barbara Engelhardt, Sivan Sabato, and Jonathan Scarlett (Eds.). PMLR, 31210–31227
Freda Shi, Xinyun Chen, Kanishka Misra, Nathan Scales, David Dohan, Ed H. Chi, Nathanael Schärli, and Denny Zhou. 2023 · 2023
Later among the works it cites.
Voyager: An open-ended embodied agent with large language models
Guanzhi Wang, Yuqi Xie, Yunfan Jiang, Ajay Mandlekar, Chaowei Xiao, Yuke Zhu, Linxi Fan, and Anima Anandkumar. 2023c · 2023
Later among the works it cites.
A survey on large language model based autonomous agents
Lei Wang, Chen Ma, Xueyang Feng, Zeyu Zhang, Hao Yang, Jingsen Zhang, Zhiyuan Chen, Jiakai Tang, Xu Chen, Yankai Lin, et al · 2023
Later among the works it cites.
Rolellm: Benchmarking, eliciting, and enhancing role-playing abilities of large language models
Zekun Moore Wang, Zhongyuan Peng, Haoran Que, Jiaheng Liu, Wangchunshu Zhou, Yuhan Wu, Hongcheng Guo, Ruitong Gan, Zehao Ni, Man Zhang, et al · 2023
Later among the works it cites.
Leveraging generative artificial intelligence to simulate student learning behavior
Songlin Xu and Xinyu Zhang. 2023 · 2023
Later among the works it cites.
Exploring large language models for communication games: An empirical study on werewolf
Yuzhuang Xu, Shuo Wang, Peng Li, Fuwen Luo, Xiaolong Wang, Weidong Liu, and Yang Liu. 2023 · 2023
Later among the works it cites.
Sotopia: Interactive evaluation for social intelligence in language agents
Xuhui Zhou, Hao Zhu, Leena Mathur, Ruohong Zhang, Haofei Yu, Zhengyang Qi, Louis-Philippe Morency, Yonatan Bisk, Daniel Fried, Graham Neubig, et al · 2023
Later among the works it cites.
Incorporating ChatGPT Into Your Teaching
[n. d.] · 2024
Closest in time.
OpenAI Chat
[n. d.] · 2024
Closest in time.
Using ChatGPT to Write Quiz Questions
[n. d.] · 2024
Closest in time.
Teach AI How to Code: Using Large Language Models as Teachable Agents for Programming Education. In Proceedings of the CHI Conference on Human Factors in Computing Systems . 1–28
Hyoungwook Jin, Seonghee Lee, Hyungyu Shin, and Juho Kim. 2024 · 2024
Closest in time.