Fetching the paper…
Reading the bibliography…
High-quality distractors are crucial to both the assessment and pedagogical value of multiple-choice questions (MCQs), where manually crafting ones that anticipate knowledge deficiencies or misconceptions among real students is difficult.
Bugs are not enough: Empirical studies of bugs, impasses and repairs in procedural skills
Kurt VanLehn. 1982 · 1982
Earlier work this paper cites.
Educational assessment of students
Anthony J. Nitko. 1996 · 1996
Earlier work this paper cites.
Classroom assessment: Concepts and applications
Peter Airasian. 2001 · 2001
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Cognitive tutor: Applied research in mathematics education
Steven Ritter, John R Anderson, Kenneth R Koedinger, and Albert Corbett. 2007 · 2007
Earlier work this paper cites.
Automatic distractor generation for multiple choice questions in standard tests
Zhaopeng Qiu, Xian Wu, and Wei Fan. 2020 · 2011
Earlier work this paper cites.
Automatic generation and delivery of multiple-choice math quizzes
Ana Paula Tomás and José Paulo Leal. 2013 · 2013
Earlier work this paper cites.
Generating multiple choice questions from ontologies: Lessons learnt
Tahani Alsubait, Bijan Parsia, and Uli Sattler. 2014 · 2014
Earlier work this paper cites.
The ASSISTments ecosystem: Building a platform that brings scientists and teachers together for minimally invasive research on human learning and teaching
Neil T Heffernan and Cristina Lindquist Heffernan. 2014 · 2014
Earlier work this paper cites.
Auto-Encoding Variational Bayes
Diederik P. Kingma and Max Welling. 2014 · 2014
Earlier work this paper cites.
Educational testing and measurement
Tom Kubiszyn and Gary Borich. 2016 · 2016
Earlier work this paper cites.
Categorical reparameterization with gumbel-softmax
Eric Jang, Shixiang Gu, and Ben Poole. 2017 · 2017
Earlier work this paper cites.
Multiple choice question generation utilizing an ontology
Katherine Stasaski and Marti A Hearst. 2017 · 2017
Earlier work this paper cites.
Automatic distractor generation for multiple-choice english vocabulary questions
Yuni Susanti, Takenobu Tokunaga, Hitoshi Nishikawa, and Hiroyuki Obari. 2018 · 2018
Earlier work this paper cites.
Diverse beam search: Decoding diverse solutions from neural sequence models
Ashwin K Vijayakumar, Michael Cogswell, Ramprasath R. Selvaraju, Qing Sun, Stefan Lee, David Crandall, and Dhruv Batra. 2018 · 2018
Earlier work this paper cites.
Generating distractors for reading comprehension questions from real examinations
Yifan Gao, Lidong Bing, Piji Li, Irwin King, and Michael R Lyu. 2019 · 2019
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter. 2019 · 2019
Earlier work this paper cites.
Skill-based career path modeling and recommendation
Aritra Ghosh, Beverly Woolf, Shlomo Zilberstein, and Andrew Lan. 2020 · 2020
Earlier work this paper cites.
Optimus: Organizing sentences via pre-trained modeling of a latent space
Chunyuan Li, Xiang Gao, Yuan Li, Baolin Peng, Xiujun Li, Yizhe Zhang, and Jianfeng Gao. 2020 · 2020
Earlier work this paper cites.
Mpnet: Masked and permuted pre-training for language understanding
Kaitao Song, Xu Tan, Tao Qin, Jianfeng Lu, and Tie-Yan Liu. 2020 · 2020
Cited alongside, same era.
Results and insights from diagnostic questions: The neurips 2020 education challenge
Zichao Wang, Angus Lamb, Evgeny Saveliev, Pashmina Cameron, Jordan Zaykov, Jose Miguel Hernandez-Lobato, Richard E Turner, Richard G Baraniuk, Craig Barton, Simon Peyton Jones, et al. 2021 · 2020
Cited alongside, same era.
Transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Remi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander Rush. 2020 · 2020
Cited alongside, same era.
Bertscore: Evaluating text generation with bert
Tianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger, and Yoav Artzi. 2020 · 2020
Cited alongside, same era.
Math multiple choice question solving and distractor generation with attentional gru networks
Q-genius: A gpt based modified mcq generator for identifying learner deficiency
Vijay Prakash, Kartikay Agrawal, and Syaamantak Das. 2023 · 2023
Later among the works it cites.
Qdg: A unified model for automatic question-distractor pairs generation
Pengju Shuai, Li Li, Sishun Liu, and Jun Shen. 2023 · 2023
Later among the works it cites.
Deduction under perturbed evidence: Probing student simulation (knowledge tracing) capabilities of large language models
Shashank Sonkar and Richard G Baraniuk. 2023 · 2023
Later among the works it cites.
Generating multiple choice questions for computing courses using large language models
Andrew Tran, Kenneth Angelikas, Egi Rama, Chiku Okechukwu, David H Smith, and Stephen MacNeil. 2023 · 2023
Later among the works it cites.
Distractor generation based on Text2Text language models with pseudo Kullback-Leibler divergence regulation
Hui-Juan Wang, Kai-Yu Hsieh, Han-Cheng Yu, Jui-Ching Tsou, Yu An Shih, Chen-Hua Huang, and Yao-Chung Fan. 2023a · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Neisarg Dave, Riley Bakes, Barton Pursel, and C Lee Giles. 2021 · 2021
Cited alongside, same era.
Dmytro Kalpakchi and Johan Boye. 2021 · 2021
Cited alongside, same era.
Diverse distractor generation for constructing high-quality multiple choice questions
Jiayuan Xie, Ningxin Peng, Yi Cai, Tao Wang, and Qingbao Huang. 2021 · 2021
Cited alongside, same era.
Cdgp: Automatic cloze distractor generation based on pre-trained language model
Shang-Hsuan Chiang, Ssu-Cheng Wang, and Yao-Chung Fan. 2022 · 2022
Cited alongside, same era.
LoRA: Low-rank adaptation of large language models
Edward J Hu, yelong shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2022 · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al. 2022 · 2022
Cited alongside, same era.
Rajkumar Ramamurthy, Prithviraj Ammanabrolu, Kianté Brantley, Jack Hessel, Rafet Sifa, Christian Bauckhage, Hannaneh Hajishirzi, and Yejin Choi. 2022 · 2022
Cited alongside, same era.
End-to-end generation of multiple-choice questions using text-to-text transfer transformer models
Ricardo Rodriguez-Torrealba, Eva Garcia-Lopez, and Antonio Garcia-Cabot. 2022 · 2022
Cited alongside, same era.
Metamath: Bootstrap your own mathematical questions for large language models
Longhui Yu, Weisen Jiang, Han Shi, Jincheng Yu, Zhengying Liu, Yu Zhang, James T Kwok, Zhenguo Li, Adrian Weller, and Weiyang Liu. 2023 · 2023
Later among the works it cites.
Khanmigo: Khan academy’s ai-powered teaching assistant
Khan Academy. 2024 · 2024
Closest in time.
Distractor generation for multiple-choice questions: A survey of methods, datasets, and evaluation
Elaf Alhazmi, Quan Z Sheng, Wei Emma Zhang, Munazza Zaib, and Ahoud Alhazmi. 2024 · 2024
Closest in time.
Qlora: Efficient finetuning of quantized llms
Tim Dettmers, Artidoro Pagnoni, Ari Holtzman, and Luke Zettlemoyer. 2024 · 2024
Closest in time.
Exploring automated distractor generation for math multiple-choice questions via large language models
Wanyong Feng, Jaewook Lee, Hunter McNichols, Alexander Scarlatos, Digory Smith, Simon Woodhead, Nancy Otero Ornelas, and Andrew Lan. 2024 · 2024
Closest in time.
Towards responsible development of generative ai for education: An evaluation-driven approach
Jurenka et al. 2024 · 2024
Closest in time.
Math multiple choice question generation via human-large language model collaboration
Jaewook Lee, Digory Smith, Simon Woodhead, and Andrew Lan. 2024 · 2024
Closest in time.
Can large language models replicate its feedback on open-ended math questions?
Hunter McNichols, Jaewook Lee, Stephen Fancsali, Steve Ritter, and Andrew Lan. 2024 · 2024
Closest in time.
An automatic question usability evaluation toolkit
Steven Moore, Eamon Costello, Huy A Nguyen, and John Stamper. 2024 · 2024
Closest in time.
Does writing with language models reduce content diversity?
Vishakh Padmakumar and He He. 2024 · 2024
Closest in time.
Fanyi Qu, Hao Sun, and Yunfang Wu. 2024 · 2024
Closest in time.
Improving automated distractor generation for math multiple-choice questions with overgenerate-and-rank
Alexander Scarlatos, Wanyong Feng, Digory Smith, Simon Woodhead, and Andrew Lan. 2024a · 2024
Closest in time.
Standardizing the measurement of text diversity: A tool and a comparative analysis of scores
Chantal Shaib, Joe Barrow, Jiuding Sun, Alexa F Siu, Byron C Wallace, and Ani Nenkova. 2024 · 2024
Closest in time.