Fetching the paper…
Reading the bibliography…
Standardised tests using short answer questions (SAQs) are common in postgraduate education.
“Cognitive Differences in Proficient and Nonproficient Essay Scorers - EDWARD W. WOLFE, CHI-WEN KAO, MICHAEL RANNEY, 1998.” https://journals.sagepub.com/doi/abs/10.1177/0741088398015004002
1998
Earlier work this paper cites.
E. M. Wolfe, “Uncovering Rater’s Cognitive Processing and Focus Using Think-Aloud Protocols,” Journal of Writing Assessment
2005
Earlier work this paper cites.
Assessing Performance: Designing, Scoring, and Validating Performance Tasks, New York, NY, US: The Guilford Press, 2009
R. L. Johnson, J. A. Penny, and B. Gordon, Assessing Performance: Designing, Scoring, and Validating Performance Tasks · 2009
Earlier work this paper cites.
I. I. Bejar, “Rater Cognition: Implications for Validity,” Educational Measurement: Issues and Practice
2012
Earlier work this paper cites.
C. Brew and C. Leacock, “Automated short answer scoring: Principles and prospects,” in Handbook of Automated Essay Evaluation: Current Applications and New Directions
2013
Earlier work this paper cites.
M. J. Peeters, S. A. Beltyukova, and B. A. Martin, “Educational Testing and Validity of Conclusions in the Scholarship of Teaching and Learning,” American Journal of Pharmaceutical Education
2013
Earlier work this paper cites.
S. Burrows, I. Gurevych, and B. Stein, “The Eras and Trends of Automatic Short Answer Grading,” International Journal of Artificial Intelligence in Education
2015
Earlier work this paper cites.
S. Burrows, I. Gurevych, and B. Stein, “The Eras and Trends of Automatic Short Answer Grading,” International Journal of Artificial Intelligence in Education
2015
Earlier work this paper cites.
J. Camacho-Collados and M. T. Pilehvar, “From Word to Sense Embeddings: A Survey on Vector Representations of Meaning,” Oct. 2018
2018
Earlier work this paper cites.
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, S. Agarwal, A. Herbert-Voss, G. Krueger, T. Henighan, R. Child, A. Ramesh, D. M. Ziegler, J. Wu, C. Winter, C. Hesse, M. Chen, E. Sigler, M. Litwin, S. Gray, B. Chess, J. Clark, C. Berner, S. McCandlish, A. Radford, I. Sutskever, and D. Amodei, “Language Models are Few-Shot Learners,” July 2020
2020
Earlier work this paper cites.
P. Liu, W. Yuan, J. Fu, Z. Jiang, H. Hayashi, and G. Neubig, “Pre-train, Prompt, and Predict: A Systematic Survey of Prompting Methods in Natural Language Processing,” July 2021
2021
Earlier work this paper cites.
A. Doewes and M. Pechenizkiy, “On the Limitations of Human-Computer Agreement in Automated Essay Scoring,” 2021
2021
Earlier work this paper cites.
J. Liu, D. Shen, Y. Zhang, B. Dolan, L. Carin, and W. Chen, “What Makes Good In-Context Examples for GPT-3?,” in Proceedings of Deep Learning Inside Out (DeeLIO 2022): The 3rd Workshop on Knowledge Extraction and Integration for Deep Learning Architectures
2022
Earlier work this paper cites.
E. Kasneci, K. Sessler, S. Küchemann, M. Bannert, D. Dementieva, F. Fischer, U. Gasser, G. Groh, S. Günnemann, E. Hüllermeier, S. Krusche, G. Kutyniok, T. Michaeli, C. Nerdel, J. Pfeffer, O. Poquet, M. Sailer, A. Schmidt, T. Seidel, M. Stadler, J. Weller, J. Kuhn, and G. Kasneci, “ChatGPT for good? On opportunities and challenges of large language models for education,” Learning and Individual Differences
2023
Cited alongside, same era.
R. A. Khan, M. Jawaid, A. R. Khan, and M. Sajjad, “ChatGPT - Reshaping medical education and clinical management,” Pakistan Journal of Medical Sciences
2023
Cited alongside, same era.
A. Mizumoto and M. Eguchi, “Exploring the potential of using an AI language model for automated essay scoring,” Research Methods in Applied Linguistics
2023
Cited alongside, same era.
A. Obata, T. Tagawa, and Y. Ono, “Assessment of ChatGPT’s Validity in Scoring Essays by Foreign Language Learners of Japanese and English,” in 2023 15th International Congress on Advanced Applied Informatics Winter (IIAI-AAI-Winter)
A. Hajikhani and C. Cole, “A critical review of large language models: Sensitivity, bias, and the path toward specialized AI,” Quantitative Science Studies
2024
Later among the works it cites.
M. I. Baig and E. Yadegaridehkordi, “ChatGPT in the higher education: A systematic literature review and research challenges,” International Journal of Educational Research
2024
Later among the works it cites.
J. Davis, L. Van Bulck, B. N. Durieux, and C. Lindvall, “The Temperature Feature of ChatGPT: Modifying Creativity for Clinical Research,” JMIR Human Factors
2024
Later among the works it cites.
M. Le and M. Davis, “ChatGPT Yields a Passing Score on a Pediatric Board Preparatory Exam but Raises Red Flags,” Global Pediatric Health
2024
Later among the works it cites.
L. Morjaria, L. Burns, K. Bracken, A. J. Levinson, Q. N. Ngo, M. Lee, and M. Sibbald, “Examining the Efficacy of ChatGPT in Marking Short-Answer Assessments in an Undergraduate Medical Program,” International Medical Education
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
T. H. Kung, M. Cheatham, A. Medenilla, C. Sillos, L. De Leon, C. Elepaño, M. Madriaga, R. Aggabao, G. Diaz-Candido, J. Maningo, and V. Tseng, “Performance of ChatGPT on USMLE: Potential for AI-assisted medical education using large language models,” PLOS Digital Health
2023
Cited alongside, same era.
F. Antaki, S. Touma, D. Milad, J. El-Khoury, and R. Duval, “Evaluating the Performance of ChatGPT in Ophthalmology: An Analysis of Its Successes and Shortcomings,” Ophthalmology Science
2023
Cited alongside, same era.
R. Bhayana, S. Krishna, and R. R. Bleakney, “Performance of ChatGPT on a Radiology Board-style Examination: Insights into Current Strengths and Limitations,” Radiology
2023
Cited alongside, same era.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention Is All You Need,” Aug. 2023
2023
Cited alongside, same era.
J. Wei, X. Wang, D. Schuurmans, M. Bosma, B. Ichter, F. Xia, E. Chi, Q. Le, and D. Zhou, “Chain-of-Thought Prompting Elicits Reasoning in Large Language Models,” Jan. 2023
2023
Cited alongside, same era.
M. Gao, J. Ruan, R. Sun, X. Yin, S. Yang, and X. Wan, “Human-like Summarization Evaluation with ChatGPT,” Apr. 2023
2023
Cited alongside, same era.
Feb. 2024
C. Impey, M. Wenger, N. Garuda, S. Golchin, and S. Stamer, Using Large Language Models for Automated Grading of Student Writing about Science · 2024
Cited alongside, same era.
P. M. Jackaria, B. H. Hajan, and A.-R. H. Mastul, “A Comparative Analysis of the Rating of College Students’ Essays by ChatGPT versus Human Raters,” International Journal of Learning, Teaching and Educational Research
2024
Cited alongside, same era.
2024
Later among the works it cites.
W. X. Zhao, K. Zhou, J. Li, T. Tang, X. Wang, Y. Hou, Y. Min, B. Zhang, J. Zhang, Z. Dong, Y. Du, C. Yang, Y. Chen, Z. Chen, J. Jiang, R. Ren, Y. Li, X. Tang, Z. Liu, P. Liu, J.-Y. Nie, and J.-R. Wen, “A Survey of Large Language Models,” Oct. 2024
2024
Later among the works it cites.
S. Schulhoff, M. Ilie, N. Balepur, K. Kahadze, A. Liu, C. Si, Y. Li, A. Gupta, H. Han, S. Schulhoff, P. S. Dulepet, S. Vidyadhara, D. Ki, S. Agrawal, C. Pham, G. Kroiz, F. Li, H. Tao, A. Srivastava, H. D. Costa, S. Gupta, M. L. Rogers, I. Goncearenco, G. Sarli, I. Galynker, D. Peskoff, M. Carpuat, J. White, S. Anadkat, A. Hoyle, and P. Resnik, “The Prompt Report: A Systematic Survey of Prompting Techniques,” Dec. 2024
2024
Later among the works it cites.
S. Minaee, T. Mikolov, N. Nikzad, M. Chenaghlu, R. Socher, X. Amatriain, and J. Gao, “Large Language Models: A Survey,” Feb. 2024
2024
Later among the works it cites.
H. Naveed, A. U. Khan, S. Qiu, M. Saqib, S. Anwar, M. Usman, N. Akhtar, N. Barnes, and A. Mian, “A Comprehensive Overview of Large Language Models,” Oct. 2024
2024
Later among the works it cites.
“Evaluating the quality of AI feedback: A comparative study of AI and human essay grading: Innovations in Education and Teaching International: Vol 0, No 0 - Get Access.” https://www.tandfonline.com/doi/full/10.1080/14703297.2024.2437122?af=R
2024
Later among the works it cites.
S. Kapoor, N. Gruver, M. Roberts, K. Collins, A. Pal, U. Bhatt, A. Weller, S. Dooley, M. Goldblum, and A. G. Wilson, “Large Language Models Must Be Taught to Know What They Don’t Know,” Dec. 2024
2024
Later among the works it cites.