Fetching the paper…
Reading the bibliography…
Developing models to automatically score students' written responses to science problems is critical for science education.
Salton, G., Buckley, C.: Term-weighting approaches in automatic text retrieval. Information processing & management (1988)
1988
Earlier work this paper cites.
Bejar, I.I.: A methodology for scoring open-ended architectural design problems. Journal of Applied Psychology (1991)
1991
Earlier work this paper cites.
Council, N.R., et al.: A framework for K-12 science education: Practices, crosscutting concepts, and core ideas. National Academies Press (2012)
2012
Earlier work this paper cites.
Haudek, K.C., et al.: What are they thinking? automated analysis of student writing about acid–base chemistry in introductory biology. Life Sciences Education (2012)
2012
Earlier work this paper cites.
Nehm, R.H., Ha, M., Mayfield, E.: Transforming biology assessment with machine learning: automated scoring of written evolutionary explanations. Journal of Science Education and Technology (2012)
2012
Earlier work this paper cites.
Pellegrino, J.W.: Proficiency in science: Assessment challenges and opportunities. Science (2013)
2013
Earlier work this paper cites.
Liu, O.L., et al.: Automated scoring of constructed-response science items: Prospects and obstacles. Educational Measurement: Issues and Practice (2014)
2014
Earlier work this paper cites.
2015
Earlier work this paper cites.
Litman, D.: Natural language processing for enhancing teaching and learning. In: Thirtieth AAAI conference on artificial intelligence (2016)
2016
Earlier work this paper cites.
Osborne, J.F., et al.: The development and validation of a learning progression for argumentation in science. Journal of research in science teaching (2016)
2016
Earlier work this paper cites.
Devlin, J., Chang, M.W., Lee, K., Toutanova, K.: Bert: Pre-training of deep bidirectional transformers for language understanding. In: ACL (2019)
2019
Earlier work this paper cites.
Gerard, L., Kidron, A., Linn, M.C.: Guiding collaborative revision of science explanations. Int. Journal of Computer-Supported Collaborative Learning (2019)
2019
Earlier work this paper cites.
Harris, C.J., et al.: Designing knowledge-in-use assessments to promote deeper learning. Educational measurement: issues and practice (2019)
2019
Earlier work this paper cites.
Lee, H.S., et al.: Automated text scoring and real-time adjustable feedback: Supporting revision of scientific arguments involving uncertainty. Science Education (2019)
2019
Earlier work this paper cites.
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J.D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.: Language models are few-shot learners. Advances in neural information processing systems (2020)
2020
Cited alongside, same era.
2020
Cited alongside, same era.
Riordan, B.e.a.: An empirical investigation of neural methods for content scoring of science explanations. In: Proceedings of the fifteenth workshop on innovative use of NLP for building educational applications (2020)
2020
Cited alongside, same era.
2020
Cited alongside, same era.
Maestrales, S.e.a.: Using machine learning to score multi-dimensional assessments of chemistry and physics. Journal of Science Education and Technology (2021)
2021
Later among the works it cites.
Omizo, R., Meeks, M., Hart-Davidson, W.: Detecting high-quality comments in written feedback with a zero shot classifier. In: ACM ICDC (2021)
2021
Later among the works it cites.
Uhl, J.D., et al.: Introductory biology undergraduate students’ mixed ideas about genetic information flow. Biochemistry and Molecular Biology Education (2021)
2021
Later among the works it cites.
2021
Later among the works it cites.
Zhai, X.: Practices and theories: How can machine learning assist in innovative assessment practices in science education. Journal of Science Education and Technology (2021)
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Schick, T., Schütze, H.: Exploiting cloze questions for few-shot text classification and natural language inference. Computing Research Repository (2020)
2020
Cited alongside, same era.
2020
Cited alongside, same era.
Wolfe, E.W., Wendler, C.L.W.: Why should we care about human raters? Applied Measurement in Education (2020)
2020
Cited alongside, same era.
Zhai, X., Yin, Y., Pellegrino, J.W., Haudek, K.C., Shi, L.: Applying machine learning in science assessment: a systematic review. Studies in Science Education (2020)
2020
Cited alongside, same era.
Haudek, K.C., Zhai, X.: Exploring the effect of assessment construct complexity on machine learning scoring of argumentation (2021)
2021
Cited alongside, same era.
2021
Cited alongside, same era.
Liu, X., et al.: Gpt understands, too. arXiv preprint arXiv:2103.10385 (2021)
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Later among the works it cites.
Zhai, X., Krajcik, J., Pellegrino, J.W.: On the validity of machine learning-based next generation science assessments: A validity inferential network. Journal of Science Education and Technology (2021)
2021
Later among the works it cites.
Zhai, X., Shi, L., Nehm, R.H.: A meta-analysis of machine learning-based science assessments: factors impacting machine-human score agreements. Journal of Science Education and Technology (2021)
2021
Later among the works it cites.
2021
Later among the works it cites.
Mayer, C.W., Ludwig, S., Brandt, S.: Prompt text classifications with transformer models! an exemplary introduction to prompt-based learning with large language models. Journal of Research on Technology in Education (2022)
2022
Later among the works it cites.
Su, Y., et al.: On transferability of prompt tuning for natural language processing. In: NACL. pp. 3949–3969 (2022)
2022
Later among the works it cites.
Zhai, X., Haudek, K.C., Ma, W.: Assessing argumentation using machine learning and cognitive diagnostic modeling. Research in Science Education (2022)
2022
Later among the works it cites.
2022
Later among the works it cites.