Fetching the paper…
Reading the bibliography…
Multiple-choice exam questions with "None of the above" (NA) options have been extensively studied in educational testing, in which existing research suggests that they better assess true knowledge.
The theory of the estimation of test reliability
G. F. Kuder and M. W. Richardson · 1937
Earlier work this paper cites.
An item-level analysis of" none of the above."
C. E. Rich and G. A. Johanson · 1990
Earlier work this paper cites.
The none-of-the-above option: An empirical study
R. B. Frary · 1991
Earlier work this paper cites.
In defence of ‘none of the above’
M. A. García-Pŕrezt · 1993
Earlier work this paper cites.
What do we know about eyewitness identification?
G. L. Wells · 1993
Earlier work this paper cites.
Logical versus empirical guidelines for writing test items: The case of" none of the above"
L. J. Gross · 1994
Earlier work this paper cites.
A review of multiple-choice item-writing guidelines for classroom assessment
T. M. Haladyna, S. M. Downing, and M. C. Rodriguez · 2002
Earlier work this paper cites.
Best practices for designing and grading exams
M. E. Piontek · 2008
Earlier work this paper cites.
The “none of the above” option in multiple-choice testing: An experimental study
D. DiBattista, J.-A. Sinnige-Egger, and G. Fortuna · 2014
Earlier work this paper cites.
How “none of the above”(nota) affects the accessibility of tested and related information in multiple-choice questions
M. F. Blendermann, J. L. Little, and K. M. Gray · 2020
Earlier work this paper cites.
Measuring massive multitask language understanding
D. Hendrycks, C. Burns, S. Basart, A. Zou, M. Mazeika, D. Song, and J. Steinhardt · 2020
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
J. E. Hu, Y. Shen, P. Wallis, Z. Allen-Zhu, Y. Li, S. Wang, and W. Chen · 2021
Earlier work this paper cites.
A. H. Jonsdottir, T. Jonmundsson, I. H. Armann, B. B. Gunnarsdottir, and G. Stefansson · 2021
Cited alongside, same era.
Language models (mostly) know what they know
S. Kadavath, T. Conerly, A. Askell, T. Henighan, D. Drain, E. Perez, N. Schiefer, Z. Hatfield-Dodds, N. DasSarma, E. Tran-Johnson, et al · 2022
Cited alongside, same era.
Large language models are zero-shot reasoners
T. Kojima, S. S. Gu, M. Reid, Y. Matsuo, and Y. Iwasawa · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
L. Ouyang, J. Wu, X. Jiang, D. Almeida, C. Wainwright, P. Mishkin, C. Zhang, S. Agarwal, K. Slama, A. Ray, et al · 2022
Cited alongside, same era.
A. Q. Jiang, A. Sablayrolles, A. Mensch, C. Bamford, D. S. Chaplot, D. d. l. Casas, F. Bressand, G. Lengyel, G. Lample, L. Saulnier, et al · 2023
axolotl, 2024
axolotl-ai-cloud · 2024
Later among the works it cites.
A. Dubey, A. Jauhri, A. Pandey, A. Kadian, A. Al-Dahle, A. Letman, A. Mathur, A. Schelten, A. Yang, A. Fan, et al · 2024
Later among the works it cites.
A. Q. Jiang, A. Sablayrolles, A. Roux, A. Mensch, B. Savary, C. Bamford, D. S. Chaplot, D. d. l. Casas, E. B. Hanna, F. Bressand, et al · 2024
Later among the works it cites.
Large language models must be taught to know what they don’t know
S. Kapoor, N. Gruver, M. Roberts, K. Collins, A. Pal, U. Bhatt, A. Weller, S. Dooley, M. Goldblum, and A. G. Wilson · 2024
Later among the works it cites.
A. Liu, B. Feng, B. Xue, B. Wang, B. Wu, C. Lu, C. Zhao, C. Deng, C. Zhang, C. Ruan, et al · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Solar 10.7 b: Scaling large language models with simple yet effective depth up-scaling
D. Kim, C. Park, S. Kim, W. Lee, W. Song, Y. Kim, H. Kim, Y. Kim, H. Lee, J. Kim, et al · 2023
Cited alongside, same era.
Efficient memory management for large language model serving with pagedattention
W. Kwon, Z. Li, S. Zhuang, Y. Sheng, L. Zheng, C. H. Yu, J. E. Gonzalez, H. Zhang, and I. Stoica · 2023
Cited alongside, same era.
Gpqa: A graduate-level google-proof q&a benchmark
D. Rein, B. L. Hou, A. C. Stickland, J. Petty, R. Y. Pang, J. Dirani, J. Michael, and S. R. Bowman · 2023
Cited alongside, same era.
Gemini: a family of highly capable multimodal models
G. Team, R. Anil, S. Borgeaud, J.-B. Alayrac, J. Yu, R. Soricut, J. Schalkwyk, A. M. Dai, A. Hauth, K. Millican, et al · 2023
Cited alongside, same era.
Large language models are not robust multiple choice selectors
C. Zheng, H. Zhou, F. Meng, J. Zhou, and M. Huang · 2023
Cited alongside, same era.
On the calibration of large language models and alignment
C. Zhu, B. Xu, Q. Wang, Y. Zhang, and Z. Mao · 2023
Cited alongside, same era.
When benchmarks are targets: Revealing the sensitivity of large language model leaderboards
N. Alzahrani, H. A. Alyahya, Y. Alnumay, S. Alrashed, S. Alsubaie, Y. Almushaykeh, F. Mirza, N. Alotaibi, N. Altwairesh, A. Alowisheq, et al · 2024
Cited alongside, same era.
Later among the works it cites.
Direct preference optimization: Your language model is secretly a reward model
R. Rafailov, A. Sharma, E. Mitchell, C. D. Manning, S. Ermon, and C. Finn · 2024
Later among the works it cites.
Mmlu-pro: A more robust and challenging multi-task language understanding benchmark
Y. Wang, X. Ma, G. Zhang, Y. Ni, A. Chandra, S. Guo, W. Ren, A. Arulraj, X. He, Z. Jiang, T. Li, M. W. Ku, K. Wang, A. Zhuang, R. R. Fan, X. Yue, and W. Chen · 2024
Later among the works it cites.
Unveiling selection biases: Exploring order and token sensitivity in large language models
S.-L. Wei, C.-K. Wu, H.-H. Huang, and H.-H. Chen · 2024
Later among the works it cites.
Calibrating reasoning in language models with internal consistency
Z. Xie, J. Guo, T. Yu, and S. Li · 2024
Later among the works it cites.
A. Yang, B. Yang, B. Zhang, B. Hui, B. Zheng, B. Yu, C. Li, D. Liu, F. Huang, H. Wei, et al · 2024
Later among the works it cites.
Yi: Open foundation models by 01. ai
A. Young, B. Chen, C. Li, C. Huang, G. Zhang, G. Zhang, G. Wang, H. Li, J. Zhu, J. Chen, et al · 2024
Later among the works it cites.
R-tuning: Instructing large language models to say ‘i don’t know’
H. Zhang, S. Diao, Y. Lin, Y. Fung, Q. Lian, X. Wang, Y. Chen, H. Ji, and T. Zhang · 2024
Later among the works it cites.