A. Søgaard, Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing
2021
Later among the works it cites.
Advances in Neural Information Processing Systems
M. Nye, M. Tessler, J. Tenenbaum, B. M. Lake, Improving coherence and consistency in neural sequence models with dual-system, neuro-symbolic reasoning · 2021
Later among the works it cites.
arXiv preprint arXiv:2104.01490
Original
R. Cao, D. Yamins, Explanatory models in neuroscience: Part 1–taking mechanistic abstraction seriously · 2021
Later among the works it cites.
arXiv preprint arXiv:2104.01489
Original
R. Cao, D. Yamins, Explanatory models in neuroscience: Part 2–constraint-based intelligibility · 2021
Later among the works it cites.
Proceedings of the National Academy of Sciences
M. Schrimpf, I. A. Blank, G. Tuckute, C. Kauf, E. A. Hosseini, N. Kanwisher, J. B. Tenenbaum, E. Fedorenko, The neural architecture of language: Integrative modeling converges on predictive processing · 2021
Later among the works it cites.
arXiv preprint arXiv:2109.01247
Original
A. Webson, E. Pavlick, Do prompt-based models really understand the meaning of their prompts? · 2021
Later among the works it cites.
arXiv preprint arXiv:2104.08773
Original
S. Mishra, D. Khashabi, C. Baral, H. Hajishirzi, Cross-task generalization via natural language crowdsourcing instructions · 2021
Later among the works it cites.
arXiv preprint arXiv:2107.06994
Original
A. J. Nam, J. L. McClelland, What underlies rapid learning and systematic generalization in humans · 2021
Later among the works it cites.
Y. Wu, M. N. Rabe, W. Li, J. Ba, R. B. Grosse, C. Szegedy, International Conference on Machine Learning
2021
Later among the works it cites.
Transactions of the Association for Computational Linguistics
T. Schick, S. Udupa, H. Schütze, Self-Diagnosis and Self-Debiasing: A Proposal for Reducing Corpus-Based Bias in NLP · 2021
Later among the works it cites.
arXiv preprint arXiv:2110.14168
Original
K. Cobbe, V. Kosaraju, M. Bavarian, J. Hilton, R. Nakano, C. Hesse, J. Schulman, Training verifiers to solve math word problems · 2021
Later among the works it cites.
Current Opinion in Behavioral Sciences
J. X. Wang, Meta-learning in natural and artificial intelligence · 2021
Later among the works it cites.
J. Dodge, M. Sap, A. Marasović, W. Agnew, G. Ilharco, D. Groeneveld, M. Mitchell, M. Gardner, Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing
2021
Later among the works it cites.
arXiv preprint arXiv:2104.08315
Original
A. Holtzman, P. West, V. Shwartz, Y. Choi, L. Zettlemoyer, Surface form competition: Why the highest probability answer isn’t always right · 2021
Later among the works it cites.
A. Hosseini, S. Reddy, D. Bahdanau, R. D. Hjelm, A. Sordoni, A. Courville, Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies
2021
Later among the works it cites.
arXiv preprint arXiv:2202.07785
Original
D. Ganguli, D. Hernandez, L. Lovitt, N. DasSarma, T. Henighan, A. Jones, N. Joseph, J. Kernion, B. Mann, A. Askell, et al · 2022
Closest in time.
arXiv preprint arXiv:2205.11916
Original
T. Kojima, S. S. Gu, M. Reid, Y. Matsuo, Y. Iwasawa, Large language models are zero-shot reasoners · 2022
Closest in time.
arXiv preprint arXiv:2206.07682
Original
J. Wei, Y. Tay, R. Bommasani, C. Raffel, B. Zoph, S. Borgeaud, D. Yogatama, M. Bosma, D. Zhou, D. Metzler, et al · 2022
Closest in time.
arXiv preprint arXiv:2202.07206
Original
Y. Razeghi, R. L. Logan IV, M. Gardner, S. Singh, Impact of pretraining term frequencies on few-shot reasoning · 2022
Closest in time.
K. Valmeekam, A. Olmo, S. Sreedharan, S. Kambhampati, Large language models still can’t plan (a benchmark for llms on planning and reasoning about change) (2022)
2022
Closest in time.
arXiv preprint arXiv:2203.15556
Original
J. Hoffmann, S. Borgeaud, A. Mensch, E. Buchatskaya, T. Cai, E. Rutherford, D. d. L. Casas, L. A. Hendricks, J. Welbl, A. Clark, et al · 2022
Closest in time.
Topics in Cognitive Science
M. H. Tessler, J. B. Tenenbaum, N. D. Goodman, Logic, probability, and pragmatics in syllogistic reasoning · 2022
Closest in time.
S. Kadavath, T. Conerly, A. Askell, T. Henighan, D. Drain, E. Perez, N. Schiefer, Z. H. Dodds, N. DasSarma, E. Tran-Johnson, S. Johnston, S. El-Showk, A. Jones, N. Elhage, T. Hume, A. Chen, Y. Bai, S. Bowman, S. Fort, D. Ganguli, D. Hernandez, J. Jacobson, J. Kernion, S. Kravec, L. Lovitt, K. Ndousse, C. Olsson, S. Ringer, D. Amodei, T. Brown, J. Clark, N. Joseph, B. Mann, S. McCandlish, C. Olah, J. Kaplan, Language models (mostly) know what they know (2022)
2022
Closest in time.
M. Binz, E. Schulz, Using cognitive psychology to understand gpt-3 (2022)
2022
Closest in time.
arXiv preprint arXiv:2201.11903
Original
J. Wei, X. Wang, D. Schuurmans, M. Bosma, E. Chi, Q. Le, D. Zhou, Chain of thought prompting elicits reasoning in large language models · 2022
Closest in time.
D. Khashabi, C. Baral, Y. Choi, H. Hajishirzi, Findings of the Association for Computational Linguistics: ACL 2022
2022
Closest in time.
arXiv preprint arXiv:2204.02329
Original
A. K. Lampinen, I. Dasgupta, S. C. Chan, K. Matthewson, M. H. Tessler, A. Creswell, J. L. McClelland, J. X. Wang, F. Hill, Can language models learn from explanations in context? · 2022
Closest in time.
arXiv preprint arXiv:2203.11171
Original
X. Wang, J. Wei, D. Schuurmans, Q. Le, E. Chi, D. Zhou, Self-consistency improves chain of thought reasoning in language models · 2022
Closest in time.
PLOS Computational Biology
Y. Li, J. L. McClelland, A weighted constraint satisfaction approach to human goal-directed decision making · 2022
Closest in time.
Nature neuroscience
A. Goldstein, Z. Zada, E. Buchnik, M. Schain, A. Price, B. Aubrey, S. A. Nastase, A. Feder, D. Emanuel, A. Cohen, et al · 2022
Closest in time.
arXiv preprint arXiv:2206.02885
Original
D. Schlangen, Norm participation grounds language · 2022
Closest in time.
arXiv preprint arXiv:2206.05802
Original
W. Saunders, C. Yeh, J. Wu, S. Bills, L. Ouyang, J. Ward, J. Leike, Self-critiquing models for assisting human evaluators · 2022
Closest in time.
arXiv preprint arXiv:2203.14465
Original
E. Zelikman, Y. Wu, N. D. Goodman, Star: Bootstrapping reasoning with reasoning · 2022
Closest in time.
Advances in Neural Information Processing Systems
S. Chan, A. Santoro, A. Lampinen, J. Wang, A. Singh, P. Richemond, J. McClelland, F. Hill, Data distributional properties drive emergent in-context learning in transformers · 2022
Closest in time.
Y. Tay, M. Dehghani, V. Q. Tran, X. Garcia, J. Wei, X. Wang, H. W. Chung, D. Bahri, T. Schuster, S. Zheng, et al
2022
Closest in time.
arXiv preprint arXiv:2304.15004
Original
R. Schaeffer, B. Miranda, S. Koyejo, Are emergent abilities of large language models a mirage? · 2023
Closest in time.
Gpt-3.5, https://platform.openai.com/docs/models/gpt-3-5
2023
Closest in time.
arXiv preprint arXiv:2305.10403
Original
R. Anil, A. M. Dai, O. Firat, M. Johnson, D. Lepikhin, A. Passos, S. Shakeri, E. Taropa, P. Bailey, Z. Chen, et al · 2023
Closest in time.
arXiv preprint arXiv:2304.03843
Original
B. Prystawski, N. D. Goodman, Why think step-by-step? reasoning emerges from the locality of experience · 2023
Closest in time.