Fetching the paper…
Reading the bibliography…
Recent advancements in artificial intelligence have led to the creation of highly capable large language models (LLMs) that can perform tasks in a human-like manner.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 1901
Earlier work this paper cites.
Winogrande: An adversarial winograd schema challenge at scale
Keisuke Sakaguchi, Ronan Le Bras, Chandra Bhagavatula, and Yejin Choi. 2019 · 1907
Earlier work this paper cites.
Delay of gratification in children
Walter Mischel, Yuichi Shoda, and Monica L Rodriguez. 1989 · 1989
Earlier work this paper cites.
Predicting adolescent cognitive and self-regulatory competencies from preschool delay of gratification: Identifying diagnostic conditions
Yuichi Shoda, Walter Mischel, and Philip K Peake. 1990 · 1990
Earlier work this paper cites.
Working memory
Alan Baddeley. 1992 · 1992
Earlier work this paper cites.
Inhibitory control as a contributor to conscience in childhood: From toddler to early school age
Grazyna Kochanska, Kathleen Murray, and Katherine C Coy. 1997 · 1997
Earlier work this paper cites.
Executive function in preschoolers: Links with theory of mind and verbal ability
Claire Hughes. 1998 · 1998
Earlier work this paper cites.
Structural and functional brain development and its relation to cognitive development
BJ Casey, Jay N Giedd, and Kathleen M Thomas. 2000 · 2000
Earlier work this paper cites.
The unity and diversity of executive functions and their contributions to complex “frontal lobe” tasks: A latent variable analysis
Akira Miyake, Naomi P Friedman, Michael J Emerson, Alexander H Witzki, Amy Howerter, and Tor D Wager. 2000 · 2000
Earlier work this paper cites.
Development of executive functions through late childhood and adolescence in an australian sample
Vicki A Anderson, Peter Anderson, Elisabeth Northam, Rani Jacobs, and Cathy Catroppa. 2001 · 2001
Earlier work this paper cites.
Individual differences in inhibitory control and children’s theory of mind
Stephanie M Carlson and Louis J Moses. 2001 · 2001
Earlier work this paper cites.
Assessment and development of executive function (ef) during childhood
Peter Anderson. 2002 · 2002
Earlier work this paper cites.
School readiness: Integrating cognition and emotion in a neurobiological conceptualization of children’s functioning at school entry
Clancy Blair. 2002 · 2002
Earlier work this paper cites.
Conditions under which young children can hold two rules in mind and inhibit a prepotent response
Adele Diamond, Natasha Kirkham, and Dima Amso. 2002 · 2002
Earlier work this paper cites.
How Neural Networks Learn from Experience
Geoffrey E. Hinton. 2002 · 2002
Earlier work this paper cites.
Self-discipline outdoes iq in predicting academic performance of adolescents
Angela L Duckworth and Martin EP Seligman. 2005 · 2005
Earlier work this paper cites.
The development of embodied cognition: Six lessons from babies
Linda Smith and Michael Gasser. 2005 · 2005
Earlier work this paper cites.
Relating effortful control, executive function, and false belief understanding to emerging math and literacy ability in kindergarten
Clancy Blair and Rachel Peters Razza. 2007 · 2007
Earlier work this paper cites.
Executive function in preschoolers: a review using an integrative framework
Nancy Garon, Susan E Bryson, and Isabel M Smith. 2008 · 2008
Earlier work this paper cites.
Fundamentals of human neuropsychology
Bryan Kolb and Ian Q Whishaw. 2009 · 2009
Earlier work this paper cites.
A developmental perspective on executive function
John R Best and Patricia H Miller. 2010 · 2010
Earlier work this paper cites.
A delay-discounting primer
Gregory J Madden and Patrick S Johnson. 2010 · 2010
Earlier work this paper cites.
Learning from experience
Jeffrey Yip and Meena Wilson. 2010 · 2010
Earlier work this paper cites.
Investigating how infants learn to search in the a-not-b task
Hanna Popick, Melody Dye, Natasha Kirkham, and Michael Ramscar. 2011 · 2011
Earlier work this paper cites.
The processes underlying flexibility in childhood
Lucy Cragg and Nicolas Chevalier. 2012 · 2012
Earlier work this paper cites.
A-not-b errors: testing the limits of natural pedagogy theory
Marion Vorms. 2012 · 2012
Earlier work this paper cites.
Executive functions
Adele Diamond. 2013 · 2013
Earlier work this paper cites.
Adolescent cognitive control and reward processing: implications for risk taking and substance use
Charles F Geier. 2013 · 2013
Earlier work this paper cites.
Triviaqa: A large scale distantly supervised challenge dataset for reading comprehension
Mandar Joshi, Eunsol Choi, Daniel S. Weld, and Luke Zettlemoyer. 2017 · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Crowdsourcing multiple choice science questions
Johannes Welbl, Nelson F. Liu, and Matt Gardner. 2017 · 2017
Cited alongside, same era.
Think you have solved question answering? try arc, the ai2 reasoning challenge
Peter Clark, Isaac Cowhey, Oren Etzioni, Tushar Khot, Ashish Sabharwal, Carissa Schoenick, and Oyvind Tafjord. 2018 · 2018
Cited alongside, same era.
Trajectories of infants’ biobehavioral development: Timing and rate of a-not-b performance gains and eeg maturation
Exploring and improving the spatial reasoning abilities of large language models
Manasi Sharma. 2023 · 2023
Later among the works it cites.
Large language models can be lazy learners: Analyze shortcuts in in-context learning
Ruixiang Tang, Dehan Kong, Longtao Huang, and Hui Xue. 2023 · 2023
Later among the works it cites.
Fast yet effective machine unlearning
Ayush K Tarun, Vikram S Chundawat, Murari Mandal, and Mohan Kankanhalli. 2023 · 2023
Later among the works it cites.
Emergent analogical reasoning in large language models
Taylor Webb, Keith J Holyoak, and Hongjing Lu. 2023 · 2023
Later among the works it cites.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc Le, and Denny Zhou. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Leigha MacNeill, Nilam Ram, Martha Ann Bell, Nathan Fox, and Koraly Perez-Edgar. 2018 · 2018
Cited alongside, same era.
MathQA: Towards interpretable math word problem solving with operation-based formalisms
Aida Amini, Saadia Gabriel, Shanchuan Lin, Rik Koncel-Kedziorski, Yejin Choi, and Hannaneh Hajishirzi. 2019 · 2019
Cited alongside, same era.
Neural substrates of early executive function development
Abigail Fiske and Karla Holmboe. 2019 · 2019
Cited alongside, same era.
Natural questions: A benchmark for question answering research
Tom Kwiatkowski, Jennimaria Palomaki, Olivia Redfield, Michael Collins, Ankur Parikh, Chris Alberti, Danielle Epstein, Illia Polosukhin, Jacob Devlin, Kenton Lee, Kristina Toutanova, Llion Jones, Matthew Kelcey, Ming-Wei Chang, Andrew M. Dai, Jakob Uszkoreit, Quoc Le, and Slav Petrov. 2019 · 2019
Cited alongside, same era.
CommonsenseQA: A question answering challenge targeting commonsense knowledge
Alon Talmor, Jonathan Herzig, Nicholas Lourie, and Jonathan Berant. 2019 · 2019
Cited alongside, same era.
Glue: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel R. Bowman. 2019 · 2019
Cited alongside, same era.
Learning to prove theorems via interacting with proof assistants
Kaiyu Yang and Jia Deng. 2019 · 2019
Cited alongside, same era.
Spatialsense: An adversarially crowdsourced benchmark for spatial relation recognition
Kaiyu Yang, Olga Russakovsky, and Jia Deng. 2019 · 2019
Cited alongside, same era.
Leandojo: Theorem proving with retrieval-augmented language models
Kaiyu Yang, Aidan Swope, Alex Gu, Rahul Chalamala, Peiyang Song, Shixing Yu, Saad Godil, Ryan J Prenger, and Animashree Anandkumar. 2023 · 2023
Later among the works it cites.
Large language models are not robust multiple choice selectors
Chujie Zheng, Hao Zhou, Fandong Meng, Jie Zhou, and Minlie Huang. 2023 · 2023
Later among the works it cites.
Llama 3 model card
AI@Meta. 2024 · 2024
Closest in time.
Exploring the psychology of llms’ moral and legal reasoning
Guilherme FCF Almeida, José Luiz Nunes, Neele Engelmann, Alex Wiegmann, and Marcelo de Araújo. 2024 · 2024
Closest in time.
The reversal curse: Llms trained on "a is b" fail to learn "b is a"
Lukas Berglund, Meg Tong, Max Kaufmann, Mikita Balesni, Asa Cooper Stickland, Tomasz Korbak, and Owain Evans. 2024 · 2024
Closest in time.
A survey on evaluation of large language models
Yupeng Chang, Xu Wang, Jindong Wang, Yuan Wu, Linyi Yang, Kaijie Zhu, Hao Chen, Xiaoyuan Yi, Cunxiang Wang, Yidong Wang, Wei Ye, Yue Zhang, Yi Chang, Philip S. Yu, Qiang Yang, and Xing Xie. 2024 · 2024
Closest in time.
Jailbreaking black box large language models in twenty queries
Patrick Chao, Alexander Robey, Edgar Dobriban, Hamed Hassani, George J. Pappas, and Eric Wong. 2024 · 2024
Closest in time.
Evaluating language models for mathematics through interactions
Katherine M. Collins, Albert Q. Jiang, Simon Frieder, Lionel Wong, Miri Zilka, Umang Bhatt, Thomas Lukasiewicz, Yuhuai Wu, Joshua B. Tenenbaum, William Hart, Timothy Gowers, Wenda Li, Adrian Weller, and Mateja Jamnik. 2024 · 2024
Closest in time.
Tao Feng, Chuanyang Jin, Jingyu Liu, Kunlun Zhu, Haoqin Tu, Zirui Cheng, Guanyu Lin, and Jiaxuan You. 2024 · 2024
Closest in time.
Energy efficient convolutions with temporal arithmetic
Rhys Gretsch, Peiyang Song, Advait Madhavan, Jeremy Lau, and Timothy Sherwood. 2024 · 2024
Closest in time.
Decomposing uncertainty for large language models through input clarification ensembling
Bairu Hou, Yujian Liu, Kaizhi Qian, Jacob Andreas, Shiyu Chang, and Yang Zhang. 2024 · 2024
Closest in time.
A peek into token bias: Large language models are not yet genuine reasoners
Bowen Jiang, Yangxinyu Xie, Zhuoqun Hao, Xiaomeng Wang, Tanwi Mallick, Weijie J. Su, Camillo J. Taylor, and Dan Roth. 2024 · 2024
Closest in time.
Machine unlearning in 2024
Ken Ziyu Liu. 2024 · 2024
Closest in time.
Marianna Nezhurina, Lucia Cipolina-Kun, Mehdi Cherti, and Jenia Jitsev. 2024 · 2024
Closest in time.
Towards large language models as copilots for theorem proving in lean
Peiyang Song, Kaiyu Yang, and Anima Anandkumar. 2024 · 2024
Closest in time.
Gemini: A family of highly capable multimodal models
Gemini Team. 2024 · 2024
Closest in time.
Large language models are latent variable models: Explaining and finding good demonstrations for in-context learning
Xinyi Wang, Wanrong Zhu, Michael Saxon, Mark Steyvers, and William Yang Wang. 2024 · 2024
Closest in time.
Livebench: A challenging, contamination-free llm benchmark
Colin White, Samuel Dooley, Manley Roberts, Arka Pal, Ben Feuer, Siddhartha Jain, Ravid Shwartz-Ziv, Neel Jain, Khalid Saifullah, Siddartha Naidu, Chinmay Hegde, Yann LeCun, Tom Goldstein, Willie Neiswanger, and Micah Goldblum. 2024 · 2024
Closest in time.
Weixiang Yan, Haitian Liu, Yunkun Wang, Yunzhe Li, Qian Chen, Wen Wang, Tingyu Lin, Weishan Zhao, Li Zhu, Hari Sundaram, and Shuiguang Deng. 2024 · 2024
Closest in time.
Investigating continual pretraining in large language models: Insights and implications
Çağatay Yıldız, Nishaanth Kanna Ravichandran, Prishruit Punia, Matthias Bethge, and Beyza Ermis. 2024 · 2024
Closest in time.
Quiet-star: Language models can teach themselves to think before speaking
Eric Zelikman, Georges Harik, Yijia Shao, Varuna Jayasiri, Nick Haber, and Noah D. Goodman. 2024 · 2024
Closest in time.
Memorybank: Enhancing large language models with long-term memory
Wanjun Zhong, Lianghong Guo, Qiqi Gao, He Ye, and Yanlin Wang. 2024 · 2024
Closest in time.