Fetching the paper…
Reading the bibliography…
Humans are efficient language learners and inherently social creatures.
Fine-tuning language models from human preferences
Daniel M Ziegler, Nisan Stiennon, Jeffrey Wu, Tom B Brown, Alec Radford, Dario Amodei, Paul Christiano, and Geoffrey Irving. 2019 · 1909
Earlier work this paper cites.
Animal intelligence: Experimental studies
Edward Thorndike. 1911 · 1911
Earlier work this paper cites.
Thought and language
Lev S Vygotsky. 1934 · 1934
Earlier work this paper cites.
Derivational complexity and order of acquisition
Roger Brown. 1970 · 1970
Earlier work this paper cites.
The acquisition of performatives prior to speech
Elizabeth Bates, Luigia Camaioni, and Virginia Volterra. 1975 · 1975
Earlier work this paper cites.
Learning how to mean
Michael Alexander Kirkwood Halliday. 1975 · 1975
Earlier work this paper cites.
Child’s talk: Learning to use language
Jerome Bruner. 1985 · 1985
Earlier work this paper cites.
The role of dialogue in providing scaffolded instruction
Annemarie Sullivan Palincsar. 1986 · 1986
Earlier work this paper cites.
The structural sources of verb meanings
Lila Gleitman. 1990 · 1990
Earlier work this paper cites.
Negative evidence and grammatical morphemes
MJ Farrar. 1992 · 1992
Earlier work this paper cites.
Negative evidence in language acquisition
Gary F Marcus. 1993 · 1993
Earlier work this paper cites.
Using language
Herbert H Clark. 1996 · 1996
Earlier work this paper cites.
Learning how to say what one means: A longitudinal study of children’s speech act use
Catherine E Snow, Barbara Alexander Pan, Alison Imbens-Bailey, and Jane Herman. 1996 · 1996
Earlier work this paper cites.
The CHILDES project: The database , volume 2
Brian MacWhinney. 2000 · 2000
Earlier work this paper cites.
Negative evidence and negative feedback: Immediate effects on the grammaticality of child speech
Matthew Saxton. 2000 · 2000
Earlier work this paper cites.
Corrective feedback in second language acquisition
Mounira El Tatawy. 2002 · 2002
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Adult reformulations of child errors as negative evidence
Michelle M Chouinard and Eve V Clark. 2003 · 2003
Earlier work this paper cites.
Early language acquisition: cracking the speech code
Patricia K Kuhl. 2004 · 2004
Earlier work this paper cites.
Negative input for grammatical errors: Effects after a lag of 12 weeks
Matthew Saxton, Phillip Backley, and Clare Gallaway. 2005 · 2005
Earlier work this paper cites.
Constructing a language: A usage-based theory of language acquisition
Michael Tomasello. 2005 · 2005
Earlier work this paper cites.
Implicit and explicit corrective feedback and the acquisition of l2 grammar
Rod Ellis, Shawn Loewen, and Rosemary Erlam. 2006 · 2006
Earlier work this paper cites.
Macarthur-bates communicative development inventories
Larry Fenson, Virginia A Marchman, Donna J Thal, Phillip S Dale, J Steven Reznick, and Elizabeth Bates. 2006 · 2006
Earlier work this paper cites.
Gaze following in human infants depends on communicative signals
Atsushi Senju and Gergely Csibra. 2008 · 2008
Earlier work this paper cites.
Infants rapidly learn word-referent mappings via cross-situational statistics
Linda Smith and Chen Yu. 2008 · 2008
Earlier work this paper cites.
Context-based word acquisition for situated dialogue in a virtual world
Shaolin Qu and Joyce Yue Chai. 2010 · 2010
Earlier work this paper cites.
Written corrective feedback in second language acquisition and writing
John Bitchener and Dana R Ferris. 2011 · 2011
Earlier work this paper cites.
A social feedback loop for speech development and its reduction in autism
Anne S Warlaumont, Jeffrey A Richards, Jill Gilkerson, and D Kimbrough Oller. 2014 · 2014
Earlier work this paper cites.
Language understanding for text-based games using deep reinforcement learning
Karthik Narasimhan, Tejas D Kulkarni, and Regina Barzilay. 2015 · 2015
Cited alongside, same era.
Aligning books and movies: Towards story-like visual explanations by watching movies and reading books
Yukun Zhu, Ryan Kiros, Rich Zemel, Ruslan Salakhutdinov, Raquel Urtasun, Antonio Torralba, and Sanja Fidler. 2015 · 2015
Cited alongside, same era.
From uh-oh to tomorrow: Predicting age of acquisition for early words across languages
Mika Braginsky, Daniel Yurovsky, Virginia A. Marchman, and Mike Frank. 2016 · 2016
Cited alongside, same era.
Deep reinforcement learning with a natural language action space
Ji He, Jianshu Chen, Xiaodong He, Jianfeng Gao, Lihong Li, Li Deng, and Mari Ostendorf. 2016 · 2016
Cited alongside, same era.
Corrective feedback in first language acquisition
Sarah Hiller. 2016 · 2016
Cited alongside, same era.
Scala: A blueprint for computational models of language acquisition in social context
Sho Tsuji, Alejandrina Cristia, and Emmanuel Dupoux. 2021 · 2021
Later among the works it cites.
Constitutional ai: Harmlessness from ai feedback
Yuntao Bai, Saurav Kadavath, Sandipan Kundu, Amanda Askell, Jackson Kernion, Andy Jones, Anna Chen, Anna Goldie, Azalia Mirhoseini, Cameron McKinnon, et al. 2022 · 2022
Later among the works it cites.
On the proper role of linguistically-oriented deep net analysis in linguistic theorizing
Marco Baroni. 2022 · 2022
Later among the works it cites.
Analyzing the mono-and cross-lingual pretraining dynamics of multilingual language models
Terra Blevins, Hila Gonen, and Luke Zettlemoyer. 2022 · 2022
Later among the works it cites.
Word acquisition in neural language models
Tyler A Chang and Benjamin K Bergen. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Angeliki Lazaridou, Nghia The Pham, and Marco Baroni. 2016 · 2016
Cited alongside, same era.
Sequence level training with recurrent neural networks
Marc’Aurelio Ranzato, Sumit Chopra, Michael Auli, and Wojciech Zaremba. 2016 · 2016
Cited alongside, same era.
High-dimensional continuous control using generalized advantage estimation
John Schulman, Philipp Moritz, Sergey Levine, Michael Jordan, and Pieter Abbeel. 2016 · 2016
Cited alongside, same era.
Neural machine translation of rare words with subword units
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Cited alongside, same era.
Dialog-based language learning
Jason E Weston. 2016 · 2016
Cited alongside, same era.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov. 2017 · 2017
Cited alongside, same era.
Beyond naïve cue combination: Salience and social cues in early word learning
Daniel Yurovsky and Michael C Frank. 2017 · 2017
Cited alongside, same era.
Computational language acquisition with theory of mind
Andy Liu, Hao Zhu, Emmy Liu, Yonatan Bisk, and Graham Neubig. 2022 · 2022
Later among the works it cites.
Chatgpt: Optimizing language models for dialogue
OpenAI. 2022 · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al. 2022 · 2022
Later among the works it cites.
Neural Network Approaches to the Study of Word Learning
Eva Portelance. 2022 · 2022
Later among the works it cites.
Towards interactive language modeling
Maartje ter Hoeve, Evgeny Kharitonov, Dieuwke Hupkes, and Emmanuel Dupoux. 2022 · 2022
Later among the works it cites.
What artificial neural networks can tell us about human language acquisition
Alex Warstadt and Samuel R Bowman. 2022 · 2022
Later among the works it cites.
Language learning from communicative goals and linguistic input
Hao Zhu, Yonatan Bisk, and Graham Neubig. 2022 · 2022
Later among the works it cites.
Pythia: A suite for analyzing large language models across training and scaling
Stella Biderman, Hailey Schoelkopf, Quentin Gregory Anthony, Herbie Bradley, Kyle O’Brien, Eric Hallahan, Mohammad Aflah Khan, Shivanshu Purohit, USVSN Sai Prashanth, Edward Raff, et al. 2023 · 2023
Later among the works it cites.
Tyler A Chang, Zhuowen Tu, and Benjamin K Bergen. 2023 · 2023
Later among the works it cites.
Language acquisition: do children and language models follow similar learning stages?
Linnea Evanson, Yair Lakretz, and Jean-Rémi King. 2023 · 2023
Later among the works it cites.
A survey of reinforcement learning from human feedback
Timo Kaufmann, Paul Weng, Viktor Bengs, and Eyke Hüllermeier. 2023 · 2023
Later among the works it cites.
Rlaif: Scaling reinforcement learning from human feedback with ai feedback
Harrison Lee, Samrat Phatale, Hassan Mansoor, Kellie Lu, Thomas Mesnard, Colton Bishop, Victor Carbune, and Abhinav Rastogi. 2023 · 2023
Later among the works it cites.
World-to-words: Grounded open vocabulary acquisition through fast mapping in vision-language models
Ziqiao Ma, Jiayi Pan, and Joyce Chai. 2023 · 2023
Later among the works it cites.
Communicative feedback in language acquisition
Mitja Nikolaus and Abdellah Fourtassi. 2023 · 2023
Later among the works it cites.
The roles of neural networks in language acquisition
Eva Portelance and Masoud Jasbi. 2023 · 2023
Later among the works it cites.
Can language models teach? teacher explanations improve student performance via personalization
Swarnadeep Saha, Peter Hase, and Mohit Bansal. 2023 · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al. 2023 · 2023
Later among the works it cites.
Findings of the babylm challenge: Sample-efficient pretraining on developmentally plausible corpora
Alex Warstadt, Aaron Mueller, Leshem Choshen, Ethan Wilcox, Chengxu Zhuang, Juan Ciro, Rafael Mosquera, Bhargavi Paranjabe, Adina Williams, Tal Linzen, et al. 2023 · 2023
Later among the works it cites.
Training trajectories of language models across scales
Mengzhou Xia, Mikel Artetxe, Chunting Zhou, Xi Victoria Lin, Ramakanth Pasunuru, Danqi Chen, Luke Zettlemoyer, and Ves Stoyanov. 2023 · 2023
Later among the works it cites.
Secrets of rlhf in large language models part i: Ppo
Rui Zheng, Shihan Dou, Songyang Gao, Yuan Hua, Wei Shen, Binghai Wang, Yan Liu, Senjie Jin, Qin Liu, Yuhao Zhou, et al. 2023 · 2023
Later among the works it cites.
DistiLLM: Towards streamlined distillation for large language models
Jongwoo Ko, Sungnyun Kim, Tianyi Chen, and Se-Young Yun. 2024 · 2024
Closest in time.
Learning the meanings of function words from grounded language using a visual question answering model
Eva Portelance, Michael C Frank, and Dan Jurafsky. 2024 · 2024
Closest in time.