Fetching the paper…
Reading the bibliography…
Interaction between caregivers and children plays a critical role in human language acquisition and development.
Roberta: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Asynchronous methods for deep reinforcement learning
Volodymyr Mnih, Adrià Puigdomènech Badia, Mehdi Mirza, Alex Graves, Timothy P. Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu. 2016 · 1937
Earlier work this paper cites.
Child’s talk: Learning to use language
Jerome Bruner. 1985 · 1985
Earlier work this paper cites.
Catastrophic interference in connectionist networks: The sequential learning problem
Michael McCloskey and Neal J Cohen. 1989 · 1989
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J. Williams. 1992 · 1992
Earlier work this paper cites.
Artificial evolution of syntactic aptitude
John Batali. 1994 · 1994
Earlier work this paper cites.
Learning to count without a counter: A case study of dynamics and activation landscapes in recurrent networks
Janet Wiles and Jeff Elman. 1995 · 1995
Earlier work this paper cites.
Active learning with statistical models
David A. Cohn, Zoubin Ghahramani, and Michael I. Jordan. 1996 · 1996
Earlier work this paper cites.
Catastrophic forgetting in connectionist networks
Robert M French. 1999 · 1999
Earlier work this paper cites.
A recurrent neural network that learns to count
Paul Rodriguez, Janet Wiles, and Jeffrey L Elman. 1999 · 1999
Earlier work this paper cites.
Kindertaalverwerving: Een handboek voor het Nederlands
Steven Gillis and Annemarie Schaerlaekens. 2000 · 2000
Earlier work this paper cites.
LSTM recurrent networks learn simple context-free and context-sensitive languages
Felix A. Gers and Jürgen Schmidhuber. 2001 · 2001
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Multi-task active learning for linguistic annotations
Roi Reichart, Katrin Tomanek, Udo Hahn, and Ari Rappoport. 2008 · 2008
Earlier work this paper cites.
Visualizing data using t-sne
Laurens Van der Maaten and Geoffrey Hinton. 2008 · 2008
Earlier work this paper cites.
Curriculum learning
Yoshua Bengio, Jérôme Louradour, Ronan Collobert, and Jason Weston. 2009 · 2009
Earlier work this paper cites.
Language input and acquisition in a mayan village: How important is directed speech?
Laura A Shneidman and Susan Goldin-Meadow. 2012 · 2012
Earlier work this paper cites.
Gumbel-max trick and weighted reservoir sampling
Tim Vieira. 2014 · 2014
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Conversation and language acquisition: A pragmatic approach
Eve V Clark. 2018 · 2018
Cited alongside, same era.
BanditSum: Extractive summarization as a contextual bandit
Yue Dong, Yikang Shen, Eric Crawford, Herke van Hoof, and Jackie Chi Kit Cheung. 2018 · 2018
Cited alongside, same era.
Visualisation and ’diagnostic classifiers’ reveal how recurrent and recursive neural networks process hierarchical structure
Dieuwke Hupkes, Sara Veldhoen, and Willem H. Zuidema. 2018 · 2018
Cited alongside, same era.
Generalization without systematicity: On the compositional skills of sequence-to-sequence recurrent networks
Teacher-student curriculum learning
Tambet Matiisen, Avital Oliver, Taco Cohen, and John Schulman. 2020 · 2020
Later among the works it cites.
Internal and external pressures on language emergence: least effort, object constancy and frequency
Diana Rodríguez Luna, Edoardo Maria Ponti, Dieuwke Hupkes, and Elia Bruni. 2020 · 2020
Later among the works it cites.
The grammar of emergent languages
Oskar van der Wal, Silvan de Boer, Elia Bruni, and Dieuwke Hupkes. 2020 · 2020
Later among the works it cites.
BLiMP: The benchmark of linguistic minimal pairs for English
Alex Warstadt, Alicia Parrish, Haokun Liu, Anhad Mohananey, Wei Peng, Sheng-Fu Wang, and Samuel R. Bowman. 2020 · 2020
Later among the works it cites.
Curriculum learning for natural language understanding
Benfeng Xu, Licheng Zhang, Zhendong Mao, Quan Wang, Hongtao Xie, and Yongdong Zhang. 2020 · 2020
Later among the works it cites.
Can transformers jump around right in natural language? assessing performance transfer from SCAN
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Brenden M. Lake and Marco Baroni. 2018 · 2018
Cited alongside, same era.
Ranking sentences for extractive summarization with reinforcement learning
Shashi Narayan, Shay B. Cohen, and Mirella Lapata. 2018 · 2018
Cited alongside, same era.
Nevergrad - A gradient-free optimization platform
J. Rapin and O. Teytaud. 2018 · 2018
Cited alongside, same era.
Segmentability differences between child-directed and adult-directed speech: A systematic test with an ecologically valid corpus
Alejandrina Cristia, Emmanuel Dupoux, Nan Bernstein Ratner, and Melanie Soderstrom. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Stochastic beams and where to find them: The gumbel-top-k trick for sampling sequences without replacement
Wouter Kool, Herke van Hoof, and Max Welling. 2019 · 2019
Cited alongside, same era.
fairseq: A fast, extensible toolkit for sequence modeling
Myle Ott, Sergey Edunov, Alexei Baevski, Angela Fan, Sam Gross, Nathan Ng, David Grangier, and Michael Auli. 2019 · 2019
Cited alongside, same era.
Rahma Chaabouni, Roberto Dessì, and Eugene Kharitonov. 2021 · 2021
Closest in time.
Co-evolution of language and agents in referential games
Gautier Dagan, Dieuwke Hupkes, and Elia Bruni. 2021 · 2021
Closest in time.
Interactive grounded language understanding in a collaborative environment: Iglu 2021
Julia Kiseleva, Ziming Li, Mohammad Aliannejadi, Shrestha Mohanty, Maartje ter Hoeve, Mikhail Burtsev, Alexey Skrynnik, Artem Zholus, Aleksandr Panov, Kavya Srinet, et al. 2022a · 2021
Closest in time.
Pitfalls of static language modelling
Angeliki Lazaridou, Adhiguna Kuncoro, Elena Gribovskaya, Devang Agrawal, Adam Liska, Tayfun Terzi, Mai Gimenez, Cyprien de Masson d’Autume, Sebastian Ruder, Dani Yogatama, Kris Cao, Tomás Kociský, Susannah Young, and Phil Blunsom. 2021 · 2021
Closest in time.
Modeling the interaction between perception-based and production-based learning in children’s early acquisition of semantic knowledge
Mitja Nikolaus and Abdellah Fourtassi. 2021 · 2021
Closest in time.
SHAPELURN: An interactive language learning game with logical inference
Katharina Stein, Leonie Harter, and Luisa Geiger. 2021 · 2021
Closest in time.
Reducing bert computation by padding removal and curriculum learning
Wei Zhang, Wei Wei, Wen Wang, Lingling Jin, and Zheng Cao. 2021 · 2021
Closest in time.
Lifelong pretraining: Continually adapting language models to emerging corpora
Xisen Jin, Dejiao Zhang, Henghui Zhu, Wei Xiao, Shang-Wen Li, Xiaokai Wei, Andrew Arnold, and Xiang Ren. 2022 · 2022
Closest in time.
Julia Kiseleva, Alexey Skrynnik, Artem Zholus, Shrestha Mohanty, Negar Arabzadeh, Marc-Alexandre Côté, Mohammad Aliannejadi, Milagro Teruel, Ziming Li, Mikhail Burtsev, et al. 2022b · 2022
Closest in time.
Beyond the imitation game: Quantifying and extrapolating the capabilities of language models
Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao, Abu Awal Md Shoeb, Abubakar Abid, Adam Fisch, Adam R. Brown, Adam Santoro, Aditya Gupta, Adrià Garriga-Alonso, Agnieszka Kluska, Aitor Lewkowycz, Akshat Agarwal, Alethea Power, Alex Ray, Alex Warstadt, Alexander W. Kocurek, Ali Safaya, Ali Tazarv, Alice Xiang, Alicia Parrish, Allen Nie, Aman Hussain, Amanda Askell, Amanda Dsouza, Ameet Rahane, Anantharaman S. Iyer, Anders Andreassen, Andrea Santilli, Andreas Stuhlmüller, Andrew M. Dai, Andrew La, Andrew K. Lampinen, Andy Zou, Angela Jiang, Angelica Chen, Anh Vuong, Animesh Gupta, Anna Gottardi, Antonio Norelli, Anu Venkatesh, Arash Gholamidavoodi, Arfa Tabassum, Arul Menezes, Arun Kirubarajan, Asher Mullokandov, Ashish Sabharwal, Austin Herrick, Avia Efrat, Aykut Erdem, Ayla Karakas, and et al. 2022 · 2022
Closest in time.
Simple recurrent networks learn context-free and context-sensitive languages by counting
Paul Rodriguez. 2001 · 2093
Closest in time.