Fetching the paper…
Reading the bibliography…
Human linguistic capacity is often characterized by compositionality and the generalization it enables -- human learners can produce and comprehend novel complex expressions by composing known parts.
Compound thoughts
Gottlob Frege. 1923 · 1923
Earlier work this paper cites.
The child’s learning of English morphology
Jean Berko. 1958 · 1958
Earlier work this paper cites.
Physical symbol systems
Allen Newell. 1980 · 1980
Earlier work this paper cites.
Computation and cognition: Issues in the foundations of cognitive science
Zenon W. Pylyshyn. 1980 · 1980
Earlier work this paper cites.
Connectionism and cognitive architecture: A critical analysis
Jerry A. Fodor and Zenon W. Pylyshyn. 1988 · 1988
Earlier work this paper cites.
The constituent structure of connectionist mental states: A reply to Fodor and Pylyshyn
Paul Smolensky. 1991 · 1991
Earlier work this paper cites.
The connectionism/classicism battle to win souls
Brian P. McLaughlin. 1993 · 1993
Earlier work this paper cites.
Twenty-five-month-old children do not have a grammatical category of verb
Raquel Olguin and Michael Tomasello. 1993 · 1993
Earlier work this paper cites.
Systematicity in connectionist language learning
Robert F. Hadley. 1994 · 1994
Earlier work this paper cites.
Are feedforward and recurrent networks systematic? Analysis and implications for a connectionist cognitive architecture
Steven Phillips. 1998 · 1998
Earlier work this paper cites.
The algebraic mind: Integrating connectionism and cognitive science
Gary Marcus. 2001 · 2001
Earlier work this paper cites.
Lack of combinatorial productivity in language processing with simple recurrent networks
Frank van der Velde, Gwendid T. van der Voort van der Kleij, and Marc de Kamps. 2004 · 2004
Earlier work this paper cites.
Compositional generalization in semantic parsing: Pre-training vs. specialized architectures
Daniel Furrer, Marc van Zee, Nathan Scales, and Nathanael Schärli. 2020 · 2007
Earlier work this paper cites.
Syntactic recursion and iteration
Fred Karlsson. 2010 · 2010
Earlier work this paper cites.
Syntactic generalization with novel intransitive verbs
Melissa Kline and Katherine Demuth. 2014 · 2014
Earlier work this paper cites.
Building machines that learn and think like people
Brenden M. Lake, Tomer D. Ullman, Joshua B. Tenenbaum, and Samuel J. Gershman. 2017 · 2017
Earlier work this paper cites.
How much does tokenization affect neural machine translation?
Miguel Domingo, Mercedes Garcıa-Martınez, Alexandre Helle, Francisco Casacuberta, and Manuel Herranz. 2018 · 2018
Earlier work this paper cites.
Improving text-to-SQL evaluation methodology
Catherine Finegan-Dollak, Jonathan K. Kummerfeld, Li Zhang, Karthik Ramanathan, Sesh Sadasivam, Rui Zhang, and Dragomir Radev. 2018 · 2018
Cited alongside, same era.
SentencePiece: A simple and language independent subword tokenizer and detokenizer for neural text processing
Taku Kudo and John Richardson. 2018 · 2018
Cited alongside, same era.
Generalization without systematicity: On the compositional skills of sequence-to-sequence recurrent networks
Brenden M. Lake and Marco Baroni. 2018 · 2018
Cited alongside, same era.
Measuring compositional generalization: A comprehensive method on realistic data
Daniel Keysers, Nathanael Schärli, Nathan Scales, Hylke Buisman, Daniel Furrer, Sergii Kashubin, Nikola Momchev, Danila Sinopalnikov, Lukasz Stafiniak, Tibor Tihon, Dmitry Tsarkov, Xiao Wang, Marc van Zee, and Olivier Bousquet. 2020 · 2020
Cited alongside, same era.
COGS: A compositional generalization challenge based on semantic interpretation
Najoung Kim and Tal Linzen. 2020 · 2020
Cited alongside, same era.
Improving compositional generalization with latent structure and data augmentation
Linlu Qiu, Peter Shaw, Panupong Pasupat, Paweł Krzysztof Nowak, Tal Linzen, Fei Sha, and Kristina Toutanova. 2021 · 2021
Later among the works it cites.
Compositional generalization and natural language variation: Can a semantic parsing approach handle both?
Peter Shaw, Ming-Wei Chang, Panupong Pasupat, and Kristina Toutanova. 2021 · 2021
Later among the works it cites.
Are pretrained convolutions better than pretrained Transformers?
Yi Tay, Mostafa Dehghani, Jai Prakash Gupta, Vamsi Aribandi, Dara Bahri, Zhen Qin, and Donald Metzler. 2021 · 2021
Later among the works it cites.
CodeT5: Identifier-aware unified pre-trained encoder-decoder models for code understanding and generation
Yue Wang, Weishi Wang, Shafiq Joty, and Steven C.H. Hoi. 2021 · 2021
Later among the works it cites.
mT5: A massively multilingual pre-trained text-to-text transformer
Linting Xue, Noah Constant, Adam Roberts, Mihir Kale, Rami Al-Rfou, Aditya Siddhant, Aditya Barua, and Colin Raffel. 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu. 2020 · 2020
Cited alongside, same era.
Lexicon learning for few shot sequence modeling
Ekin Akyürek and Jacob Andreas. 2021 · 2021
Cited alongside, same era.
Systematic generalization with Edge Transformers
Leon Bergen, Timothy O’Donnell, and Dzmitry Bahdanau. 2021 · 2021
Cited alongside, same era.
Meta-learning to compositionally generalize
Henry Conklin, Bailin Wang, Kenny Smith, and Ivan Titov. 2021 · 2021
Cited alongside, same era.
The devil is in the detail: Simple tricks improve systematic generalization of Transformers
Róbert Csordás, Kazuki Irie, and Juergen Schmidhuber. 2021 · 2021
Cited alongside, same era.
Unlocking compositional generalization in pre-trained models using intermediate representations
Jonathan Herzig, Peter Shaw, Ming-Wei Chang, Kelvin Guu, Panupong Pasupat, and Yuan Zhang. 2021 · 2021
Cited alongside, same era.
Initializing new word embeddings for pretrained language models
John Hewitt. 2021 · 2021
Cited alongside, same era.
Later among the works it cites.
SyGNS: A systematic generalization testbed based on natural language semantics
Hitomi Yanaka, Koji Mineshima, and Kentaro Inui. 2021 · 2021
Later among the works it cites.
Learning to generalize compositionally by transferring across semantic parsing tasks
Wang Zhu, Peter Shaw, Tal Linzen, and Fei Sha. 2021 · 2021
Later among the works it cites.
Unobserved local structures make compositional generalization hard
Ben Bogin, Shivanshu Gupta, and Jonathan Berant. 2022 · 2022
Closest in time.
Compositional semantic parsing with large language models
Andrew Drozdov, Nathanael Schärli, Ekin Akyürek, Nathan Scales, Xinying Song, Xinyun Chen, Olivier Bousquet, and Denny Zhou. 2022 · 2022
Closest in time.
LAGr: Label aligned graphs for better systematic generalization in semantic parsing
Dora Jambor and Dzmitry Bahdanau. 2022 · 2022
Closest in time.
Making transformers solve compositional tasks
Santiago Ontañon, Joshua Ainslie, Zachary Fisher, and Vaclav Cvicek. 2022 · 2022
Closest in time.
Do language models learn position-role mappings?
Jackson Petty, Michael Wilson, and Robert Frank. 2022 · 2022
Closest in time.
Evaluating the impact of model scale for compositional generalization in semantic parsing
Linlu Qiu, Peter Shaw, Panupong Pasupat, Tianze Shi, Jonathan Herzig, Emily Pitler, Fei Sha, and Kristina Toutanova. 2022 · 2022
Closest in time.
Compositional generalization requires compositional parsers
Pia Weißenhorn, Yuekun Yao, Lucia Donatelli, and Alexander Koller. 2022 · 2022
Closest in time.
ByT5: Towards a Token-Free Future with Pre-trained Byte-to-Byte Models
Linting Xue, Aditya Barua, Noah Constant, Rami Al-Rfou, Sharan Narang, Mihir Kale, Adam Roberts, and Colin Raffel. 2022 · 2022
Closest in time.
Disentangled sequence to sequence learning for compositional generalization
Hao Zheng and Mirella Lapata. 2022 · 2022
Closest in time.