Fetching the paper…
Reading the bibliography…
Neural network models often generalize poorly to mismatched domains or distributions.
Well-read students learn better: The impact of student initialization on knowledge distillation
Iulia Turc, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 1908
Earlier work this paper cites.
Measuring compositional generalization: A comprehensive method on realistic data
Daniel Keysers, Nathanael Schärli, Nathan Scales, Hylke Buisman, Daniel Furrer, Sergii Kashubin, Nikola Momchev, Danila Sinopalnikov, Lukasz Stafiniak, Tibor Tihon, Dmitry Tsarkov, Xiao Wang, Marc van Zee, and Olivier Bousquet · 1912
Earlier work this paper cites.
Using Inductive Logic Programming to Automate the Construction of Natural Language Parsers
John M. Zelle · 1995
Earlier work this paper cites.
Multitask learning
Rich Caruana · 1998
Earlier work this paper cites.
Learning to learn
Sebastian Thrun and Lorien Pratt · 1998
Earlier work this paper cites.
Using multiple clause constructors in inductive logic programming for semantic parsing
Lappoon R Tang and Raymond J Mooney · 2001
Earlier work this paper cites.
Task clustering and gating for Bayesian multitask learning
Bart Bakker and Tom Heskes · 2003
Earlier work this paper cites.
Learning to transform natural to formal languages
Rohit J Kate, Yuk Wah Wong, and Raymond J Mooney · 2005
Earlier work this paper cites.
Compositional generalization in semantic parsing: Pre-training vs. specialized architectures
Daniel Furrer, Marc van Zee, Nathan Scales, and Nathanael Schaerli · 2007
Earlier work this paper cites.
Learning phrase representations using RNN encoder–decoder for statistical machine translation
Kyunghyun Cho, Bart van Merriënboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Earlier work this paper cites.
A joint many-task model: Growing a neural network for multiple NLP tasks
Kazuma Hashimoto, Caiming Xiong, Yoshimasa Tsuruoka, and Richard Socher · 2016
Earlier work this paper cites.
Deep multi-task learning with low level tasks supervised at lower layers
Anders Sogaard and Yoav Goldberg · 2016
Earlier work this paper cites.
An overview of multi-task learning in deep neural networks
Sebastian Ruder · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Improving text-to-SQL evaluation methodology
Catherine Finegan-Dollak, Jonathan K. Kummerfeld, Li Zhang, Karthik Ramanathan, Sesh Sadasivam, Rui Zhang, and Dragomir Radev · 2018
Earlier work this paper cites.
Meta-learning for low-resource neural machine translation
Jiatao Gu, Yong Wang, Yun Chen, Victor O. K. Li, and Kyunghyun Cho · 2018
Earlier work this paper cites.
Generalization without systematicity: On the compositional skills of sequence-to-sequence recurrent networks
Brenden M. Lake and Marco Baroni · 2018
Cited alongside, same era.
Sentence encoders on stilts: Supplementary training on intermediate labeled-data tasks
Jason Phang, Thibault Févry, and Samuel R Bowman · 2018
Cited alongside, same era.
Syntactic scaffolds for semantic structures
Swabha Swayamdipta, Sam Thomson, Kenton Lee, Luke Zettlemoyer, Chris Dyer, and Noah A. Smith · 2018
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Cited alongside, same era.
Compositional generalization through meta sequence-to-sequence learning
Brenden M Lake · 2019
Cited alongside, same era.
COGS: A compositional generalization challenge based on semantic interpretation
Najoung Kim and Tal Linzen · 2020
Later among the works it cites.
Compositional generalization by learning analytical expressions
Qian Liu, Shengnan An, Jian-Guang Lou, Bei Chen, Zeqi Lin, Yan Gao, Bin Zhou, Nanning Zheng, and Dongmei Zhang · 2020
Later among the works it cites.
Learning compositional rules via neural program synthesis
Maxwell I Nye, Armando Solar-Lezama, Joshua B Tenenbaum, and Brenden M Lake · 2020
Later among the works it cites.
Improving compositional generalization in semantic parsing
Inbar Oren, Jonathan Herzig, Nitish Gupta, Matt Gardner, and Jonathan Berant · 2020
Later among the works it cites.
Intermediate-task transfer learning with pretrained language models: When and why does it work?
Yada Pruksachatkun, Jason Phang, Haokun Liu, Phu Mon Htut, Xiaoyi Zhang, Richard Yuanzhe Pang, Clara Vania, Katharina Kann, and Samuel R. Bowman · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Compositional generalization for primitive substitutions
Yuanpeng Li, Liang Zhao, Jianyu Wang, and Joel Hestness · 2019
Cited alongside, same era.
When does label smoothing help?
R. Müller, Simon Kornblith, and Geoffrey E. Hinton · 2019
Cited alongside, same era.
Compositional generalization in a deep seq2seq model by separating syntax and semantics
Jake Russin, Jason Jo, Randall C O’Reilly, and Yoshua Bengio · 2019
Cited alongside, same era.
Learning to recombine and resample data for compositional generalization
Ekin Akyürek, Afra Feyza Akyurek, and Jacob Andreas · 2020
Cited alongside, same era.
Good-enough compositional data augmentation
Jacob Andreas · 2020
Cited alongside, same era.
Low-resource domain adaptation for compositional task-oriented semantic parsing
Xilun Chen, Asish Ghoshal, Yashar Mehdad, Luke Zettlemoyer, and Sonal Gupta · 2020
Cited alongside, same era.
Permutation equivariant models for compositional generalization in language
Jonathan Gordon, David Lopez-Paz, Marco Baroni, and Diane Bouchacourt · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu · 2020
Later among the works it cites.
Exploring and predicting transferability across NLP tasks
Tu Vu, Tong Wang, Tsendsuren Munkhdalai, Alessandro Sordoni, Adam Trischler, Andrew Mattarella-Micke, Subhransu Maji, and Mohit Iyyer · 2020
Later among the works it cites.
Compositional generalization via semantic tagging
Hao Zheng and Mirella Lapata · 2020
Later among the works it cites.
Muppet: Massive multi-task representations with pre-finetuning
Armen Aghajanyan, Anchit Gupta, Akshat Shrivastava, Xilun Chen, Luke Zettlemoyer, and Sonal Gupta · 2021
Closest in time.
Meta-learning to compositionally generalize
Henry Conklin, Bailin Wang, Kenny Smith, and Ivan Titov · 2021
Closest in time.
Unlocking compositional generalization in pre-trained models using intermediate representations
Jonathan Herzig, Peter Shaw, Ming-Wei Chang, Kelvin Guu, Panupong Pasupat, and Yuan Zhang · 2021
Closest in time.
Learning algebraic recombination for compositional generalization, 2021
Chenyao Liu, Shengnan An, Zeqi Lin, Qian Liu, Bei Chen, Jian-Guang Lou, Lijie Wen, Nanning Zheng, and Dongmei Zhang · 2021
Closest in time.
Finding needles in a haystack: Sampling structurally-diverse training sets from synthetic data for compositional generalization, 2021
Inbar Oren, Jonathan Herzig, and Jonathan Berant · 2021
Closest in time.
Compositional generalization and natural language variation: Can a semantic parsing approach handle both?
Peter Shaw, Ming-Wei Chang, Panupong Pasupat, and Kristina Toutanova · 2021
Closest in time.
Compositional generalization for neural semantic parsing via span-level supervised attention
Pengcheng Yin, Hao Fang, Graham Neubig, Adam Pauls, Emmanouil Antonios Platanios, Yu Su, Sam Thomson, and Jacob Andreas · 2021
Closest in time.