Fetching the paper…
Reading the bibliography…
Flexible neural sequence models outperform grammar- and automaton-based counterparts on a variety of tasks.
Comparative study of the distribution of flora in a portion of the Alps and Jura
Paul Jaccard · 1901
Earlier work this paper cites.
Closure: Assessing systematic generalization of clevr models
Dzmitry Bahdanau, Harm de Vries, Timothy J O’Donnell, Shikhar Murty, Philippe Beaudoin, Yoshua Bengio, and Aaron Courville · 1912
Earlier work this paper cites.
Remarks on some nonparametric estimates of a density function
Murray Rosenblatt · 1956
Earlier work this paper cites.
The child’s learning of english morphology
Jean Berko · 1958
Earlier work this paper cites.
Binary codes capable of correcting deletions, insertions, and reversals
Vladimir I Levenshtein · 1966
Earlier work this paper cites.
Acquiring a single new word
Susan Carey and Elsa Bartlett · 1978
Earlier work this paper cites.
On learning the past tenses of english verbs
David E. Rumelhart and James L. McClelland · 1986
Earlier work this paper cites.
Connectionism and cognitive architecture: A critical analysis
Jerry A Fodor, Zenon W Pylyshyn, et al · 1988
Earlier work this paper cites.
Training with noise is equivalent to tikhonov regularization
Chris M Bishop · 1995
Earlier work this paper cites.
Learning string-edit distance
Eric Sven Ristad and Peter N Yianilos · 1998
Earlier work this paper cites.
Learning to learn: Introduction and overview
Sebastian Thrun and Lorien Pratt · 1998
Earlier work this paper cites.
Learning from imbalanced data sets: a comparison of various strategies
Nathalie Japkowicz et al · 2000
Earlier work this paper cites.
Lstm recurrent networks learn simple context-free and context-sensitive languages
Felix A Gers and E Schmidhuber · 2001
Earlier work this paper cites.
Smote: synthetic minority over-sampling technique
Nitesh V Chawla, Kevin W Bowyer, Lawrence O Hall, and W Philip Kegelmeyer · 2002
Earlier work this paper cites.
Polynomial identification in the limit of substitutable context-free languages
Alexander Clark and Rémi Eyraud · 2007
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V Le · 2014
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2015
Earlier work this paper cites.
Neural module networks
Jacob Andreas, Marcus Rohrbach, Trevor Darrell, and Dan Klein · 2016
Earlier work this paper cites.
Generating sentences from a continuous space
Samuel Bowman, Luke Vilnis, Oriol Vinyals, Andrew Dai, Rafal Jozefowicz, and Samy Bengio · 2016
Earlier work this paper cites.
Incorporating copying mechanism in sequence-to-sequence learning
Jiatao Gu, Zhengdong Lu, Hang Li, and Victor O.K. Li · 2016
Cited alongside, same era.
Data recombination for neural semantic parsing
Robin Jia and Percy Liang · 2016
Cited alongside, same era.
Key-value memory networks for directly reading documents
Alexander Miller, Adam Fisch, Jesse Dodge, Amir-Hossein Karimi, Antoine Bordes, and Jason Weston · 2016
Cited alongside, same era.
Compositional reasoning in early childhood
Steven Piantadosi and Richard Aslin · 2016
Cited alongside, same era.
Meta-learning with memory-augmented neural networks
Adam Santoro, Sergey Bartunov, Matthew Botvinick, Daan Wierstra, and Timothy Lillicrap · 2016
Cited alongside, same era.
Learning to reinforcement learn
Jane X. Wang, Zeb Kurth-Nelson, Hubert Soyer, Joel Z. Leibo, Dhruva Tirumala, Rémi Munos, Charles Blundell, Dharshan Kumaran, and Matt M. Botvinick · 2016
Cited alongside, same era.
Generating sentences by editing prototypes
Kelvin Guu, Tatsunori B. Hashimoto, Yonatan Oren, and Percy Liang · 2018
Later among the works it cites.
A retrieve-and-edit framework for predicting structured outputs
Tatsunori B Hashimoto, Kelvin Guu, Yonatan Oren, and Percy S Liang · 2018
Later among the works it cites.
Data augmentation by pairing samples for images classification
Hiroshi Inoue · 2018
Later among the works it cites.
UniMorph 2.0: Universal Morphology
Christo Kirov, Ryan Cotterell, John Sylak-Glassman, Géraldine Walther, Ekaterina Vylomova, Patrick Xia, Manaal Faruqui, Sabrina J. Mielke, Arya McCarthy, Sandra Kübler, David Yarowsky, Jason Eisner, and Mans Hulden · 2018
Later among the works it cites.
Generalization without systematicity: On the compositional skills of sequence-to-sequence recurrent networks
Brenden Lake and Marco Baroni · 2018
Later among the works it cites.
Rearranging the familiar: Testing compositional generalization in recurrent networks
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sequence-based structured prediction for semantic parsing
Chunyang Xiao, Marc Dymetman, and Claire Gardent · 2016
Cited alongside, same era.
Knet: beginning deep learning with 100 lines of julia
Deniz Yuret · 2016
Cited alongside, same era.
Julia: A fresh approach to numerical computing
Jeff Bezanson, Alan Edelman, Stefan Karpinski, and Viral B Shah · 2017
Cited alongside, same era.
Making neural programming architectures generalize via recursion
Jonathon Cai, Richard Shin, and Dawn Song · 2017
Cited alongside, same era.
Data augmentation for low-resource neural machine translation
Marzieh Fadaee, Arianna Bisazza, and Christof Monz · 2017
Cited alongside, same era.
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning
Justin Johnson, Bharath Hariharan, Laurens van der Maaten, Li Fei-Fei, C Lawrence Zitnick, and Ross Girshick · 2017
Cited alongside, same era.
João Loula, Marco Baroni, and Brenden Lake · 2018
Later among the works it cites.
Interactive supercomputing on 40,000 cores for machine learning and data analysis
Albert Reuther, Jeremy Kepner, Chansup Byun, Siddharth Samsi, William Arcand, David Bestor, Bill Bergeron, Vijay Gadepally, Michael Houle, Matthew Hubbell, et al · 2018
Later among the works it cites.
On the practical computational power of finite precision RNNs for language recognition
Gail Weiss, Yoav Goldberg, and Eran Yahav · 2018
Later among the works it cites.
Retrieve and refine: Improved sequence generation models for dialogue
Jason Weston, Emily Dinan, and Alexander Miller · 2018
Later among the works it cites.
Morphological analysis using a sequence decoder
Ekin Akyürek, Erenay Dayanık, and Deniz Yuret · 2019
Later among the works it cites.
Human few-shot learning of compositional instructions
B. Lake, Tal Linzen, and M. Baroni · 2019
Later among the works it cites.
Compositional generalization through meta sequence-to-sequence learning
Brenden M Lake · 2019
Later among the works it cites.
Compositional generalization in a deep seq2seq model by separating syntax and semantics
Jake Russin, Jason Jo, Randall C O’Reilly, and Yoshua Bengio · 2019
Later among the works it cites.
Good-enough compositional data augmentation
Jacob Andreas · 2020
Closest in time.
Permutation equivariant models for compositional generalization in language
Jonathan Gordon, David Lopez-Paz, Marco Baroni, and Diane Bouchacourt · 2020
Closest in time.
Realm: Retrieval-augmented language model pre-training
Kelvin Guu, Kenton Lee, Zora Tung, Panupong Pasupat, and Ming-Wei Chang · 2020
Closest in time.
Measuring compositional generalization: A comprehensive method on realistic data
Daniel Keysers, Nathanael Schärli, Nathan Scales, Hylke Buisman, Daniel Furrer, Sergii Kashubin, Nikola Momchev, Danila Sinopalnikov, Lukasz Stafiniak, Tibor Tihon, Dmitry Tsarkov, Xiao Wang, Marc van Zee, and Olivier Bousquet · 2020
Closest in time.
Generalization through memorization: Nearest neighbor language models
Urvashi Khandelwal, Omer Levy, Dan Jurafsky, Luke Zettlemoyer, and Mike Lewis · 2020
Closest in time.
Posterior control of blackbox generation
Xiang Lisa Li and Alexander M Rush · 2020
Closest in time.