Fetching the paper…
Reading the bibliography…
While neural network models have been successfully applied to domains that require substantial generalisation skills, recent studies have implied that they struggle when solving the task they are trained on requires inferring its underlying compositional structure.
Three models for the description of language
Chomsky, N.: · 1956
Earlier work this paper cites.
Language identification in the limit
Gold, E.M., et al.: · 1967
Earlier work this paper cites.
Modeling by shortest data description
Rissanen, J.: · 1978
Earlier work this paper cites.
Inductive inference: Theory and methods
Angluin, D., Smith, C.H.: · 1983
Earlier work this paper cites.
Connectionism and cognitive architecture: A critical analysis
Fodor, J.A., Pylyshyn, Z.W.: · 1988
Earlier work this paper cites.
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning
Johnson, J., Hariharan, B., van der Maaten, L., Fei-Fei, L., Zitnick, C.L., Girshick, R.: · 1997
Earlier work this paper cites.
Long short-term memory
Hochreiter, S., Schmidhuber, J.: · 1997
Earlier work this paper cites.
Universal artificial intelligence: Sequential decisions based on algorithmic probability
Hutter, M.: · 2004
Earlier work this paper cites.
Motor primitives in vertebrates and invertebrates
Flash, T., Hochner, B.: · 2005
Earlier work this paper cites.
Learning continuous phrase representations and syntactic parsing with recursive neural networks
Socher, R., Manning, C.D., Ng, A.Y.: · 2010
Earlier work this paper cites.
Compositionality
Szabó, Z.G.: · 2010
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Sutskever, I., Vinyals, O., Le, Q.V.: · 2014
Earlier work this paper cites.
Empirical evaluation of gated recurrent neural networks on sequence modeling
Chung, J., Gulcehre, C., Cho, K., Bengio, Y.: · 2014
Cited alongside, same era.
Learning phrase representations using RNN encoder–decoder for statistical machine translation
Cho, K., van Merriënboer, B., Gulcehre, C., Schwenk, F.B.H., Bengio, Y.: · 2014
Cited alongside, same era.
Human-level concept learning through probabilistic program induction
Lake, B.M., Salakhutdinov, R., Tenenbaum, J.B.: · 2015
Cited alongside, same era.
Neural programmer-interpreters
Reed, S., De Freitas, N.: · 2015
Cited alongside, same era.
Kurach, K., Andrychowicz, M., Sutskever, I.: · 2015
Cited alongside, same era.
Supervised attentions for neural machine translation
Mi, H., Wang, Z., Ittycheriah, A.: · 2016
Later among the works it cites.
Proceedings of the first workshop on building linguistically generalizable nlp systems
Bender, E., Daumé III, H., Ettinger, A., Rao, S.: · 2017
Later among the works it cites.
Lake, B.M., Baroni, M.: · 2017
Later among the works it cites.
Making neural programming architectures generalize via recursion
Cai, J., Shin, R., Song, D.: · 2017
Later among the works it cites.
Exploring human-like attention supervision in visual question answering
Qiao, T., Dong, J., Xu, D.: · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Neural programmer: Inducing latent programs with gradient descent
Neelakantan, A., Le, Q.V., Sutskever, I.: · 2015
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Bahdanau, D., Cho, K., Bengio, Y.: · 2015
Cited alongside, same era.
Effective approaches to attention-based neural machine translation
Luong, T., Pham, H., Manning, C.D.: · 2015
Cited alongside, same era.
Mastering the game of go with deep neural networks and tree search
Silver, D., Huang, A., Maddison, C.J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al.: · 2016
Cited alongside, same era.
Probing the compositionality of intuitive functions
Schulz, E., Tenenbaum, J., Duvenaud, D.K., Speekenbrink, M., Gershman, S.J.: · 2016
Cited alongside, same era.
Modular multitask reinforcement learning with policy sketches
Andreas, J., Klein, D., Levine, S.: · 2016
Cited alongside, same era.
Vqs: Linking segmentations to questions and answers for supervised attention in vqa and question-focused semantic segmentation
Gan, C., Li, Y., Li, H., Sun, C., Gong, B.: · 2017
Later among the works it cites.
Categorical reparameterization with gumbel-softmax
Jang, E., Gu, S., Poole, B.: · 2017
Later among the works it cites.
Memorize or generalize? searching for a compositional rnn in a haystack
Liška, A., Kruszewski, G., Baroni, M.: · 2018
Closest in time.
Lstms can learn syntax-sensitive dependencies well, but modeling structure makes them better
Kuncoro, A., Dyer, C., Hale, J., Yogatama, D., Clark, S., Blunsom, P.: · 2018
Closest in time.
The fine line between linguistic generalization and failure in seq2seq-attention models
Weber, N., Shekhar, L., Balasubramanian, N.: · 2018
Closest in time.
Compositional attention networks for machine reasoning
Hudson, D.A., Manning, C.D.: · 2018
Closest in time.