Fetching the paper…
Reading the bibliography…
While long short-term memory (LSTM) neural net architectures are designed to capture sequence information, human language is generally composed of hierarchical structures.
Recognizing well-parenthesized expressions in the streaming model
F. Magniez, C. Mathieu, and A. Nayak. 2014 · 1905
Earlier work this paper cites.
The algebraic theory of context-free languages
Noam Chomsky and Marcel P Schützenberger. 1963 · 1963
Earlier work this paper cites.
Context-free languages and pushdown automata
Jean-Michel Autebert, Jean Berstel, and Luc Boasson. 1997 · 1997
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Fractal encoding of context-free grammars in connectionist networks
Whitney Tabor. 2000 · 2000
Earlier work this paper cites.
Lstm recurrent networks learn simple context-free and context-sensitive languages
Felix A Gers and E Schmidhuber. 2001 · 2001
Earlier work this paper cites.
Stack-like and queue-like dynamics in recurrent neural networks
André Grüning. 2006 · 2006
Earlier work this paper cites.
Recurrent neural network based language model
Tomáš Mikolov, Martin Karafiát, Lukáš Burget, Jan Černockỳ, and Sanjeev Khudanpur. 2010 · 2010
Earlier work this paper cites.
Processing of nested and cross-serial dependencies: an automaton perspective on srn behaviour
Christo Kirov and Robert Frank. 2012 · 2012
Earlier work this paper cites.
Lstm neural networks for language modeling
Martin Sundermeyer, Ralf Schlüter, and Hermann Ney. 2012 · 2012
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
Alex Graves, Greg Wayne, and Ivo Danihelka. 2014 · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Cited alongside, same era.
Structures, not strings: linguistics as part of the cognitive sciences
Martin BH Everaert, Marinus AC Huybregts, Noam Chomsky, Robert C Berwick, and Johan J Bolhuis. 2015 · 2015
Cited alongside, same era.
Inferring algorithmic patterns with stack-augmented recurrent nets
Armand Joulin and Tomas Mikolov. 2015 · 2015
Cited alongside, same era.
Visualizing and understanding recurrent networks
Andrej Karpathy, Justin Johnson, and Fei-Fei Li. 2015 · 2015
Cited alongside, same era.
End-to-end memory networks
Sainbayar Sukhbaatar, arthur szlam, Jason Weston, and Rob Fergus. 2015 · 2015
Cited alongside, same era.
Increasing the interpretability of recurrent neural networks using hidden markov models
Viktoriya Krakovna and Finale Doshi-Velez. 2016 · 2016
Later among the works it cites.
Assessing the ability of lstms to learn syntax-sensitive dependencies
Tal Linzen, Emmanuel Dupoux, and Yoav Goldberg. 2016 · 2016
Later among the works it cites.
Does string-based neural mt learn source syntax?
Xing Shi, Inkit Padhi, and Kevin Knight. 2016 · 2016
Later among the works it cites.
What do neural machine translation models learn about morphology?
Yonatan Belinkov, Nadir Durrani, Fahim Dalvi, Hassan Sajjad, and James Glass. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kai Sheng Tai, Richard Socher, and Christopher D. Manning. 2015 · 2015
Cited alongside, same era.
Grammar as a foreign language
Oriol Vinyals, Łukasz Kaiser, Terry Koo, Slav Petrov, Ilya Sutskever, and Geoffrey Hinton. 2015 · 2015
Cited alongside, same era.
Why only us: Language and evolution
Robert C Berwick and Noam Chomsky. 2016 · 2016
Cited alongside, same era.
Capacity and trainability in recurrent neural networks
Jasmine Collins, Jascha Sohl-Dickstein, and David Sussillo. 2016 · 2016
Cited alongside, same era.
Recurrent neural network grammars
Chris Dyer, Adhiguna Kuncoro, Miguel Ballesteros, and Noah A Smith. 2016 · 2016
Cited alongside, same era.
Character-aware neural language models
Yoon Kim, Yacine Jernite, David Sontag, and Alexander M Rush. 2016 · 2016
Cited alongside, same era.
Simple and accurate dependency parsing using bidirectional lstm feature representations
Eliyahu Kiperwasser and Yoav Goldberg. 2016 · 2016
Cited alongside, same era.
Moshe Looks, Marcello Herreshoff, DeLesley Hutchins, and Peter Norvig. 2017 · 2017
Later among the works it cites.
Visualizing the hidden activity of artificial neural networks
P. E. Rauber, S. G. Fadel, A. X. Falcão, and A. C. Telea. 2017 · 2017
Later among the works it cites.
Don’t decay the learning rate, increase the batch size
Samuel L Smith, Pieter-Jan Kindermans, and Quoc V Le. 2017 · 2017
Later among the works it cites.
Can recurrent neural networks learn nested recursion?
Jean-Philippe Bernardy. 2018 · 2018
Closest in time.
What is the significance of the chomsky-schützenberger theorem about representing context-free languages?
Michal Forišek. 2018 · 2018
Closest in time.
Memorize or generalize? searching for a compositional RNN in a haystack
Adam Liska, Germán Kruszewski, and Marco Baroni. 2018 · 2018
Closest in time.
Simple recurrent networks learn context-free and context-sensitive languages by counting
Paul Rodriguez. 2001 · 2093
Closest in time.