Fetching the paper…
Reading the bibliography…
We model the recursive production property of context-free grammars for natural and synthetic languages.
Unsupervised recurrent neural network grammars
Yoon Kim, Alexander M Rush, Lei Yu, Adhiguna Kuncoro, Chris Dyer, and Gábor Melis. 2019b · 1904
Earlier work this paper cites.
Compositional generalization in a deep seq2seq model by separating syntax and semantics
Jake Russin, Jason Jo, and Randall C O’Reilly. 2019 · 1904
Earlier work this paper cites.
An inequality with applications to statistical estimation for probabilistic functions of markov processes and to a model for ecology
Leonard E Baum and John Alonzo Eagon. 1967 · 1967
Earlier work this paper cites.
Error bounds for convolutional codes and an asymptotically optimum decoding algorithm
Andrew Viterbi. 1967 · 1967
Earlier work this paper cites.
Growth transformations for functions on manifolds
Leonard E Baum and George Sell. 1968 · 1968
Earlier work this paper cites.
Trainable grammars for speech recognition
James K Baker. 1979 · 1979
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams. 1992 · 1992
Earlier work this paper cites.
A constructive definition of dirichlet priors
Jayaram Sethuraman. 1994 · 1994
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Natural language grammar induction using a constituent-context model
Dan Klein and Christopher D Manning. 2001 · 2001
Earlier work this paper cites.
R Thomas McCoy, Robert Frank, and Tal Linzen. 2020 · 2001
Earlier work this paper cites.
Hierarchical topic models and the nested chinese restaurant process
Thomas L Griffiths, Michael I Jordan, Joshua B Tenenbaum, and David M Blei. 2004 · 2004
Earlier work this paper cites.
Natural language grammar induction with a generative constituent-context model
Dan Klein and Christopher D Manning. 2005 · 2005
Earlier work this paper cites.
An all-subtrees approach to unsupervised parsing
Rens Bod. 2006 · 2006
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Alex Graves, Santiago Fernández, Faustino Gomez, and Jürgen Schmidhuber. 2006 · 2006
Earlier work this paper cites.
The infinite markov model
Daichi Mochihashi and Eiichiro Sumita. 2008 · 2008
Earlier work this paper cites.
Tree-structured stick breaking for hierarchical data
Zoubin Ghahramani, Michael I Jordan, and Ryan P Adams. 2010 · 2010
Earlier work this paper cites.
Learning continuous phrase representations and syntactic parsing with recursive neural networks
Richard Socher, Christopher D Manning, and Andrew Y Ng. 2010 · 2010
Cited alongside, same era.
Auto-encoding variational bayes
Diederik P Kingma and Max Welling. 2013 · 2013
Cited alongside, same era.
Recursive deep models for semantic compositionality over a sentiment treebank
Richard Socher, Alex Perelygin, Jean Wu, Jason Chuang, Christopher D Manning, Andrew Ng, and Christopher Potts. 2013 · 2013
Cited alongside, same era.
On the properties of neural machine translation: Encoder-decoder approaches
Kyunghyun Cho, Bart Van Merriënboer, Dzmitry Bahdanau, and Yoshua Bengio. 2014 · 2014
Cited alongside, same era.
A large annotated corpus for learning natural language inference
Samuel R Bowman, Gabor Angeli, Christopher Potts, and Christopher D Manning. 2015 · 2015
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Later among the works it cites.
Jump to better conclusions: Scan both left and right
Joost Bastings, Marco Baroni, Jason Weston, Kyunghyun Cho, and Douwe Kiela. 2018 · 2018
Later among the works it cites.
Tree-to-tree neural networks for program translation
Xinyun Chen, Chang Liu, and Dawn Song. 2018 · 2018
Later among the works it cites.
Top-down tree structured decoding with syntactic connections for neural machine translation and parsing
Jetic Gū, Hassan S. Shavarani, and Anoop Sarkar. 2018 · 2018
Later among the works it cites.
Learning hierarchical structures on-the-fly with a recurrent-recursive model for sequences
Athul Paul Jacob, Zhouhan Lin, Alessandro Sordoni, and Yoshua Bengio. 2018 · 2018
Later among the works it cites.
Ordered neurons: Integrating tree structures into recurrent neural networks
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Tree-structured decoding with doubly-recurrent neural networks
David Alvarez-Melis and Tommi S Jaakkola. 2016 · 2016
Cited alongside, same era.
Jimmy Lei Ba, Jamie Ryan Kiros, and Geoffrey E Hinton. 2016 · 2016
Cited alongside, same era.
Hierarchical multiscale recurrent neural networks
Junyoung Chung, Sungjin Ahn, and Yoshua Bengio. 2016 · 2016
Cited alongside, same era.
Recurrent neural network grammars
Chris Dyer, Adhiguna Kuncoro, Miguel Ballesteros, and Noah A Smith. 2016 · 2016
Cited alongside, same era.
Multi30k: Multilingual English-German image descriptions
Desmond Elliott, Stella Frank, Khalil Sima’an, and Lucia Specia. 2016 · 2016
Cited alongside, same era.
Tree-to-sequence attentional neural machine translation
Akiko Eriguchi, Kazuma Hashimoto, and Yoshimasa Tsuruoka. 2016 · 2016
Cited alongside, same era.
Sequence-level knowledge distillation
Yoon Kim and Alexander M Rush. 2016 · 2016
Cited alongside, same era.
Yikang Shen, Shawn Tan, Alessandro Sordoni, and Aaron Courville. 2018 · 2018
Later among the works it cites.
Do latent tree learning models identify meaningful structure in sentences?
Adina Williams, Andrew Drozdov*, and Samuel R Bowman. 2018 · 2018
Later among the works it cites.
Non-monotonic sequential text generation
Kyunghyun Cho, Hal Daumé III, Sean Welleck, et al. 2019 · 2019
Later among the works it cites.
Top-down structurally-constrained neural response generation with lexicalized probabilistic context-free grammar
Wenchao Du and Alan W Black. 2019 · 2019
Later among the works it cites.
Insertion-based decoding with automatically inferred generation order
Jiatao Gu, Qi Liu, and Kyunghyun Cho. 2019 · 2019
Later among the works it cites.
Compositional generalization for primitive substitutions
Yuanpeng Li, Liang Zhao, Jianyu Wang, and Joel Hestness. 2019 · 2019
Later among the works it cites.
Ordered memory
Yikang Shen, Shawn Tan, Arian Hosseini, Zhouhan Lin, Alessandro Sordoni, and Aaron C Courville. 2019 · 2019
Later among the works it cites.
Insertion transformer: Flexible sequence generation via insertion operations
Mitchell Stern, William Chan, Jamie Kiros, and Jakob Uszkoreit. 2019 · 2019
Later among the works it cites.
Xlnet: Generalized autoregressive pretraining for language understanding
Zhilin Yang, Zihang Dai, Yiming Yang, Jaime Carbonell, Russ R Salakhutdinov, and Quoc V Le. 2019 · 2019
Later among the works it cites.
Exploiting syntactic structure for better language modeling: A syntactic distance approach
Wenyu Du, Zhouhan Lin, Yikang Shen, Timothy J. O’Donnell, Yoshua Bengio, and Yue Zhang. 2020 · 2020
Closest in time.
Latent-variable non-autoregressive neural machine translation with deterministic inference using a delta posterior
Raphael Shu, Jason Lee, Hideki Nakayama, and Kyunghyun Cho. 2020 · 2020
Closest in time.