Fetching the paper…
Reading the bibliography…
The design of neural architectures for structured objects is typically guided by experimental insights rather than a formal process.
Long short-term memory
Hochreiter, Sepp and Schmidhuber, Jürgen · 1997
Earlier work this paper cites.
Text classification using string kernels
Lodhi, Huma, Saunders, Craig, Shawe-Taylor, John, Cristianini, Nello, and Watkins, Chris · 2002
Earlier work this paper cites.
On graph kernels: Hardness results and efficient alternatives
Gärtner, Thomas, Flach, Peter, and Wrobel, Stefan · 2003
Earlier work this paper cites.
Expressivity versus efficiency of graph kernels
Ramon, Jan and Gärtner, Thomas · 2003
Earlier work this paper cites.
Kernel methods for deep learning
Cho, Youngmin and Saul, Lawrence K · 2009
Earlier work this paper cites.
Graph kernels
Vishwanathan, S Vichy N, Schraudolph, Nicol N, Kondor, Risi, and Borgwardt, Karsten M · 2010
Earlier work this paper cites.
Learning kernel-based halfspaces with the 0-1 loss
Shalev-Shwartz, Shai, Shamir, Ohad, and Sridharan, Karthik · 2011
Earlier work this paper cites.
Weisfeiler-lehman graph kernels
Shervashidze, Nino, Schweitzer, Pascal, Leeuwen, Erik Jan van, Mehlhorn, Kurt, and Borgwardt, Karsten M · 2011
Earlier work this paper cites.
Semi-supervised recursive autoencoders for predicting sentiment distributions
Socher, Richard, Pennington, Jeffrey, Huang, Eric H, Ng, Andrew Y, and Manning, Christopher D · 2011
Earlier work this paper cites.
Improving neural networks by preventing co-adaptation of feature detectors
Hinton, Geoffrey E, Srivastava, Nitish, Krizhevsky, Alex, Sutskever, Ilya, and Salakhutdinov, Ruslan R · 2012
Earlier work this paper cites.
Spectral networks and locally connected networks on graphs
Bruna, Joan, Zaremba, Wojciech, Szlam, Arthur, and LeCun, Yann · 2013
Earlier work this paper cites.
Recursive deep models for semantic compositionality over a sentiment treebank
Socher, Richard, Perelygin, Alex, Wu, Jean, Chuang, Jason, Manning, Christopher D., Ng, Andrew Y., and Potts, Christopher · 2013
Earlier work this paper cites.
Empirical evaluation of gated recurrent neural networks on sequence modeling
Chung, Junyoung, Gulcehre, Caglar, Cho, KyungHyun, and Bengio, Yoshua · 2014
Earlier work this paper cites.
Deep recursive neural networks for compositionality in language
Irsoy, Ozan and Cardie, Claire · 2014
Earlier work this paper cites.
A neural network for factoid question answering over paragraphs
Iyyer, Mohit, Boyd-Graber, Jordan, Claudino, Leonardo, Socher, Richard, and Daumé III, Hal · 2014
Earlier work this paper cites.
A convolutional neural network for modelling sentences
Kalchbrenner, Nal, Grefenstette, Edward, and Blunsom, Phil · 2014
Earlier work this paper cites.
Convolutional neural networks for sentence classification
Kim, Yoon · 2014
Cited alongside, same era.
Distributed representations of sentences and documents
Le, Quoc and Mikolov, Tomas · 2014
Cited alongside, same era.
Glove: Global vectors for word representation
Pennington, Jeffrey, Socher, Richard, and Manning, Christopher D · 2014
Cited alongside, same era.
Recurrent neural network regularization
Zaremba, Wojciech, Sutskever, Ilya, and Vinyals, Oriol · 2014
Cited alongside, same era.
Deep convolutional networks are hierarchical kernel machines
Anselmi, Fabio, Rosasco, Lorenzo, Tan, Cheston, and Poggio, Tomaso · 2015
Cited alongside, same era.
Convolutional networks on graphs for learning molecular fingerprints
Improved semantic representations from tree-structured long short-term memory networks
Tai, Kai Sheng, Socher, Richard, and Manning, Christopher D · 2015
Later among the works it cites.
Strongly-typed recurrent neural networks
Balduzzi, David and Ghifary, Muhammad · 2016
Later among the works it cites.
Long short-term memory networks for machine reading
Cheng, Jianpeng, Dong, Li, and Lapata, Mirella · 2016
Later among the works it cites.
Discriminative embeddings of latent variable models for structured data
Dai, Hanjun, Dai, Bo, and Song, Le · 2016
Later among the works it cites.
Daniely, Amit, Frostig, Roy, and Singer, Yoram · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Duvenaud, David K, Maclaurin, Dougal, Iparraguirre, Jorge, Bombarell, Rafael, Hirzel, Timothy, Aspuru-Guzik, Alán, and Adams, Ryan P · 2015
Cited alongside, same era.
Transition-based dependency parsing with stack long short-term memory
Dyer, Chris, Ballesteros, Miguel, Ling, Wang, Matthews, Austin, and Smith, Noah A · 2015
Cited alongside, same era.
Greff, Klaus, Srivastava, Rupesh Kumar, Koutník, Jan, Steunebrink, Bas R, and Schmidhuber, Jürgen · 2015
Cited alongside, same era.
Steps toward deep kernel methods from infinite neural networks
Hazan, Tamir and Jaakkola, Tommi · 2015
Cited alongside, same era.
Deep convolutional networks on graph-structured data
Henaff, Mikael, Bruna, Joan, and LeCun, Yann · 2015
Cited alongside, same era.
Deep unordered composition rivals syntactic methods for text classification
Iyyer, Mohit, Manjunatha, Varun, Boyd-Graber, Jordan, and Daumé III, Hal · 2015
Cited alongside, same era.
Character-aware neural language models
Kim, Yoon, Jernite, Yacine, Sontag, David, and Rush, Alexander M · 2015
Cited alongside, same era.
Recurrent neural network grammars
Dyer, Chris, Kuncoro, Adhiguna, Ballesteros, Miguel, and Smith, Noah A · 2016
Later among the works it cites.
A theoretically grounded application of dropout in recurrent neural networks
Gal, Yarin and Ghahramani, Zoubin · 2016
Later among the works it cites.
Improper deep kernels
Heinemann, Uri, Livni, Roi, Eban, Elad, Elidan, Gal, and Globerson, Amir · 2016
Later among the works it cites.
Ask me anything: Dynamic memory networks for natural language processing
Kumar, Ankit, Irsoy, Ozan, Ondruska, Peter, Iyyer, Mohit, James Bradbury, Ishaan Gulrajani, Zhong, Victor, Paulus, Romain, and Socher, Richard · 2016
Later among the works it cites.
Pointer sentinel mixture models
Merity, Stephen, Xiong, Caiming, Bradbury, James, and Socher, Richard · 2016
Later among the works it cites.
Using the output embedding to improve language models
Press, Ofir and Wolf, Lior · 2016
Later among the works it cites.
Value iteration networks
Tamar, Aviv, Levine, Sergey, Abbeel, Pieter, Wu, Yi, and Thomas, Garrett · 2016
Later among the works it cites.
ℓ 1 \ell_{1} -regularized neural networks are improperly learnable in polynomial time
Zhang, Yuchen, Lee, Jason D., and Jordan, Michael I · 2016
Later among the works it cites.
Zilly, Julian Georg, Srivastava, Rupesh Kumar, Koutník, Jan, and Schmidhuber, Jürgen · 2016
Later among the works it cites.
Neural architecture search with reinforcement learning
Zoph, Barret and Le, Quoc V · 2016
Later among the works it cites.
Recurrent additive networks
Lee, Kenton, Levy, Omer, and Zettlemoyer, Luke · 2017
Closest in time.