Fetching the paper…
Reading the bibliography…
We describe DyNet, a toolkit for implementing neural network models based on dynamic declaration of network structure.
A simple automatic derivative evaluation program
R.E. Wengert · 1964
Earlier work this paper cites.
A method of solving a convex programming problem with convergence rate o (1/k2)
Yurii Nesterov · 1983
Earlier work this paper cites.
Finding structure in time
Jeffrey L Elman · 1990
Earlier work this paper cites.
Automatic differentiation of algorithms: theory, implementation, and application
Andreas Griewank · 1991
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Classes for fast maximum entropy training
Joshua Goodman · 2001
Earlier work this paper cites.
Torch: a modular machine learning software library
Ronan Collobert, Samy Bengio, and Johnny Mariéthoz · 2002
Earlier work this paper cites.
A neural probabilistic language model
Yoshua Bengio, Réjean Ducharme, Pascal Vincent, and Christian Jauvin · 2003
Earlier work this paper cites.
Compiling comp ling: Practical weighted dynamic programming and the Dyna language
Jason Eisner, Eric Goldlust, and Noah A. Smith · 2005
Earlier work this paper cites.
Connectionist temporal classification: Labelling unsegmented sequence data with recurrent neural networks
Alex Graves, Santiago Fernández, Faustino Gomez, and Jürgen Schmidhuber · 2006
Earlier work this paper cites.
Theano: A CPU and GPU math compiler in Python
James Bergstra, Olivier Breuleux, Frédéric Bastien, Pascal Lamblin, Razvan Pascanu, Guillaume Desjardins, Joseph Turian, David Warde-Farley, and Yoshua Bengio · 2010
Earlier work this paper cites.
Neural conditional random fields
Trinh-Minh-Tri Do and Thierry Artières · 2010
Earlier work this paper cites.
Eigen v3
Gaël Guennebaud, Benoît Jacob, et al · 2010
Earlier work this paper cites.
Natural language processing (almost) from scratch
Ronan Collobert, Jason Weston, Léon Bottou, Michael Karlen, Koray Kavukcuoglu, and Pavel Kuksa · 2011
Earlier work this paper cites.
Adaptive subgradient methods for online learning and stochastic optimization
John Duchi, Elad Hazan, and Yoram Singer · 2011
Earlier work this paper cites.
Extensions of recurrent neural network language model
Tomáš Mikolov, Stefan Kombrink, Lukáš Burget, Jan Černockỳ, and Sanjeev Khudanpur · 2011
Earlier work this paper cites.
Hogwild: A lock-free approach to parallelizing stochastic gradient descent
Benjamin Recht, Christopher Re, Stephen Wright, and Feng Niu · 2011
Earlier work this paper cites.
Pegasos: Primal estimated sub-gradient solver for SVM
Shai Shalev-Shwartz, Yoram Singer, Nathan Srebro, and Andrew Cotter · 2011
Earlier work this paper cites.
Parsing natural scenes and natural language with recursive neural networks
Richard Socher, Cliff C Lin, Chris Manning, and Andrew Y Ng · 2011
Earlier work this paper cites.
Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups
Geoffrey Hinton, Li Deng, Dong Yu, George E Dahl, Abdel-rahman Mohamed, Navdeep Jaitly, Andrew Senior, Vincent Vanhoucke, Patrick Nguyen, Tara N Sainath, et al · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton · 2012
Earlier work this paper cites.
Learning multilingual named entity recognition from Wikipedia
Joel Nothman, Nicky Ringland, Will Radford, Tara Murphy, and James R. Curran · 2012
Earlier work this paper cites.
LSTM neural networks for language modeling
Martin Sundermeyer, Ralf Schlüter, and Hermann Ney · 2012
Cited alongside, same era.
Statistical parametric speech synthesis using deep neural networks
Heiga Zen, Andrew Senior, and Mike Schuster · 2013
Cited alongside, same era.
Empirical evaluation of gated recurrent neural networks on sequence modeling
Junyoung Chung, Caglar Gulcehre, KyungHyun Cho, and Yoshua Bengio · 2014
Cited alongside, same era.
Fast reverse-mode automatic differentiation using expression templates in C++
Robin J. Hogan · 2014
Cited alongside, same era.
A neural network for factoid question answering over paragraphs
Mohit Iyyer, Jordan Boyd-Graber, Leonardo Claudino, Richard Socher, and Hal Daumé III · 2014
Cited alongside, same era.
Transition-based dependency parsing with heuristic backtracking
Jacob Buckman, Miguel Ballesteros, and Chris Dyer · 2016
Later among the works it cites.
Incorporating structural alignment biases into an attentional neural translation model
Trevor Cohn, Cong Duy Vu Hoang, Ekaterina Vymolova, Kaisheng Yao, Chris Dyer, and Gholamreza Haffari · 2016
Later among the works it cites.
Recurrent neural network grammars
Chris Dyer, Adhiguna Kuncoro, Miguel Ballesteros, and Noah A. Smith · 2016
Later among the works it cites.
Morphological inflection generation using character sequence to sequence learning
Manaal Faruqui, Yulia Tsvetkov, Graham Neubig, and Chris Dyer · 2016
Later among the works it cites.
A neural network for coordination boundary prediction
Jessica Ficler and Yoav Goldberg · 2016
Later among the works it cites.
Semi supervised preposition-sense disambiguation using multilingual data
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Diederik Kingma and Jimmy Ba · 2014
Cited alongside, same era.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc VV Le · 2014
Cited alongside, same era.
An introduction to computational networks and the computational network toolkit
Dong Yu, Adam Eversole, Mike Seltzer, Kaisheng Yao, Oleksii Kuchaiev, Yu Zhang, Frank Seide, Zhiheng Huang, Brian Guenter, Huaming Wang, Jasha Droppo, Geoffrey Zweig, Chris Rossbach, Jie Gao, Andreas Stolcke, Jon Currey, Malcolm Slaney, Guoguo Chen, Amit Agarwal, Chris Basoglu, Marko Padmilac, Alexey Kamenev, Vladimir Ivanov, Scott Cypher, Hari Parthasarathi, Bhaskar Mitra, Baolin Peng, and Xuedong Huang · 2014
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2015
Cited alongside, same era.
Mxnet: A flexible and efficient machine learning library for heterogeneous distributed systems
Tianqi Chen, Mu Li, Yutian Li, Min Lin, Naiyan Wang, Minjie Wang, Tianjun Xiao, Bing Xu, Chiyuan Zhang, and Zheng Zhang · 2015
Cited alongside, same era.
Transition-based dependency parsing with stack long short-term memory
Chris Dyer, Miguel Ballesteros, Wang Ling, Austin Matthews, and Noah A. Smith · 2015
Cited alongside, same era.
Approximation-aware dependency parsing by belief propagation
Matthew R. Gormley, Mark Dredze, and Jason Eisner · 2015
Cited alongside, same era.
Hila Gonen and Yoav Goldberg · 2016
Later among the works it cites.
Easy-first dependency parsing with hierarchical tree LSTMs
Eliyahu Kiperwasser and Yoav Goldberg · 2016
Later among the works it cites.
Simple and accurate dependency parsing using bidirectional LSTM feature representations
Eliyahu Kiperwasser and Yoav Goldberg · 2016
Later among the works it cites.
Improving sentence compression by learning to predict gaze
Sigrid Klerke, Yoav Goldberg, and Anders Søgaard · 2016
Later among the works it cites.
Segmental recurrent neural networks
Lingpeng Kong, Chris Dyer, and Noah A. Smith · 2016
Later among the works it cites.
Neural architectures for named entity recognition
Guillaume Lample, Miguel Ballesteros, Sandeep Subramanian, Kazuya Kawakami, and Chris Dyer · 2016
Later among the works it cites.
Semantic object parsing with graph lstm
Xiaodan Liang, Xiaohui Shen, Jiashi Feng, Liang Lin, and Shuicheng Yan · 2016
Later among the works it cites.
Generalizing and hybridizing count-based and neural language models
Graham Neubig and Chris Dyer · 2016
Later among the works it cites.
Multilingual part-of-speech tagging with bidirectional long short-term memory models and auxiliary loss
Barbara Plank, Anders Søgaard, and Yoav Goldberg · 2016
Later among the works it cites.
Cogalex-v shared task: Lexnet - integrated path-based and distributional method for the identification of semantic relations
Vered Shwartz and Ido Dagan · 2016
Later among the works it cites.
Improving hypernymy detection with an integrated path-based and distributional method
Vered Shwartz, Yoav Goldberg, and Ido Dagan · 2016
Later among the works it cites.
Mastering the game of Go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Later among the works it cites.
Deep multi-task learning with low level tasks supervised at lower layers
Anders Søgaard and Yoav Goldberg · 2016
Later among the works it cites.
Asynchronous parallel learning for neural networks and structured models with dense features
Xu Sun · 2016
Later among the works it cites.
Greedy, joint syntactic-semantic parsing with stack lstms
Swabha Swayamdipta, Miguel Ballesteros, Chris Dyer, and Noah A. Smith · 2016
Later among the works it cites.
Deep learning with dynamic computation graphs
Moshe Looks, Marcello Herreshoff, DeLesley Hutchins, and Peter Norvig · 2017
Closest in time.