Fetching the paper…
Reading the bibliography…
In this paper, we revisit the recurrent back-propagation (RBP) algorithm, discuss the conditions under which it applies as well as how to satisfy them in deep neural networks.
Revisiting semi-supervised learning with graph embeddings
Yang, Z., Cohen, W. W., and Salakhutdinov, R · 1930
Earlier work this paper cites.
Forces in molecules
Feynman, R. P · 1939
Earlier work this paper cites.
Methods of conjugate gradients for solving linear systems , volume 49
Hestenes, M. R. and Stiefel, E · 1952
Earlier work this paper cites.
Principles of mathematical analysis , volume 3
Rudin, W · 1964
Earlier work this paper cites.
Neural networks and physical systems with emergent collective computational abilities
Hopfield, J. J · 1982
Earlier work this paper cites.
Lsqr: An algorithm for sparse linear equations and sparse least squares
Paige, C. C. and Saunders, M. A · 1982
Earlier work this paper cites.
Neurons with graded response have collective computational properties like those of two-state neurons
Hopfield, J. J · 1984
Earlier work this paper cites.
A self-optimizing, nonsymmetrical neural net for content addressable memory and pattern recognition
Lapedes, A. and Farber, R · 1986
Earlier work this paper cites.
A learning rule for asynchronous perceptrons with feedback in a combinatorial environment
Almeida, L. B · 1987
Earlier work this paper cites.
Generalization of back-propagation to recurrent neural networks
Pineda, F. J · 1987
Earlier work this paper cites.
Evaluating the matrix polynomial i+a+. . .+a n-1
Westreich, D · 1989
Earlier work this paper cites.
Backpropagation through time: what it does and how to do it
Werbos, P. J · 1990
Earlier work this paper cites.
An efficient gradient-based algorithm for on-line training of recurrent network trajectories
Williams, R. J. and Peng, J · 1990
Earlier work this paper cites.
Neural Networks and Learning Machines
Haykin, S · 1993
Cited alongside, same era.
Fast exact multiplication by the hessian
Pearlmutter, B. A · 1994
Cited alongside, same era.
Backpropagation: theory, architectures, and applications
Chauvin, Y. and Rumelhart, D. E · 1995
Cited alongside, same era.
Long short-term memory
Hochreiter, S. and Schmidhuber, J · 1997
Cited alongside, same era.
Fast curvature matrix-vector products for second-order gradient descent
Schraudolph, N. N · 2002
Cited alongside, same era.
Efficient multiple hyperparameter learning for log-linear models
Foo, C.-s., Do, C. B., and Ng, A. Y · 2008
Cited alongside, same era.
Visualizing data using t-sne
Learning phrase representations using rnn encoder-decoder for statistical machine translation
Cho, K., Van Merriënboer, B., Gulcehre, C., Bahdanau, D., Bougares, F., Schwenk, H., and Bengio, Y · 2014
Later among the works it cites.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J · 2014
Later among the works it cites.
Sequence to sequence learning with neural networks
Sutskever, I., Vinyals, O., and Le, Q. V · 2014
Later among the works it cites.
Gradient-based hyperparameter optimization through reversible learning
Maclaurin, D., Duvenaud, D., and Adams, R · 2015
Later among the works it cites.
Training recurrent networks online without backtracking
Ollivier, Y., Tallec, C., and Charpiat, G · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Maaten, L. v. d. and Hinton, G · 2008
Cited alongside, same era.
Learning optimized map estimates in continuously-valued mrf models
Samuel, K. G. and Tappen, M. F · 2009
Cited alongside, same era.
The graph neural network model
Scarselli, F., Gori, M., Tsoi, A. C., Hagenbuchner, M., and Monfardini, G · 2009
Cited alongside, same era.
Generic methods for optimization-based modeling
Domke, J · 2012
Cited alongside, same era.
Matrix computations , volume 3
Golub, G. H. and Van Loan, C. F · 2012
Cited alongside, same era.
Practical bayesian optimization of machine learning algorithms
Snoek, J., Larochelle, H., and Adams, R. P · 2012
Cited alongside, same era.
Deep learning , volume 1
Goodfellow, I., Bengio, Y., Courville, A., and Bengio, Y · 2016
Later among the works it cites.
Gated graph sequence neural networks
Li, Y., Tarlow, D., Brockschmidt, M., and Zemel, R · 2016
Later among the works it cites.
Gated graph sequence neural networks
Li, Y., Tarlow, D., Brockschmidt, M., and Zemel, R · 2016
Later among the works it cites.
Optnet: Differentiable optimization as a layer in neural networks
Amos, B. and Kolter, J. Z · 2017
Later among the works it cites.
Unbiasing truncated backpropagation through time
Tallec, C. and Ollivier, Y · 2017
Later among the works it cites.
On the computation of neumann series
Vassil, S. D. and Diego, F. G. C · 2017
Later among the works it cites.
Unbiasing truncated backpropagation through time
Tallec, C. and Ollivier, Y · 2017
Later among the works it cites.
Graph partition neural networks for semi-supervised classification
Liao, R., Brockschmidt, M., Tarlow, D., Gaunt, A., Urtasun, R., and Zemel, R · 2018
Closest in time.