Fetching the paper…
Reading the bibliography…
The recently proposed Unbiased Online Recurrent Optimization algorithm (UORO, arXiv:1702.05043) uses an unbiased approximation of RTRL to achieve fully online gradient-based learning in RNNs.
Steps toward artificial intelligence
Marvin Minsky · 1961
Earlier work this paper cites.
Neuronlike adaptive elements that can solve difficult learning control problems
Andrew G Barto, Richard S Sutton, and Charles W Anderson · 1983
Earlier work this paper cites.
Temporal credit assignment in reinforcement learning
Richard S Sutton · 1984
Earlier work this paper cites.
Learning representations by back-propagating errors
David E Rumelhart, Geoffrey E Hinton, and Ronald J Williams · 1986
Earlier work this paper cites.
Learning to predict by the methods of temporal differences
Richard S Sutton · 1988
Earlier work this paper cites.
A learning algorithm for continually running fully recurrent neural networks
Ronald J Williams and David Zipser · 1989
Earlier work this paper cites.
Finding structure in time
Jeffrey L Elman · 1990
Earlier work this paper cites.
Backpropagation through time: what it does and how to do it
Paul J Werbos · 1990
Earlier work this paper cites.
An efficient gradient-based algorithm for on-line training of recurrent network trajectories
Ronald J Williams and Jing Peng · 1990
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams · 1992
Earlier work this paper cites.
Gradient-based learning algorithms for recurrent networks and their computational complexity
Ronald J Williams and David Zipser · 1995
Cited alongside, same era.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Cited alongside, same era.
On the improvement of the real time recurrent learning algorithm for recurrent neural networks
Man-Wai Mak, Kim-Wing Ku, and Yee-Ling Lu · 1999
Cited alongside, same era.
Actor-critic algorithms
Vijay R Konda and John N Tsitsiklis · 2000
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Cited alongside, same era.
A simple way to initialize recurrent networks of rectified linear units
The reversible residual network: Backpropagation without storing activations
Aidan N Gomez, Mengye Ren, Raquel Urtasun, and Roger B Grosse · 2017
Later among the works it cites.
Decoupled neural interfaces using synthetic gradients
Max Jaderberg, Wojciech Marian Czarnecki, Simon Osindero, Oriol Vinyals, Alex Graves, David Silver, and Koray Kavukcuoglu · 2017
Later among the works it cites.
Around the Use of Gradients in Machine Learning
Pierre-Yves Massé · 2017
Later among the works it cites.
Rudder: Return decomposition for delayed rewards
Jose A Arjona-Medina, Michael Gillhofer, Michael Widrich, Thomas Unterthiner, and Sepp Hochreiter · 2018
Later among the works it cites.
Scientific computing: an introductory survey , volume 80
Michael T Heath · 2018
Later among the works it cites.
Optimizing agent behavior over long time scales by transporting value
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Quoc V Le, Navdeep Jaitly, and Geoffrey E Hinton · 2015
Cited alongside, same era.
Training recurrent networks online without backtracking
Yann Ollivier, Corentin Tallec, and Guillaume Charpiat · 2015
Cited alongside, same era.
Training deep nets with sublinear memory cost
Tianqi Chen, Bing Xu, Chiyuan Zhang, and Carlos Guestrin · 2016
Cited alongside, same era.
Memory-efficient backpropagation through time
Audrunas Gruslys, Rémi Munos, Ivo Danihelka, Marc Lanctot, and Alex Graves · 2016
Cited alongside, same era.
A review of matrix scaling and sinkhorn’s normal form for matrices and positive maps
Martin Idel · 2016
Cited alongside, same era.
Chia-Chun Hung, Timothy Lillicrap, Josh Abramson, Yan Wu, Mehdi Mirza, Federico Carnevale, Arun Ahuja, and Greg Wayne · 2018
Later among the works it cites.
Sparse attentive backtracking: Temporal credit assignment through reminding
Nan Rosemary Ke, Anirudh Goyal, Olexa Bilaniuk, Jonathan Binas, Michael C Mozer, Chris Pal, and Yoshua Bengio · 2018
Later among the works it cites.
Reversible recurrent neural networks
Matthew MacKay, Paul Vicol, Jimmy Ba, and Roger B Grosse · 2018
Later among the works it cites.
Approximating real-time recurrent learning with random kronecker factors
Asier Mujika, Florian Meier, and Angelika Steger · 2018
Later among the works it cites.
Unbiased online recurrent optimization
Corentin Tallec and Yann Ollivier · 2018
Later among the works it cites.