Fetching the paper…
Reading the bibliography…
Using unitary (instead of general) matrices in artificial neural networks (ANNs) is a promising way to solve the gradient explosion/vanishing problem, as well as to enable ANNs to learn long-term correlations in the data.
Untersuchungen zu dynamischen neuronalen netzen
Hochreiter, Sepp · 1991
Earlier work this paper cites.
Darpa timit acoustic-phonetic continous speech corpus cd-rom. nist speech disc 1-1.1
Garofolo, John S, Lamel, Lori F, Fisher, William M, Fiscus, Jonathon G, and Pallett, David S · 1993
Earlier work this paper cites.
Learning long-term dependencies with gradient descent is difficult
Bengio, Yoshua, Simard, Patrice, and Frasconi, Paolo · 1994
Earlier work this paper cites.
Experimental realization of any discrete unitary operator
Reck, Michael, Zeilinger, Anton, Bernstein, Herbert J., and Bertani, Philip · 1994
Earlier work this paper cites.
Long short-term memory
Hochreiter, Sepp and Schmidhuber, Jürgen · 1997
Earlier work this paper cites.
Time–frequency feature representation using energy concentration: An overview of recent advances
Sejdić, Ervin, Djurović, Igor, and Jiang, Jin · 2007
Earlier work this paper cites.
Natural language processing (almost) from scratch
Collobert, Ronan, Weston, Jason, Bottou, Léon, Karlen, Michael, Kavukcuoglu, Koray, and Kuksa, Pavel · 2011
Earlier work this paper cites.
Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups
Hinton, Geoffrey, Deng, Li, Yu, Dong, Dahl, George E, Mohamed, Abdel-rahman, Jaitly, Navdeep, Senior, Andrew, Vanhoucke, Vincent, Nguyen, Patrick, Sainath, Tara N, et al · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Krizhevsky, Alex, Sutskever, Ilya, and Hinton, Geoffrey E · 2012
Cited alongside, same era.
Exact solutions to the nonlinear dynamics of learning in deep linear neural networks
Saxe, Andrew M, McClelland, James L, and Ganguli, Surya · 2013
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Bahdanau, Dzmitry, Cho, Kyunghyun, and Bengio, Yoshua · 2014
Cited alongside, same era.
On the properties of neural machine translation: Encoder-decoder approaches
Cho, Kyunghyun, Van Merriënboer, Bart, Bahdanau, Dzmitry, and Bengio, Yoshua · 2014
Cited alongside, same era.
Sequence to sequence learning with neural networks
Sutskever, Ilya, Vinyals, Oriol, and Le, Quoc V · 2014
Cited alongside, same era.
Long-term recurrent convolutional networks for visual recognition and description
Donahue, Jeffrey, Anne Hendricks, Lisa, Guadarrama, Sergio, Rohrbach, Marcus, Venugopalan, Subhashini, Saenko, Kate, and Darrell, Trevor · 2015
Later among the works it cites.
A simple way to initialize recurrent networks of rectified linear units
Le, Quoc V, Jaitly, Navdeep, and Hinton, Geoffrey E · 2015
Later among the works it cites.
Deep learning
LeCun, Yann, Bengio, Yoshua, and Hinton, Geoffrey · 2015
Later among the works it cites.
An optimal design for universal multiport interferometers, 2016
Clements, William R., Humphreys, Peter C., Metcalf, Benjamin J., Kolthammer, W. Steven, and Walmsley, Ian A · 2016
Closest in time.
Orthogonal rnns and long-memory tasks
Henaff, Mikael, Szlam, Arthur, and LeCun, Yann · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Unitary evolution recurrent neural networks
Arjovsky, Martin, Shah, Amar, and Bengio, Yoshua · 2015
Cited alongside, same era.
Bidirectional recurrent neural networks as generative models
Berglund, Mathias, Raiko, Tapani, Honkala, Mikko, Kärkkäinen, Leo, Vetek, Akos, and Karhunen, Juha T · 2015
Cited alongside, same era.
Fast approximation of rotations and hessians matrices
Mathieu, Michael and LeCun, Yann
Cited in the paper.
Fast approximation of rotations and hessians matrices
Mathieu, Michal and LeCun, Yann
Cited in the paper.
Closest in time.
Deep learning with coherent nanophotonic circuits
Shen, Yichen, Harris, Nicholas C, Skirlo, Scott, Prabhu, Mihika, Baehr-Jones, Tom, Hochberg, Michael, Sun, Xin, Zhao, Shijie, Larochelle, Hugo, Englund, Dirk, et al · 2016
Closest in time.
Full-capacity unitary recurrent neural networks
Wisdom, Scott, Powers, Thomas, Hershey, John, Le Roux, Jonathan, and Atlas, Les · 2016
Closest in time.