Fetching the paper…
Reading the bibliography…
Layer normalization is a recently introduced technique for normalizing the activities of neurons in deep neural networks to improve the training speed and stability.
D. B. Paul and J. M. Baker, “The design for the wall street journal-based csr corpus,” in
1992
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,”
1997
Earlier work this paper cites.
M. Schuster and K. K. Paliwal, “Bidirectional recurrent neural networks,”
1997
Earlier work this paper cites.
A. Graves, S. Fernández, and J. Schmidhuber,
2005
Earlier work this paper cites.
L. van der Maaten and G. E. Hinton, “Visualizing high-dimensional data using t-sne,”
2008
Earlier work this paper cites.
N. Dehak, P. Kenny, R. Dehak, P. Dumouchel, and P. Ouellet, “Front-end factor analysis for speaker verification,”
2010
Earlier work this paper cites.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz
2011
Earlier work this paper cites.
G. Hinton, L. Deng, D. Yu, G. E. Dahl, A.-r. Mohamed, N. Jaitly, A. Senior, V. Vanhoucke, P. Nguyen, T. N. Sainath
2012
Earlier work this paper cites.
G. Saon, H. Soltau, D. Nahamoo, and M. Picheny, “Speaker adaptation of neural network acoustic models using i-vectors.” in
2013
Earlier work this paper cites.
H. Liao, “Speaker adaptation of context dependent deep neural networks,” in
2013
Earlier work this paper cites.
A. Graves, N. Jaitly, and A.-r. Mohamed, “Hybrid speech recognition with deep bidirectional lstm,” in
2013
Cited alongside, same era.
H. Sak, A. W. Senior, and F. Beaufays, “Long short-term memory recurrent neural network architectures for large scale acoustic modeling.” in
2014
Cited alongside, same era.
V. Gupta, P. Kenny, P. Ouellet, and T. Stafylakis, “I-vector-based speaker adaptation of deep neural networks for french broadcast audio transcription,” in
2014
Cited alongside, same era.
P. Swietojanski and S. Renals, “Learning hidden unit contributions for unsupervised speaker adaptation of neural network acoustic models,” in
2014
Cited alongside, same era.
A. Rousseau, P. Deléglise, and Y. Estève, “Enhancing the ted-lium corpus with selected data for language modeling and more ted talks.” in
2014
Cited alongside, same era.
P. Swietojanski, J. Li, and S. Renals, “Learning hidden unit contributions for unsupervised acoustic model adaptation,”
2016
Later among the works it cites.
L. J. Ba, R. Kiros, and G. E. Hinton, “Layer normalization,”
2016
Later among the works it cites.
2016
Later among the works it cites.
2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,”
2014
Cited alongside, same era.
Y. Miao, H. Zhang, and F. Metze, “Speaker adaptive training of deep neural network acoustic models using i-vectors,”
2015
Cited alongside, same era.
S. Ioffe and C. Szegedy, “Batch normalization: Accelerating deep network training by reducing internal covariate shift.” in
2015
Cited alongside, same era.
S. Dieleman, J. Schlüter, C. Raffel, E. Olson, S. K. Sønderby, D. Nouri, D. Maturana, M. Thoma, E. Battenberg, J. Kelly, J. D. Fauw, M. Heilman, D. M. de Almeida, B. McFee, H. Weideman, G. Takács, P. de Rivaz, J. Crall, G. Sanders, K. Rasul, C. Liu, G. French, and J. Degrave, “Lasagne: First release.” Aug. 2015. [Online]. Available:
2015
Cited alongside, same era.
T. Tan, Y. Qian, D. Yu, S. Kundu, L. Lu, K. C. Sim, X. Xiao, and Y. Zhang, “Speaker-aware training of lstm-rnns for acoustic modelling,” in
2016
Cited alongside, same era.
2016
Later among the works it cites.
V. Dumoulin, J. Shlens, and M. Kudlur, “A learned representation for artistic style,”
2016
Later among the works it cites.
D. Ha, A. Dai, and Q. V. Le, “Hypernetworks,”
2016
Later among the works it cites.
M. Henaff, A. Szlam, and Y. LeCun, “Orthogonal rnns and long-memory tasks,”
2016
Later among the works it cites.
2016
Later among the works it cites.