Fetching the paper…
Reading the bibliography…
Learning good representations is of crucial importance in deep learning.
L. Bahl, P. Brown, P. de Souza, and R. Mercer, “Maximum mutual information estimation of hidden markov model parameters for speech recognition,” in
1986
Earlier work this paper cites.
J. S. Garofolo, L. F. Lamel, W. M. Fisher, J. G. Fiscus, D. S. Pallett, and N. L. Dahlgren, “DARPA TIMIT Acoustic Phonetic Continuous Speech Corpus CDROM,” 1993
1993
Earlier work this paper cites.
L. Paninski, “Estimation of entropy and mutual information,”
2003
Earlier work this paper cites.
G. Hinton, S. Osindero, and Y. Teh, “A fast learning algorithm for deep belief nets,” vol. 18, 2006, pp. 1527–1554
2006
Earlier work this paper cites.
Y. Bengio, P. L., D. Popovici, and H. Larochelle, “Greedy layer-wise training of deep networks,” in
2007
Earlier work this paper cites.
D. Applebaum,
2008
Earlier work this paper cites.
G. Dahl, D. Yu, L. Deng, and A. Acero, “Context-dependent pre-trained deep neural networks for large vocabulary speech recognition,”
2012
Earlier work this paper cites.
M. Ravanelli, A. Sosi, P. Svaizer, and M. Omologo, “Impulse response estimation for robust speech recognition in a reverberant environment,” in
2012
Earlier work this paper cites.
A. K. Sarkar, D. Matrouf, P. Bousquet, and J. Bonastre, “Study of the effect of i-vector modeling on short and mismatch utterance duration for speaker verification,” in
2012
Earlier work this paper cites.
D. P. Kingma and M. Welling, “Auto-Encoding Variational Bayes,”
2013
Earlier work this paper cites.
A. L. Maas, A. Y. Hannun, and A. Y. Ng, “Rectifier nonlinearities improve neural network acoustic models,” in
2013
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” in
2014
Earlier work this paper cites.
J. B. Kinney and G. S. Atwal, “Equitability, mutual information, and the maximal information coefficient,”
2014
Earlier work this paper cites.
Z. Tüske, P. Golik, R. Schlüter, and H. Ney, “Acoustic modeling with deep neural networks using raw time signal for LVCSR,” in
2014
Earlier work this paper cites.
E. Variani, X. Lei, E. McDermott, I. L. Moreno, and J. Gonzalez-Dominguez, “Deep neural networks for small footprint text-dependent speaker verification,” in
2014
Earlier work this paper cites.
D. Yu and L. Deng,
2015
Cited alongside, same era.
M. McLaren, Y. Lei, and L. Ferrer, “Advances in deep neural network approaches to speaker recognition,” in
2015
Cited alongside, same era.
F. Richardson, D. Reynolds, and N. Dehak, “Deep neural network approaches to speaker and language recognition,”
2015
Cited alongside, same era.
D. Palaz, M. Magimai-Doss, and R. Collobert, “Analysis of CNN-based speech recognition system using raw speech as input,” in
2015
Cited alongside, same era.
T. N. Sainath, R. J. Weiss, A. W. Senior, K. W. Wilson, and O. Vinyals, “Learning the speech front-end with raw waveform CLDNNs,” in
2015
Cited alongside, same era.
M. Ravanelli, P. Brakel, M. Omologo, and Y. Bengio, “A network of deep neural networks for distant speech recognition,” in
2017
Later among the works it cites.
P. Brakel and Y. Bengio, “Learning independent features with adversarial nets for non-linear ica,”
2017
Later among the works it cites.
2017
Later among the works it cites.
A. Nagrani, J. S. Chung, and A. Zisserman, “Voxceleb: a large-scale speaker identification dataset,” in
2017
Later among the works it cites.
M. I. Belghazi, A. Baratin, S. Rajeshwar, S. Ozair, Y. Bengio, A. Courville, and R. D. Hjelm, “Mutual information neural estimation,” in
2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
M. Ravanelli, L. Cristoforetti, R. Gretter, M. Pellin, A. Sosi, and M. Omologo, “The DIRHA-ENGLISH corpus and related tasks for distant-speech recognition in domestic environments,” in
2015
Cited alongside, same era.
S. Ioffe and C. Szegedy, “Batch normalization: Accelerating deep network training by reducing internal covariate shift,” in
2015
Cited alongside, same era.
I. Goodfellow, Y. Bengio, and A. Courville,
2016
Cited alongside, same era.
Y. Ganin, E. Ustinova, H. Ajakan, P. Germain, H. Larochelle, F. Laviolette, M. Marchand, and V. Lempitsky, “Domain-adversarial training of neural networks,”
2016
Cited alongside, same era.
2016
Cited alongside, same era.
G. Trigeorgis, F. Ringeval, R. Brueckner, E. Marchi, M. A. Nicolaou, B. Schuller, and S. Zafeiriou, “Adieu features? end-to-end speech emotion recognition using a deep convolutional recurrent network,” in
2016
Cited alongside, same era.
2018
Closest in time.
R. D. Hjelm, A. Fedorov, S. Lavoie-Marchildon, K. Grewal, A. Trischler, and Y. Bengio, “Learning deep representations by mutual information estimation and maximization,”
2018
Closest in time.
M. Ravanelli and Y. Bengio, “Speaker Recognition from raw waveform with SincNet,” in
2018
Closest in time.
M. Ravanelli and Y.Bengio, “Interpretable Convolutional Filters with SincNet,” in
2018
Closest in time.
P. Velickovic, W. Fedus, W. L. Hamilton, P. Liò, Y. Bengio, and R. D. Hjelm, “Deep graph infomax,”
2018
Closest in time.
H. Muckenhirn, M. Magimai-Doss, and S. Marcel, “Towards directly modeling raw speech signal for speaker verification using CNNs,” in
2018
Closest in time.
2018
Closest in time.
N. Le and J. Odobez, “Robust and discriminative speaker embedding via intra-class distance variance regularization,” in
2018
Closest in time.
M. Ravanelli, T. Parcollet, and Y. Bengio, “The PyTorch-Kaldi Speech Recognition Toolkit,” in
2019
Closest in time.