Fetching the paper…
Reading the bibliography…
The vector representations of fixed dimensionality for words (in text) offered by Word2Vec have been shown to be very useful in many application scenarios, in particular due to the semantic information they carry.
H. Sakoe and S. Chiba, “Dynamic programming algorithm optimization for spoken word recognition,”
1978
Earlier work this paper cites.
Y. Bengio, P. Simard, and P. Frasconi, “Learning long-term dependencies with gradient descent is difficult,”
1994
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,”
1997
Earlier work this paper cites.
F. Gers, J. Schmidhuber
2000
Earlier work this paper cites.
G. E. Hinton and R. R. Salakhutdinov, “Reducing the dimensionality of data with neural networks,”
2006
Earlier work this paper cites.
J. Schmidhuber, D. Wierstra, M. Gagliolo, and F. Gomez, “Training recurrent networks by evolino,”
2007
Earlier work this paper cites.
P. Vincent, H. Larochelle, Y. Bengio, and P.-A. Manzagol, “Extracting and composing robust features with denoising autoencoders,” in
2008
Earlier work this paper cites.
C. D. Manning, P. Raghavan, and H. Schütze,
2008
Earlier work this paper cites.
N. Dehak, R. Dehak, P. Kenny, N. Brummer, P. Ouellet, and P. Dumouchel, “Support vector machines versus fast scoring in the low-dimensional total variability space for speaker verification,” in
2009
Earlier work this paper cites.
B. Schuller, S. Steidl, and A. Batliner, “The INTERSPEECH 2009 emotion challenge,” in
2009
Earlier work this paper cites.
J. Bayer, D. Wierstra, J. Togelius, and J. Schmidhuber, “Evolving memory cell structures for sequence learning,” in
2009
Earlier work this paper cites.
P. Vincent, H. Larochelle, I. Lajoie, Y. Bengio, and P.-A. Manzagol, “Stacked denoising autoencoders: Learning useful representations in a deep network with a local denoising criterion,”
2010
Earlier work this paper cites.
J. Bergstra, O. Breuleux, F. Bastien, P. Lamblin, R. Pascanu, G. Desjardins, J. Turian, D. Warde-Farley, and Y. Bengio, “Theano: a CPU and GPU math expression compiler,” in
2010
Earlier work this paper cites.
A. Norouzian, A. Jansen, R. Rose, and S. Thomas, “Exploiting discriminative point process models for spoken term detection,” in
2012
Cited alongside, same era.
P. Baldi, “Autoencoders, unsupervised learning, and deep architectures,”
2012
Cited alongside, same era.
2012
Cited alongside, same era.
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean, “Distributed representations of words and phrases and their compositionality,” in
2013
Cited alongside, same era.
2013
Cited alongside, same era.
H. Sak, A. Senior, and F. Beaufays, “Long short-term memory recurrent neural network architectures for large scale acoustic modeling,” in
2014
Later among the works it cites.
P. Doetsch, M. Kozielski, and H. Ney, “Fast and robust training of recurrent neural networks for offline handwriting recognition,” in
2014
Later among the works it cites.
2014
Later among the works it cites.
C.-T. Chung, C.-A. Chan, and L.-S. Lee, “Unsupervised spoken term detection with spoken queries by multi-level acoustic patterns with varying model granularity,” in
2014
Later among the works it cites.
K. Levin, A. Jansen, and B. Van Durme, “Segmental acoustic indexing for zero resource keyword search,” in
2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
H.-Y. Lee and L.-S. Lee, “Enhanced spoken term detection using support vector machines and weighted pseudo examples,”
2013
Cited alongside, same era.
I.-F. Chen and C.-H. Lee, “A hybrid HMM/DNN approach to keyword spotting of short words,” in
2013
Cited alongside, same era.
K. Levin, K. Henry, A. Jansen, and K. Livescu, “Fixed-dimensional acoustic embeddings of variable-length segments in low-resource settings,” in
2013
Cited alongside, same era.
A. Graves, “Generating sequences with recurrent neural networks,”
2013
Cited alongside, same era.
Q. V. Le and T. Mikolov, “Distributed representations of sentences and documents,”
2014
Cited alongside, same era.
S. Bengio and G. Heigold, “Word embeddings for speech recognition,” in
2014
Cited alongside, same era.
I. Sutskever, O. Vinyals, and Q. V. Le, “Sequence to sequence learning with neural networks,” in
2014
Cited alongside, same era.
Later among the works it cites.
G. Chen, C. Parada, and T. N. Sainath, “Query-by-example keyword spotting using long short-term memory networks,” in
2015
Later among the works it cites.
J. Li, M.-T. Luong, and D. Jurafsky, “A hierarchical neural autoencoder for paragraphs and documents,” in
2015
Later among the works it cites.
R. Kiros, Y. Zhu, R. Salakhutdinov, R. S. Zemel, A. Torralba, R. Urtasun, and S. Fidler, “Skip-thought vectors,” in
2015
Later among the works it cites.
2015
Later among the works it cites.
2015
Later among the works it cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: an ASR corpus based on public domain audio books,” in
2015
Later among the works it cites.
H. Kamper, W. Wang, and K. Livescu, “Deep convolutional acoustic word embeddings using word-pair side information,” in
2016
Closest in time.