Fetching the paper…
Reading the bibliography…
In this paper, we propose a novel deep neural network architecture, Speech2Vec, for learning fixed-length vector representations of audio segments excised from a speech corpus, where the vectors contain semantic information pertaining to the underlying spoken words, and are close to other vectors in the embedding space if their corresponding underlying spoken words are semantically similar.
H. Rubenstein and J. B. Goodenough, “Contextual correlates of synonymy,”
1965
Earlier work this paper cites.
G. A. Miller and W. G. Charles, “Contextual correlates of semantic similarity,”
1991
Earlier work this paper cites.
J. L. Myers and A. D. Well,
1995
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,”
1997
Earlier work this paper cites.
D. Yang and D. M. Powers, “Verb similarity on the taxonomy of wordnet,” in
2006
Earlier work this paper cites.
L. v. d. Maaten and G. Hinton, “Visualizing data using t-sne,”
2008
Earlier work this paper cites.
E. Agirre, E. Alfonseca, K. Hall, J. Kravalova, M. Paşca, and A. Soroa, “A study on similarity and relatedness using distributional and wordnet-based approaches,” in
2009
Earlier work this paper cites.
K. Radinsky, E. Agichtein, E. Gabrilovich, and S. Markovitch, “A word at a time: computing word relatedness using temporal semantic analysis,” in
2011
Earlier work this paper cites.
E. Bruni, G. Boleda, M. Baroni, and N.-K. Tran, “Distributional semantics in technicolor,” in
2012
Earlier work this paper cites.
G. Halawi, G. Dror, E. Gabrilovich, and Y. Koren, “Large-scale learning of word relatedness with constraints,” in
2012
Earlier work this paper cites.
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean, “Distributed representations of words and phrases and their compositionality,” in
2013
Earlier work this paper cites.
K. Levin, K. Henry, A. Jansen, and K. Livescu, “Fixed-dimensional acoustic embeddings of variable-length segments in low-resource settings,” in
2013
Earlier work this paper cites.
M.-T. Luong, R. Socher, and C. D. Manning, “Better word representations with recursive neural networks for morphology,” in
2013
Earlier work this paper cites.
J. Pennington, R. Socher, and C. Manning, “Glove: Global vectors for word representation,” in
2014
Earlier work this paper cites.
S. Bengio and G. Heigold, “Word embeddings for speech recognition,” in
2014
Earlier work this paper cites.
I. Sutskever, O. Vinyals, and Q. V. Le, “Sequence to sequence learning with neural networks,” in
2014
Earlier work this paper cites.
K. Cho, B. van Merriënboer, Ç. Gülçehre, D. Bahdanau, F. Bougares, H. Schwenk, and Y. Bengio, “Learning phrase representations using rnn encoder–decoder for statistical machine translation,” in
2014
Cited alongside, same era.
D. Bahdanau, K. Cho, and Y. Bengio, “Neural machine translation by jointly learning to align and translate,” in
2014
Cited alongside, same era.
M. Faruqui and C. Dyer, “Community evaluation and exchange of word vectors at wordvectors.org,” in
2014
Cited alongside, same era.
S. Baker, R. Reichart, and A. Korhonen, “An unsupervised model for instance level subcategorization acquisition,” in
2014
Cited alongside, same era.
M. Ballesteros, C. Dyer, and N. A. Smith, “Improved transition-based parsing by modeling characters instead of words with LSTMs,” in
2015
Cited alongside, same era.
H. Kamper, W. Wang, and K. Livescu, “Deep convolutional acoustic word embeddings using word-pair side information,” in
2016
Later among the works it cites.
D. Harwath, A. Torralba, and J. Glass, “Unsupervised learning of spoken language with visual context,” in
2016
Later among the works it cites.
D. Gerz, I. Vulić, F. Hill, R. Reichart, and A. Korhonen, “Simverb-3500: A large-scale evaluation set of verb similarity,” in
2016
Later among the works it cites.
B.-H. Tseng, S.-S. Shen, H.-Y. Lee, and L.-S. Lee, “Towards machine comprehension of spoken content: Initial TOEFL listening comprehension test by machine,” in
2016
Later among the works it cites.
P. Bojanowski, E. Grave, A. Joulin, and T. Mikolov, “Enriching word vectors with subword information,”
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. Harwath and J. Glass, “Deep multimodal semantic embeddings for speech and images,” in
2015
Cited alongside, same era.
T. Luong, H. Pham, and C. D. Manning, “Effective approaches to attention-based neural machine translation,” in
2015
Cited alongside, same era.
S. Venugopalan, M. Rohrbach, J. Donahue, R. Mooney, T. Darrell, and K. Saenko, “Sequence to sequence-video to text,” in
2015
Cited alongside, same era.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: an ASR corpus based on public domain audio books,” in
2015
Cited alongside, same era.
T. Schnabel, I. Labutov, D. Mimno, and T. Joachims, “Evaluation methods for unsupervised word embeddings,” in
2015
Cited alongside, same era.
F. Hill, R. Reichart, and A. Korhonen, “Simlex-999: Evaluating semantic models with (genuine) similarity estimation,”
2015
Cited alongside, same era.
G. Lample, M. Ballesteros, S. Subramanian, K. Kawakami, and C. Dyer, “Neural architectures for named entity recognition,” in
2016
Cited alongside, same era.
X. Yu and N. T. Vu, “Character composition model with convolutional neural networks for dependency parsing on morphologically rich languages,” in
2017
Later among the works it cites.
W. He, W. Wang, and K. Livescu, “Multi-view recurrent neural acoustic word embeddings,” in
2017
Later among the works it cites.
D. Harwath and J. Glass, “Learning word-like units from joint audio-visual analysis,” in
2017
Later among the works it cites.
Y.-A. Chung and J. Glass, “Learning word embeddings from speech,” in
2017
Later among the works it cites.
I. Konstas, S. Iyer, M. Yatskar, Y. Choi, and L. Zettlemoyer, “Neural amr: Sequence-to-sequence models for parsing and generation,” in
2017
Later among the works it cites.
A. Paszke, S. Gross, S. Chintala, G. Chanan, E. Yang, Z. DeVito, Z. Lin, A. Desmaison, L. Antiga, and A. Lerer, “Automatic differentiation in pytorch,” in
2017
Later among the works it cites.
2017
Later among the works it cites.
H. Kamper, K. Livescu, and S. Goldwater, “An embedded segmental k-means model for unsupervised segmentation and clustering of speech,” in
2017
Later among the works it cites.
H. Kamper, A. Jansen, and S. Goldwater, “A segmental framework for fully-unsupervised large-vocabulary speech recognition,”
2017
Later among the works it cites.
S. Subramanian, A. Trischler, Y. Bengio, and C. Pal, “Learning general purpose distributed sentence representations via large scale multi-task learning,” in
2018
Closest in time.
Y.-A. Chung, H.-Y. Lee, and J. Glass, “Supervised and unsupervised transfer learning for question answering,” in
2018
Closest in time.