Fetching the paper…
Reading the bibliography…
We propose a new unsupervised model for mapping a variable-duration speech segment to a fixed-dimensional representation.
Signature verification using a "Siamese" time delay neural network
J. Bromley, J. W. Bentz, L. Bottou, I. Guyon, Y. LeCun, C. Moore, E. Säckinger, and R. Shah · 1993
Earlier work this paper cites.
The Buckeye corpus of conversational speech: labeling conventions and a test of transcriber reliability
M. Pitt, K. Johnson, E. Hume, S. F. Kiesling, and W. Raymond · 2005
Earlier work this paper cites.
Improved acoustic word embeddings for zero-resource languages using multilingual transfer
H. Kamper, Y. Matusevych, and S. Goldwater · 2006
Earlier work this paper cites.
Unsupervised pattern discovery in speech
A. S. Park and J. R. Glass · 2007
Earlier work this paper cites.
Towards spoken term discovery at scale with zero resources
A. Jansen, K. Church, and H. Hermansky · 2010
Earlier work this paper cites.
Rapid evaluation of speech representations for spoken term discovery
M. A. Carlin, S. Thomas, A. Jansen, and H. Hermansky · 2011
Earlier work this paper cites.
Efficient spoken term discovery using randomized algorithms
A. Jansen and B. Van Durme · 2011
Earlier work this paper cites.
Fixed-dimensional acoustic embeddings of variable-length segments in low-resource settings
K. Levin, K. Henry, A. Jansen, and K. Livescu · 2013
Earlier work this paper cites.
The NCHLT speech corpus of the South African languages
E. Barnard, M. Davel, C. J. van Heerden, F. Wet, and J. Badenhorst · 2014
Earlier work this paper cites.
Word embeddings for speech recognition
S. Bengio and G. Heigold · 2014
Earlier work this paper cites.
Automatic speech recognition for under-resourced languages: A survey
L. Besacier, E. Barnard, A. Karpov, and T. Schultz · 2014
Earlier work this paper cites.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
K. Cho, B. van Merrienboer, C. Gülçehre, D. Bahdanau, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Earlier work this paper cites.
Auto encoding variational Bayes
D. P. Kingma and M. Welling · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2015
Cited alongside, same era.
Generating sentences from a continuous space
S. R. Bowman, L. Vilnis, O. Vinyals, A. M. Dai, R. Józefowicz, and S. Bengio · 2016
Cited alongside, same era.
Unsupervised learning of audio segment representations using sequence-to-sequence recurrent neural networks
Y.-A. Chung, C.-C. Wu, C.-H. Shen, H.-Y. Lee, and L. Lee · 2016
Cited alongside, same era.
Discriminative acoustic word embeddings: Recurrent neural network-based approaches
S. Settle and K. Livescu · 2016
Cited alongside, same era.
Multi-view recurrent neural acoustic word embeddings
W. He, W. Wang, and K. Livescu · 2017
Cited alongside, same era.
beta-VAE: Learning basic visual concepts with a constrained variational framework
I. Higgins, L. Matthey, A. Pal, C. Burgess, X. Glorot, M. Botvinick, S. Mohamed, and A. Lerchner · 2017
Learning acoustic word embeddings with temporal context for query-by-example speech search
Y. Yuan, C.-C. Leung, L. Xie, H. Chen, B. Ma, and H. Li · 2018
Later among the works it cites.
Unsupervised speech representation learning using WaveNet autoencoders
J. Chorowski, R. J. Weiss, S. Bengio, and A. van den Oord · 2019
Later among the works it cites.
Lagging inference networks and posterior collapse in variational autoencoders
J. He, D. Spokoyny, G. Neubig, and T. Berg-Kirkpatrick · 2019
Later among the works it cites.
Additional shared decoder on Siamese multi-view encoders for learning acoustic word embeddings
M. Jung, H. jun Lim, J. Goo, Y. Jung, and H. Kim · 2019
Later among the works it cites.
Truly unsupervised acoustic word embeddings using weak top-down constraints in encoder-decoder models
H. Kamper · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
An embedded segmental k-means model for unsupervised segmentation and clustering of speech
H. Kamper, K. Livescu, and S. Goldwater · 2017
Cited alongside, same era.
Query-by-example search with discriminative neural acoustic word embeddings
S. Settle, K. Levin, H. Kamper, and K. Livescu · 2017
Cited alongside, same era.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. Gomez, L. Kaiser, and I. Polosukhin · 2017
Cited alongside, same era.
Query-by-example spoken term detection using attention-based multi-hop networks
C.-W. Ao and H.-y. Lee · 2018
Cited alongside, same era.
Speech2vec: A sequence-to-sequence framework for learning word embeddings from speech
Y.-A. Chung and J. Glass · 2018
Cited alongside, same era.
Learning acoustic word embeddings with phonetically associated triplet network
H. Lim, Y. Kim, Y. Jung, M. Jung, and H. Kim · 2018
Cited alongside, same era.
J. Lucas, G. Tucker, R. B. Grosse, and M. Norouzi · 2019
Later among the works it cites.
Preventing posterior collapse with delta-VAEs
A. Razavi, A. v. d. Oord, B. Poole, and O. Vinyals · 2019
Later among the works it cites.
Acoustically grounded word embeddings for improved acoustics-to-word speech recognition
S. Settle, K. Audhkhasi, K. Livescu, and M. Picheny · 2019
Later among the works it cites.
Linguistically-informed training of acoustic word embeddings for low-resource languages
Z. Yang and J. Hirschberg · 2019
Later among the works it cites.
G. Beguš · 2020
Closest in time.
Analyzing autoencoder-based acoustic word embeddings
Y. Matusevych, H. Kamper, and S. Goldwater · 2020
Closest in time.
Whole-word segmental speech recognition with acoustic word embeddings
B. Shi, S. Settle, and K. Livescu · 2021
Closest in time.