Fetching the paper…
Reading the bibliography…
In this paper, we investigate the feasibility of applying few-shot learning algorithms to a speech task.
I. Szoke, P. Schwarz, P. Matejka, L. Burget, M. Karafiát, M. Fapso, and J. Cernocky, “Comparison of keyword spotting approaches for informal continuous speech,” in
2005
Earlier work this paper cites.
S. Ioffe and C. Szegedy, “Batch normalization: Accelerating deep network training by reducing internal covariate shift,” in
2015
Earlier work this paper cites.
T. Ko, V. Peddinti, D. Povey, and S. Khudanpur, “Audio augmentation for speech recognition,” in
2015
Earlier work this paper cites.
O. Vinyals, C. Blundell, T. Lillicrap, D. Wierstra
2016
Earlier work this paper cites.
A. Santoro, S. Bartunov, M. Botvinick, D. Wierstra, and T. Lillicrap, “Meta-learning with memory-augmented neural networks,” in
2016
Earlier work this paper cites.
K. Audhkhasi, A. Rosenberg, A. Sethy, B. Ramabhadran, and B. Kingsbury, “End-to-end asr-free keyword search from speech,”
2017
Earlier work this paper cites.
A. Rosenberg, K. Audhkhasi, A. Sethy, B. Ramabhadran, and M. Picheny, “End-to-end speech recognition and keyword search on low-resource languages,” in
2017
Earlier work this paper cites.
J. Trmal, M. Wiesner, V. Peddinti, X. Zhang, P. Ghahremani, Y. Wang, V. Manohar, H. Xu, D. Povey, and S. Khudanpur, “The kaldi openkws system: Improving low resource keyword search.” in
2017
Earlier work this paper cites.
J. Snell, K. Swersky, and R. Zemel, “Prototypical networks for few-shot learning,” in
2017
Cited alongside, same era.
T. Munkhdalai and H. Yu, “Meta networks,” in
2017
Cited alongside, same era.
S. Ravi and H. Larochelle, “Optimization as a model for few-shot learning,” in
2017
Cited alongside, same era.
C. Finn, P. Abbeel, and S. Levine, “Model-agnostic meta-learning for fast adaptation of deep networks,” in
2017
Cited alongside, same era.
Z. Li, F. Zhou, F. Chen, and H. Li, “Meta-sgd: Learning to learn quickly for few-shot learning,”
2017
Cited alongside, same era.
T. Ko, V. Peddinti, D. Povey, M. L. Seltzer, and S. Khudanpur, “A study on data augmentation of reverberant speech for robust speech recognition,” in
A. Nichol, J. Achiam, and J. Schulman, “On first-order meta-learning algorithms,”
2018
Closest in time.
P. Warden, “Speech commands: A dataset for limited-vocabulary speech recognition,”
2018
Closest in time.
N. Sacchi, A. Nanchen, M. Jaggi, and M. Cernak, “Open-vocabulary keyword spotting with audio and text embeddings,” in
2019
Closest in time.
C.-C. Kao, M. Sun, Y. Gao, S. Vitaladevuni, and C. Wang, “Sub-band convolutional neural networks for small-footprint spoken term classification,” in
2019
Closest in time.
Y. Zhu, T. Ko, and B. Mak, “Mixup learning strategies for text-independent speaker verification,” in
2019
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
F. Sung, Y. Yang, L. Zhang, T. Xiang, P. H. Torr, and T. M. Hospedales, “Learning to compare: Relation network for few-shot learning,” in
2018
Cited alongside, same era.
N. Mishra, M. Rohaninejad, X. Chen, and P. Abbeel, “A simple neural attentive meta-learner,” in
2018
Cited alongside, same era.
“Tensorflow: Simple audio recognition,”
Cited in the paper.
E. Sharma, G. Ye, W. Wei, R. Zhao, Y. Tian, J. Wu, L. He, E. Lin, and Y. Gong, “Adaptation of rnn transducer with text-to-speech technology for keyword spotting,” in
2020
Closest in time.
T. Ko, Y. Chen, and Q. Li, “Prototypical networks for small footprint text-independent speaker verification,” in
2020
Closest in time.