Fetching the paper…
Reading the bibliography…
Recently, a variety of acoustic tasks and related applications arised.
G. Tzanetakis and P. Cook, “Marsyas: A framework for audio analysis,” Organised sound , vol. 4, no. 3, pp. 169–175, 2000
2000
Earlier work this paper cites.
K. Papineni, S. Roukos, T. Ward et al. , “Bleu: a method for automatic evaluation of machine translation,” in Proceedings of the 40th annual meeting on association for computational linguistics . Association for Computational Linguistics, 2002, pp. 311–318
2002
Earlier work this paper cites.
C. Busso, M. Bulut, C.-C. Lee et al. , “Iemocap: Interactive emotional dyadic motion capture database,” Language resources and evaluation , vol. 42, no. 4, p. 335, 2008
2008
Earlier work this paper cites.
V. Rozgic, S. Ananthakrishnan, S. Saleem et al. , “Ensemble of svm trees for multimodal emotion recognition,” in Signal Information Processing Association Summit And Conference , 2012
2012
Earlier work this paper cites.
2014
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey et al. , “Librispeech: an asr corpus based on public domain audio books,” in 2015 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2015, pp. 5206–5210
2015
Earlier work this paper cites.
K. J. Piczak, “Esc: Dataset for environmental sound classification,” in Proceedings of the 23rd ACM international conference on Multimedia , 2015, pp. 1015–1018
2015
Earlier work this paper cites.
R. Xia and Y. Liu, “Leveraging valence and activation information via multi-task learning for categorical emotion recognition,” in 2015 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2015, pp. 5301–5305
2015
Earlier work this paper cites.
R. Sennrich, B. Haddow, and A. Birch, “Neural machine translation of rare words with subword units,” in Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Berlin, Germany: Association for Computational Linguistics, Aug. 2016, pp. 1715–1725
2016
Cited alongside, same era.
A. Vaswani, N. Shazeer, N. Parmar et al. , “Attention is all you need,” in Advances in neural information processing systems , 2017, pp. 5998–6008
2017
Cited alongside, same era.
M. A. Di Gangi, R. Cattoni, L. Bentivogli et al. , “Must-c: a multilingual speech translation corpus,” in 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . Association for Computational Linguistics, 2019, pp. 2012–2017
2017
Cited alongside, same era.
G. Dekkers, S. Lauwereins, B. Thoen et al. , “The SINS database for detection of daily activities in a home environment using an acoustic sensor network,” in Proceedings of the Detection and Classification of Acoustic Scenes and Events 2017 Workshop (DCASE2017) , November 2017, pp. 32–36
H.-W. Liao, J.-Y. Huang, S.-S. Lan et al. , “DCASE 2018 task 5 challenge technical report: Sound event classification by a deep neural network with attention and minimum variance distortionless response enhancement,” DCASE2018 Challenge, Tech. Rep., September 2018
2018
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
M. Neumann and T. Vu, “Improving speech emotion recognition with unsupervised representation learning on unlabeled speech,” in International Conference on Acoustics, Speech, and Signal Processing, 2019 , 2019
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
T. Inoue, P. Vinayavekhin, S. Wang et al. , “Domestic activities classification based on CNN using shuffling and mixing data augmentation,” DCASE2018 Challenge, Tech. Rep., September 2018
2018
Cited alongside, same era.
H. Liu, F. Wang, X. Liu et al. , “An ensemble system for domestic activity recognition,” DCASE2018 Challenge, Tech. Rep., September 2018
2018
Cited alongside, same era.
M. A. D. Gangi, M. Negri, and M. Turchi, “Adapting transformer to end-to-end spoken language translation,” in Interspeech 2019 , 2019
2019
Later among the works it cites.
R. Pappagari, T. Wang, J. Villalba et al. , “x-vectors meet emotions: A study on dependencies between emotion and speaker recognition,” in ICASSP 2020-2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2020, pp. 7169–7173
2020
Closest in time.