K. He, X. Zhang, S. Ren, and J. Sun, “Identity Mappings in Deep Residual Networks,” in
2016
Cited alongside, same era.
J. F. Gemmeke, D. P. W. Ellis, D. Freedman, A. Jansen, W. Lawrence, R. C. Moore, M. Plakal, and M. Ritter, “Audio Set: An ontology and human-labeled dataset for audio events,” in
2017
Cited alongside, same era.
e. a. B. Rocha, “A Respiratory Sound Database for the Development of Automated Classification,” in
2017
Cited alongside, same era.
S. Zagoruyko and N. Komodakis, “Wide Residual Networks,” 2017
2017
Cited alongside, same era.
J. Li, W. Dai, F. Metze, S. Qu, and S. Das, “A comparison of Deep Learning methods for environmental sound detection,” in
2017
Cited alongside, same era.
S. Livingstone and F. Russo, “The Ryerson Audio-Visual Database of Emotional Speech and Song (RAVDESS): A dynamic, multimodal set of facial and vocal expressions in North American English,”
2018
Cited alongside, same era.
Y. Tokozume, Y. Ushiku, and T. Harada, “Learning from Between-class Examples for Deep Sound Recognition,” in
2018
Cited alongside, same era.
N. Turpault, R. Serizel, A. Shah, and J. Salamon, “Sound event detection in domestic environments with weakly labeled data and soundscape synthesis,” in
2019
Cited alongside, same era.
M. Tan and Q. Le, “EfficientNet: Rethinking Model Scaling for Convolutional Neural Networks,” in
2019
Cited alongside, same era.
W. C. Daniel S. Park, Y. Zhang, C.-C. Chiu, B. Zoph, E. D. Cubuk, and Q. V. Le, “SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition,” in
2019
Cited alongside, same era.