Fetching the paper…
Reading the bibliography…
Environmental Sound Classification (ESC) is a challenging field of research in non-speech audio processing.
D. E. Rumelhart, G. E. Hinton, and R. J. Williams, “Learning representations by back-propagating errors,” nature , vol. 323, no. 6088, pp. 533–536, 1986
1986
Earlier work this paper cites.
E. Baum and F. Wilczek, “Supervised learning of probability distributions by neural networks,” in Neural information processing systems , 1987, pp. 52–61
1987
Earlier work this paper cites.
R. Radhakrishnan, A. Divakaran, and A. Smaragdis, “Audio analysis for surveillance applications,” in IEEE Workshop on Applications of Signal Processing to Audio and Acoustics, 2005. IEEE, 2005, pp. 158–161
2005
Earlier work this paper cites.
R. Hadsell, S. Chopra, and Y. LeCun, “Dimensionality reduction by learning an invariant mapping,” in 2006 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR’06) , vol. 2. IEEE, 2006, pp. 1735–1742
2006
Earlier work this paper cites.
H. Li, S. Ishikawa, Q. Zhao, M. Ebana, H. Yamamoto, and J. Huang, “Robot navigation and sound based position identification,” in 2007 IEEE International Conference on Systems, Man and Cybernetics
2007
Earlier work this paper cites.
R. Salakhutdinov and G. Hinton, “Learning a nonlinear embedding by preserving class neighbourhood structure,” in Artificial Intelligence and Statistics , 2007, pp. 412–419
2007
Earlier work this paper cites.
S. Chu, S. Narayanan, and C.-C. J. Kuo, “Environmental sound recognition with time–frequency audio features,” IEEE Transactions on Audio, Speech, and Language Processing , vol. 17, no. 6, pp. 1142–1158, 2009
2009
Earlier work this paper cites.
P. Dhanalakshmi, S. Palanivel, and V. Ramalingam, “Classification of audio signals using aann and gmm,” Applied soft computing , vol. 11, no. 1, pp. 716–723, 2011
2011
Earlier work this paper cites.
B. Uzkent, B. D. Barkana, and H. Cevikalp, “Non-speech environmental sound classification using svms with a new set of features,” International Journal of Innovative Computing, Information and Control , vol. 8, no. 5, pp. 3511–3524, 2012
2012
Earlier work this paper cites.
P. Intani and T. Orachon, “Crime warning system using image and sound processing,” in 2013 13th International Conference on Control, Automation and Systems (ICCAS 2013) . IEEE, 2013, pp. 1751–1753
2013
Earlier work this paper cites.
J. Salamon, C. Jacoby, and J. P. Bello, “A dataset and taxonomy for urban sound research,” in Proceedings of the 22nd ACM international conference on Multimedia , 2014, pp. 1041–1044
2014
Earlier work this paper cites.
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, “Dropout: a simple way to prevent neural networks from overfitting,” The journal of machine learning research , vol. 15, no. 1, pp. 1929–1958, 2014
2014
Earlier work this paper cites.
R. Dobre, V. Niţă, A. Ciobanu, C. Negrescu, and D. Stanomir, “Low computational method for siren detection,” in 2015 IEEE 21st International Symposium for Design and Technology in Electronic Packaging (SIITME) . IEEE, 2015, pp. 291–295
2015
Earlier work this paper cites.
P. Foggia, N. Petkov, A. Saggese, N. Strisciuglio, and M. Vento, “Reliable detection of audio events in highly noisy environments,” Pattern Recognition Letters , vol. 65, pp. 22–28, 2015
2015
Earlier work this paper cites.
K. J. Piczak, “Esc: Dataset for environmental sound classification,” in Proceedings of the 23rd ACM international conference on Multimedia , 2015, pp. 1015–1018
2015
Earlier work this paper cites.
A. Mesaros, T. Heittola, O. Dikmen, and T. Virtanen, “Sound event detection in real life recordings using coupled matrix factorization of spectral representations and class activity annotations,” in 2015 IEEE international conference on acoustics, speech and signal processing (ICASSP) . IEEE, 2015, pp. 151–155
2015
Cited alongside, same era.
J. Salamon and J. P. Bello, “Unsupervised feature learning for urban sound classification,” in 2015 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2015, pp. 171–175
2015
Cited alongside, same era.
K. J. Piczak, “Environmental sound classification with convolutional neural networks,” in 2015 IEEE 25th International Workshop on Machine Learning for Signal Processing (MLSP) . IEEE, 2015, pp. 1–6
2015
Cited alongside, same era.
E. Hoffer and N. Ailon, “Deep metric learning using triplet network,” in International Workshop on Similarity-Based Pattern Recognition . Springer, 2015, pp. 84–92
2015
B. Zhu, C. Wang, F. Liu, J. Lei, Z. Huang, Y. Peng, and F. Li, “Learning environmental sounds with multi-scale convolutional neural network,” in 2018 International Joint Conference on Neural Networks (IJCNN) . IEEE, 2018, pp. 1–8
2018
Later among the works it cites.
Z. Zhang, S. Xu, S. Cao, and S. Zhang, “Deep convolutional neural network with mixup for environmental sound classification,” in Chinese Conference on Pattern Recognition and Computer Vision (PRCV) . Springer, 2018, pp. 356–367
2018
Later among the works it cites.
A. Nasiri, J. Bao, D. Mccleeary, S.-Y. M. Louis, X. Huang, and J. Hu, “Online damage monitoring of sic f-sic m composite materials using acoustic emission and deep learning,” IEEE Access , vol. 7, pp. 140 534–140 541, 2019
2019
Later among the works it cites.
A. Nasiri, Y. Cui, Z. Liu, J. Jin, Y. Zhao, and J. Hu, “Audiomask: Robust sound event detection using mask r-cnn and frame-level classifier,” in 2019 IEEE 31st International Conference on Tools with Artificial Intelligence (ICTAI) . IEEE, 2019, pp. 485–492
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
V. Bisot, R. Serizel, S. Essid, and G. Richard, “Acoustic scene classification with matrix factorization for unsupervised feature learning,” in 2016 IEEE international conference on acoustics, speech and signal processing (ICASSP) . IEEE, 2016, pp. 6445–6449
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Cited alongside, same era.
E. Cakir, S. Adavanne, G. Parascandolo, K. Drossos, and T. Virtanen, “Convolutional recurrent neural networks for bird audio detection,” in 2017 25th European Signal Processing Conference (EUSIPCO) . IEEE, 2017, pp. 1744–1748
2017
Cited alongside, same era.
S. Hershey, S. Chaudhuri, D. P. Ellis, J. F. Gemmeke, A. Jansen, R. C. Moore, M. Plakal, D. Platt, R. A. Saurous, B. Seybold et al. , “Cnn architectures for large-scale audio classification,” in 2017 ieee international conference on acoustics, speech and signal processing (icassp) . IEEE, 2017, pp. 131–135
2017
Cited alongside, same era.
A. Mesaros, T. Heittola, E. Benetos, P. Foster, M. Lagrange, T. Virtanen, and M. D. Plumbley, “Detection and classification of acoustic scenes and events: Outcome of the dcase 2016 challenge,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 26, no. 2, pp. 379–393, 2017
2017
Cited alongside, same era.
Y. Tokozume and T. Harada, “Learning environmental sounds with end-to-end convolutional neural network,” in 2017 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2017, pp. 2721–2725
2017
Cited alongside, same era.
J. Salamon and J. P. Bello, “Deep convolutional neural networks and data augmentation for environmental sound classification,” IEEE Signal Processing Letters , vol. 24, no. 3, pp. 279–283, 2017
2017
Cited alongside, same era.
Y. Tokozume, Y. Ushiku, and T. Harada, “Learning from between-class examples for deep sound recognition,” in International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
2019
Later among the works it cites.
B. da Silva, A. W Happi, A. Braeken, and A. Touhafi, “Evaluation of classical machine learning techniques towards urban sound recognition on embedded systems,” Applied Sciences , vol. 9, no. 18, p. 3885, 2019
2019
Later among the works it cites.
Z. Zhang, S. Xu, S. Zhang, T. Qiao, and S. Cao, “Learning attentive representations for environmental sound classification,” IEEE Access , vol. 7, pp. 130 327–130 339, 2019
2019
Later among the works it cites.
2020
Later among the works it cites.
V. Abrol and P. Sharma, “Learning hierarchy aware embedding from raw audio for acoustic scene classification,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 28, pp. 1964–1973, 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
K. He, H. Fan, Y. Wu, S. Xie, and R. Girshick, “Momentum contrast for unsupervised visual representation learning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 9729–9738
2020
Later among the works it cites.
N. Frosst, N. Papernot, and G. Hinton, “Analyzing and improving representations with the soft nearest neighbor loss,” in International Conference on Machine Learning , 2019, pp. 2012–2020
2020
Later among the works it cites.
2020
Later among the works it cites.
T. Chen, S. Kornblith, M. Norouzi, and G. Hinton, “A simple framework for contrastive learning of visual representations,” in International conference on machine learning . PMLR, 2020, pp. 1597–1607
2020
Later among the works it cites.
2020
Later among the works it cites.