Fetching the paper…
Reading the bibliography…
Audio content analysis in terms of sound events is an important research problem for a variety of applications.
E. Wold, T. Blum, D. Keislar, and J. Wheaten, “Content-based classification, search, and retrieval of audio,” IEEE multimedia , vol. 3, no. 3, pp. 27–36, 1996
1996
Earlier work this paper cites.
G. Guo and S. Z. Li, “Content-based audio classification and retrieval by support vector machines,” IEEE transactions on Neural Networks , vol. 14, no. 1, pp. 209–215, 2003
2003
Earlier work this paper cites.
S. Andrews, I. Tsochantaridis, and T. Hofmann, “Support vector machines for multiple-instance learning,” in Advances in neural information processing systems , 2003, pp. 577–584
2003
Earlier work this paper cites.
Z.-H. Zhou, “Multi-instance learning: A survey,” Department of Computer Science & Technology, Nanjing University, Tech. Rep , 2004
2004
Earlier work this paper cites.
C. Buckley and E. Voorhees, “Retrieval evaluation with incomplete information,” in Proceedings of the 27th annual international ACM SIGIR conference on Research and development in information retrieval . ACM, 2004, pp. 25–32
2004
Earlier work this paper cites.
T. Fawcett, “Roc graphs: Notes and practical considerations for researchers,” Machine learning , vol. 31, no. 1, pp. 1–38, 2004
2004
Earlier work this paper cites.
C. Clavel, T. Ehrette, and G. Richard, “Events detection for an audio-based surveillance system,” in Multimedia and Expo, 2005. ICME 2005. IEEE International conference on . IEEE, 2005, pp. 1306–1309
2005
Earlier work this paper cites.
C. Zieger and M. Omologo, “Acoustic event detection-itc-irst aed database,” Internal ITC report , 2005
2005
Earlier work this paper cites.
G. Valenzise, L. Gerosa, M. Tagliasacchi, F. Antonacci, and A. Sarti, “Scream and gunshot detection and localization for audio-surveillance systems,” in Advanced Video and Signal Based Surveillance, 2007. AVSS 2007. IEEE Conference on . IEEE, 2007, pp. 21–26
2007
Earlier work this paper cites.
L. Gerosa, G. Valenzise, M. Tagliasacchi, F. Antonacci, and A. Sarti, “Scream and gunshot detection in noisy environments,” in Signal Processing Conference, 2007 15th European . IEEE, 2007, pp. 1216–1220
2007
Earlier work this paper cites.
J. Tang, S. Yan, R. Hong, G.-J. Qi, and T.-S. Chua, “Inferring semantic concepts from community-contributed images and noisy tags,” in Proceedings of the 17th ACM international conference on Multimedia . ACM, 2009, pp. 223–232
2009
Earlier work this paper cites.
Y. Zigel, D. Litvak, and I. Gannot, “A method for automatic fall detection of elderly people using floor vibrations and sound—proof of concept on human mimicking doll falls,” IEEE Transactions on Biomedical Engineering , vol. 56, no. 12, pp. 2858–2867, 2009
2009
Earlier work this paper cites.
Y. Li, Z. Zeng, M. Popescu, and K. Ho, “Acoustic fall detection using a circular microphone array,” in Engineering in Medicine and Biology Society (EMBC), 2010 Annual International Conference of the IEEE . IEEE, 2010, pp. 2242–2245
2010
Earlier work this paper cites.
X. Zhuang, X. Zhou, M. A. Hasegawa-Johnson, and T. S. Huang, “Real-world acoustic event detection,” Pattern Recognition Letters , vol. 31, no. 12, pp. 1543–1551, 2010
2010
Earlier work this paper cites.
V. Nair and G. E. Hinton, “Rectified linear units improve restricted boltzmann machines,” in Proceedings of the 27th international conference on machine learning (ICML-10) , 2010, pp. 807–814
2010
Earlier work this paper cites.
T. Heittola, A. Mesaros, T. Virtanen, and A. Eronen, “Sound event detection in multisource environments using source separation,” in Machine Listening in Multisource Environments , 2011
2011
Earlier work this paper cites.
S. Pancoast and M. Akbacak, “Bag-of-audio-words approach for multimedia event classification,” in Thirteenth Annual Conference of the International Speech Communication Association , 2012
2012
Cited alongside, same era.
A. Kumar, P. Dighe, R. Singh, S. Chaudhuri, and B. Raj, “Audio event detection from acoustic unit occurrence patterns,” in Acoustics, Speech and Signal Processing (ICASSP), 2012 IEEE International Conference on . IEEE, 2012, pp. 489–492
2012
Cited alongside, same era.
J. F. Gemmeke, L. Vuegen, P. Karsmakers, B. Vanrumste et al. , “An exemplar-based nmf approach to audio event detection,” in Applications of Signal Processing to Audio and Acoustics (WASPAA), 2013 IEEE Workshop on . IEEE, 2013, pp. 1–4
2013
Cited alongside, same era.
J. Salamon, C. Jacoby, and J. P. Bello, “A dataset and taxonomy for urban sound research,” in Proceedings of the 22nd ACM international conference on Multimedia . ACM, 2014, pp. 1041–1044
2014
Cited alongside, same era.
A. Kumar and B. Raj, “Audio event detection using weakly labeled data,” in Proceedings of the 2016 ACM on Multimedia Conference , ser. MM ’16. New York, NY, USA: ACM, 2016, pp. 1038–1047. [Online]. Available: http://doi.acm.org/10.1145/2964284.2964310
2016
Later among the works it cites.
C. Debes, A. Merentitis, S. Sukhanov, M. Niessen, N. Frangiadakis, and A. Bauer, “Monitoring activities of daily living in smart homes: Understanding human behavior,” IEEE Signal Processing Magazine , vol. 33, no. 2, pp. 81–94, 2016
2016
Later among the works it cites.
S. Greene, H. Thapliyal, and D. Carpenter, “Iot-based fall detection for smart home environments,” in Nanoelectronic and Information Systems (iNIS), 2016 IEEE International Symposium on . IEEE, 2016, pp. 23–28
2016
Later among the works it cites.
2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Z. Feng, S. Feng, R. Jin, and A. K. Jain, “Image tag completion by noisy matrix recovery,” in European Conference on Computer Vision . Springer, 2014, pp. 424–438
2014
Cited alongside, same era.
J. Maxime, X. Alameda-Pineda, L. Girin, and R. Horaud, “Sound representation and classification benchmark for domestic robots,” in Robotics and Automation (ICRA), 2014 IEEE International Conference on . IEEE, 2014, pp. 6285–6292
2014
Cited alongside, same era.
D. Stowell and M. D. Plumbley, “Automatic large-scale classification of bird sounds is strongly improved by unsupervised feature learning,” PeerJ , vol. 2, p. e488, 2014
2014
Cited alongside, same era.
X. Lu, Y. Tsao, S. Matsuda, and C. Hori, “Sparse representation based on a bag of spectral exemplars for acoustic event detection,” in Acoustics, Speech and Signal Processing (ICASSP), 2014 IEEE International Conference on . IEEE, 2014, pp. 6255–6259
2014
Cited alongside, same era.
2014
Cited alongside, same era.
2014
Cited alongside, same era.
D. Stowell, D. Giannoulis, E. Benetos, M. Lagrange, and M. D. Plumbley, “Detection and classification of acoustic scenes and events,” IEEE Transactions on Multimedia , vol. 17, no. 10, pp. 1733–1746, 2015
2015
Cited alongside, same era.
K. J. Piczak, “Esc: Dataset for environmental sound classification,” in Proceedings of the 23rd ACM international conference on Multimedia . ACM, 2015, pp. 1015–1018
2015
Cited alongside, same era.
Later among the works it cites.
A. Kumar and B. Raj, “Weakly supervised scalable audio content analysis,” in Multimedia and Expo (ICME), 2016 IEEE International Conference on . IEEE, 2016, pp. 1–6
2016
Later among the works it cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Later among the works it cites.
——, “Deep cnn framework for audio event recognition using weakly labeled web data,” in Machine Learning for Audio, 2017 NIPS Workshop on . NIPS, 2017
2017
Later among the works it cites.
A. Mesaros, T. Heittola, A. Diment, B. Elizalde, A. Shah, E. Vincent, B. Raj, and T. Virtanen, “Dcase 2017 challenge setup: Tasks, datasets and baseline system,” in DCASE 2017-Workshop on Detection and Classification of Acoustic Scenes and Events , 2017
2017
Later among the works it cites.
J. F. Gemmeke, D. P. W. Ellis, D. Freedman, A. Jansen, W. Lawrence, R. C. Moore, M. Plakal, and M. Ritter, “Audio set: An ontology and human-labeled dataset for audio events,” in Proc. IEEE ICASSP 2017 , New Orleans, LA, 2017
2017
Later among the works it cites.
Z. Lu, Z. Fu, T. Xiang, P. Han, L. Wang, and X. Gao, “Learning from weak and noisy labels for semantic segmentation,” IEEE transactions on pattern analysis and machine intelligence , vol. 39, no. 3, pp. 486–500, 2017
2017
Later among the works it cites.
S. Hershey, S. Chaudhuri, D. P. Ellis, J. F. Gemmeke, A. Jansen, R. C. Moore, M. Plakal, D. Platt, R. A. Saurous, B. Seybold et al. , “Cnn architectures for large-scale audio classification,” in Acoustics, Speech and Signal Processing (ICASSP), 2017 IEEE International Conference on . IEEE, 2017, pp. 131–135
2017
Later among the works it cites.
T.-W. Su, J.-Y. Liu, and Y.-H. Yang, “Weakly-supervised audio event detection using event-specific gaussian filters and fully convolutional networks,” in Acoustics, Speech and Signal Processing (ICASSP), 2017 IEEE International Conference on . IEEE, 2017, pp. 791–795
2017
Later among the works it cites.
2017
Later among the works it cites.
A. Kumar and B. Raj, “Audio event and scene recognition: A unified approach using strongly and weakly labeled data,” in Neural Networks (IJCNN), 2017 International Joint Conference on . IEEE, 2017, pp. 3475–3482
2017
Later among the works it cites.
Y. Wu and T. Lee, “Reducing Model Complexity for DNN Based Large-Scale Audio Classification,” ArXiv e-prints , Nov. 2017
2017
Later among the works it cites.
A. Kumar, M. Khadkevich, and C. Fugen, “Knowledge transfer from weakly labeled audio using convolutional neural network for sound events and scenes,” in Acoustics, Speech and Signal Processing (ICASSP), 2012 IEEE International Conference on . IEEE, 2018
2018
Closest in time.