Fetching the paper…
Reading the bibliography…
Sound event detection (SED) entails two subtasks: recognizing what types of sound events are present in an audio stream (audio tagging), and pinpointing their onset and offset times (localization).
“A framework for multiple-instance learning”
Oded Maron and Tom“’as Lozano-P“’erez · 1998
Earlier work this paper cites.
“Multiple instance boosting for object detection”
Cha Zhang, John Platt and Paul Viola · 2006
Earlier work this paper cites.
“Simultaneous learning and alignment: Multi-instance and multi-pose learning”
Boris Babenko, Piotr Doll“’ar, Zhuowen Tu and Serge Belongie · 2008
Earlier work this paper cites.
“Multiple instance classification: Review, taxonomy and comparative study”
Jaume Amores · 2013
Earlier work this paper cites.
“XSEDE: Accelerating scientific discovery”
John Towns et al · 2014
Earlier work this paper cites.
“Exploiting spectro-temporal locality in deep learning based acoustic event detection”
Miquel Espi, Masakiyo Fujimoto, Keisuke Kinoshita and Tomohiro Nakatani · 2015
Earlier work this paper cites.
“Batch normalization: Accelerating deep network training by reducing internal covariate shift”
Sergey Ioffe and Christian Szegedy · 2015
Earlier work this paper cites.
“DCASE 2016 sound event detection system based on convolutional neural network”, 2016
Arseniy Gorin, Nurtas Makhazhanov and Nickolay Shmyrev · 2016
Earlier work this paper cites.
“Audio-based multimedia event detection using deep recurrent neural networks”
Yun Wang, Leonardo Neves and Florian Metze · 2016
Earlier work this paper cites.
“Recurrent neural networks for polyphonic sound event detection in real life recordings”
Giambattista Parascandolo, Heikki Huttunen and Tuomas Virtanen · 2016
Earlier work this paper cites.
“Sound event detection in multichannel audio using spatial and harmonic features”
Sharath Adavanne et al · 2016
Earlier work this paper cites.
“Bidirectional LSTM-HMM hybrid system for polyphonic sound event detection”
Tomoki Hayashi et al · 2016
Cited alongside, same era.
“Audio event detection using weakly labeled data”
Anurag Kumar and Bhiksha Raj · 2016
Cited alongside, same era.
“Metrics for polyphonic sound event detection”
Annamaria Mesaros, Toni Heittola and Tuomas Virtanen · 2016
Cited alongside, same era.
“Convolutional recurrent neural networks for polyphonic sound event detection”
Emre Cakr et al · 2017
Cited alongside, same era.
“Audio Set: An ontology and human-labeled dataset for audio events”
Jort Gemmeke et al · 2017
Cited alongside, same era.
“DCASE 2017 challenge setup: Tasks, datasets and baseline system”
Annamaria Mesaros et al · 2017
Cited alongside, same era.
“Comparing the max and noisy-or pooling functions in multiple instance learning for weakly supervised sequence learning tasks”
Yun Wang, Juncheng Li and Florian Metze · 2018
Closest in time.
“A closer look at weak label learning for audio events”
Ankit Shah, Anurag Kumar, Alexander Hauptmann and Bhiksha Raj · 2018
Closest in time.
“Large-scale weakly supervised audio classification using gated convolutional neural network”
Yong Xu, Qiuqiang Kong, Wenwu Wang and Mark Plumbley · 2018
Closest in time.
“Audio set classification with attention model: A probabilistic perspective”
Qiuqiang Kong, Yong Xu, Wenwu Wang and Mark Plumbley · 2018
Closest in time.
“Multi-level attention model for weakly supervised audio classification”
Changsong Yu, Karim Barsim, Qiuqiang Kong and Bin Yang · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Weakly-supervised audio event detection using event-specific Gaussian filters and fully convolutional networks”
Ting-Wei Su, Jen-Yu Liu and Yi-Hsuan Yang · 2017
Cited alongside, same era.
“Deep learning for DCASE2017 challenge”, 2017
An Dang, Toan Vu and Jia-Ching Wang · 2017
Cited alongside, same era.
“DCASE 2017 submission: Multiple instance learning for sound event detection”, 2017
Justin Salamon, Brian McFee and Peter Li · 2017
Cited alongside, same era.
“Automatic differentiation in PyTorch”
Adam Paszke et al · 2017
Cited alongside, same era.
“CNN architectures for large-scale audio classification”
Shawn Hershey · 2017
Cited alongside, same era.
“Knowledge transfer from weakly labeled audio using convolutional neural network for sound events and scenes”
Anurag Kumar, Maksim Khadkevich and Christian F“”ugen · 2018
Closest in time.
“Reducing model complexity for DNN based large-scale audio classification”
Yuzhong Wu and Tan Lee · 2018
Closest in time.
“Class-aware self-attention for audio event recognition”
Shizhe Chen, Jia Chen, Qin Jin and Alexander Hauptmann · 2018
Closest in time.
“Learning to recognize transient sound events using attentional supervision”
Szu-Yu Chou, Jyh-Shing Jang and Yi-Hsuan Yang · 2018
Closest in time.
“Adaptive pooling operators for weakly labeled sound event detection”
Brian McFee, Justin Salamon and Juan Bello · 2018
Closest in time.
“Polyphonic sound event detection with weak labeling”, 2018
Yun Wang · 2018
Closest in time.