Fetching the paper…
Reading the bibliography…
There are still two problems in SDD causing some inaccurate results: (1) In the process of feature extraction, with the layer-by-layer acquisition of semantic information, local information is gradually lost, resulting into less representative feature maps; (2) During the Non-Maximum Suppression (NMS) algorithm due to inconsistency in classification and regression tasks, the classification confidence and predicted detection position cannot accurately indicate the position of the prediction boxes.
K. He, X. Zhang, S. Ren, J. Sun, Spatial pyramid pooling in deep convolutional networks for visual recognition, IEEE transactions on pattern analysis and machine intelligence 37 (9) (2015) 1904–1916
1916
Earlier work this paper cites.
W. Khan, K. Raj, T. Kumar, A. M. Roy, B. Luo, Introducing urdu digits dataset with demonstration of an efficient and robust noisy decoder-based pseudo example generator, Symmetry 14 (10) (2022) 1976
1976
Earlier work this paper cites.
S. Lazebnik, C. Schmid, J. Ponce, Beyond bags of features: Spatial pyramid matching for recognizing natural scene categories, in: 2006 IEEE computer society conference on computer vision and pattern recognition (CVPR’06), Vol. 2, IEEE, 2006, pp. 2169–2178
2006
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, L. Fei-Fei, Imagenet: A large-scale hierarchical image database, in: 2009 IEEE conference on computer vision and pattern recognition, Ieee, 2009, pp. 248–255
2009
Earlier work this paper cites.
Z. Fu, G. Lu, K. M. Ting, D. Zhang, A survey of audio-based music classification and annotation, IEEE transactions on multimedia 13 (2) (2010) 303–319
2010
Earlier work this paper cites.
M. Everingham, L. Van Gool, C. K. Williams, J. Winn, A. Zisserman, The pascal visual object classes (voc) challenge, International journal of computer vision 88 (2) (2010) 303–338
2010
Earlier work this paper cites.
P. F. Felzenszwalb, R. B. Girshick, D. McAllester, D. Ramanan, Object detection with discriminatively trained part-based models, IEEE transactions on pattern analysis and machine intelligence 32 (9) (2010) 1627–1645
2010
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, G. E. Hinton, Imagenet classification with deep convolutional neural networks, Advances in neural information processing systems 25 (2012)
2012
Earlier work this paper cites.
J. R. Uijlings, K. E. Van De Sande, T. Gevers, A. W. Smeulders, Selective search for object recognition, International journal of computer vision 104 (2) (2013) 154–171
2013
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, C. L. Zitnick, Microsoft coco: Common objects in context, in: European conference on computer vision, Springer, 2014, pp. 740–755
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
S. Ren, K. He, R. Girshick, J. Sun, Faster r-cnn: Towards real-time object detection with region proposal networks, in: Advances in neural information processing systems, 2015, pp. 91–99
2015
Earlier work this paper cites.
R. Girshick, Fast r-cnn, in: Proceedings of the IEEE international conference on computer vision, 2015, pp. 1440–1448
2015
Earlier work this paper cites.
K. Lenc, A. Vedaldi, R-cnn minus r, arXiv preprint arXiv:1506.06981 (2015)
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
M. Crocco, M. Cristani, A. Trucco, V. Murino, Audio surveillance: A systematic review, ACM Computing Surveys (CSUR) 48 (4) (2016) 1–46
2016
Earlier work this paper cites.
J. Redmon, S. Divvala, R. Girshick, A. Farhadi, You only look once: Unified, real-time object detection, in: Proceedings of the IEEE conference on computer vision and pattern recognition, 2016, pp. 779–788
2016
Earlier work this paper cites.
W. Liu, D. Anguelov, D. Erhan, C. Szegedy, S. Reed, C.-Y. Fu, A. C. Berg, Ssd: Single shot multibox detector, in: European conference on computer vision, Springer, 2016, pp. 21–37
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, J. Sun, Deep residual learning for image recognition, in: Proceedings of the IEEE conference on computer vision and pattern recognition, 2016, pp. 770–778
2016
Earlier work this paper cites.
J. Dai, Y. Li, K. He, J. Sun, R-fcn: Object detection via region-based fully convolutional networks, in: Advances in neural information processing systems, 2016, pp. 379–387
2016
Earlier work this paper cites.
K. U. Sharma, N. V. Thakur, A review and an approach for object detection in images, International Journal of Computational Vision and Robotics 7 (1/2) (2017) 196–237
2017
Earlier work this paper cites.
T.-Y. Lin, P. Dollár, R. Girshick, K. He, B. Hariharan, S. Belongie, Feature pyramid networks for object detection, in: Proceedings of the IEEE conference on computer vision and pattern recognition, 2017, pp. 2117–2125
2017
Earlier work this paper cites.
T.-Y. Lin, P. Goyal, R. Girshick, K. He, P. Dollár, Focal loss for dense object detection, in: Proceedings of the IEEE international conference on computer vision, 2017, pp. 2980–2988
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Cited alongside, same era.
Y. Zhu, C. Zhao, J. Wang, X. Zhao, Y. Wu, H. Lu, Couplenet: Coupling global structure with local parts for object detection, in: Proceedings of the IEEE International Conference on Computer Vision, 2017, pp. 4126–4134
2017
Cited alongside, same era.
J. Dai, H. Qi, Y. Xiong, Y. Li, G. Zhang, H. Hu, Y. Wei, Deformable convolutional networks, in: Proceedings of the IEEE international conference on computer vision, 2017, pp. 764–773
2017
Cited alongside, same era.
A. Asvadi, L. Garrote, C. Premebida, P. Peixoto, U. J. Nunes, Multimodal vehicle detection: fusing 3d-lidar and color camera data, Pattern Recognition Letters 115 (2018) 20–29
2018
Cited alongside, same era.
A. Chandio, Y. Shen, M. Bendechache, I. Inayat, T. Kumar, Audd: audio urdu digits dataset for automatic audio urdu digit recognition, Applied Sciences 11 (19) (2021) 8842
2021
Later among the works it cites.
S. Minaee, N. Kalchbrenner, E. Cambria, N. Nikzad, M. Chenaghlu, J. Gao, Deep learning–based text classification: a comprehensive review, ACM Computing Surveys (CSUR) 54 (3) (2021) 1–40
2021
Later among the works it cites.
S. Selva Birunda, R. Kanniga Devi, A review on word embedding techniques for text classification, Innovative Data Communication Technologies and Application (2021) 267–281
2021
Later among the works it cites.
N. Aslam, I. Ullah Khan, F. S. Alotaibi, L. A. Aldaej, A. K. Aldubaikil, Fake detect: A deep learning ensemble model for fake news detection, complexity 2021 (2021)
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
B. Jiang, R. Luo, J. Mao, T. Xiao, Y. Jiang, Acquisition of localization confidence for accurate object detection, in: Proceedings of the European conference on computer vision (ECCV), 2018, pp. 784–799
2018
Cited alongside, same era.
Z. Cai, N. Vasconcelos, Cascade r-cnn: Delving into high quality object detection, in: Proceedings of the IEEE conference on computer vision and pattern recognition, 2018, pp. 6154–6162
2018
Cited alongside, same era.
S. Liu, D. Huang, et al., Receptive field block net for accurate and fast object detection, in: Proceedings of the European Conference on Computer Vision (ECCV), 2018, pp. 385–400
2018
Cited alongside, same era.
S. Zhang, L. Wen, X. Bian, Z. Lei, S. Z. Li, Single-shot refinement neural network for object detection, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, 2018, pp. 4203–4212
2018
Cited alongside, same era.
S. Liu, L. Qi, H. Qin, J. Shi, J. Jia, Path aggregation network for instance segmentation, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, 2018, pp. 8759–8768
2018
Cited alongside, same era.
T. Kong, F. Sun, C. Tan, H. Liu, W. Huang, Deep feature pyramid reconfiguration for object detection, in: Proceedings of the European Conference on Computer Vision (ECCV), 2018, pp. 169–185
2018
Cited alongside, same era.
N. Ma, X. Zhang, H.-T. Zheng, J. Sun, Shufflenet v2: Practical guidelines for efficient cnn architecture design, in: Proceedings of the European Conference on Computer Vision (ECCV), 2018, pp. 116–131
2018
Cited alongside, same era.
I. Ullah, S. Khan, M. Imran, Y.-K. Lee, Rweetminer: Automatic identification and categorization of help requests on twitter during disasters, Expert Systems with Applications 176 (2021) 114787
2021
Later among the works it cites.
A. M. Roy, Finite element framework for efficient design of three dimensional multicomponent composite helicopter rotor blade system, Eng 2 (1) (2021) 69–79
2021
Later among the works it cites.
A. M. Roy, J. Bhaduri, A deep learning enabled multi-class plant disease detection model based on computer vision, AI 2 (3) (2021) 413–428
2021
Later among the works it cites.
Y. Liu, P. Sun, N. Wergeles, Y. Shang, A survey and performance evaluation of deep learning methods for small object detection, Expert Systems with Applications 172 (2021) 114602
2021
Later among the works it cites.
Y. Ji, H. Zhang, Z. Zhang, M. Liu, Cnn-based encoder-decoder networks for salient object detection: A comprehensive review and recent advances, Information Sciences 546 (2021) 835–857
2021
Later among the works it cites.
A. Hiemann, T. Kautz, T. Zottmann, M. Hlawitschka, Enhancement of speed and accuracy trade-off for sports ball detection in videos—finding fast moving, small objects in real time, Sensors 21 (9) (2021) 3214
2021
Later among the works it cites.
S. K. Pal, A. Pramanik, J. Maiti, P. Mitra, Deep learning in multi-object detection and tracking: state of the art, Applied Intelligence 51 (9) (2021) 6400–6429
2021
Later among the works it cites.
S. Jamil, M. S. Abbas, A. M. Roy, Distinguishing malicious drones using vision transformer, AI 3 (2) (2022) 260–273
2022
Closest in time.
2022
Closest in time.
A. M. Roy, An efficient multi-scale CNN model with intrinsic feature integration for motor imagery EEG subject classification in brain-machine interfaces, Biomedical Signal Processing and Control 74 (2022) 103496
2022
Closest in time.
A. M. Roy, A multi-scale fusion cnn model based on adaptive transfer learning for multi-class mi-classification in bci system, BioRxiv (2022)
2022
Closest in time.
A. M. Roy, Adaptive transfer learning-based multiscale feature fused deep convolutional neural network for eeg mi multiclassification in brain–computer interface, Engineering Applications of Artificial Intelligence 116 (2022) 105347
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
Y. Jiang, W. Zhang, K. Fu, Q. Zhao, Meanet: Multi-modal edge-aware network for light field salient object detection, Neurocomputing 491 (2022) 78–90
2022
Closest in time.
A. M. Roy, R. Bose, J. Bhaduri, A fast accurate fine-grain object detection model based on YOLOv4 deep neural network, Neural Computing and Applications (2022) 1–27
2022
Closest in time.
A. M. Roy, J. Bhaduri, Real-time growth stage detection model for high degree of occultation using densenet-fused YOLOv4, Computers and Electronics in Agriculture 193 (2022) 106694
2022
Closest in time.
S. S. A. Zaidi, M. S. Ansari, A. Aslam, N. Kanwal, M. Asghar, B. Lee, A survey of modern deep learning based object detection models, Digital Signal Processing (2022) 103514
2022
Closest in time.