Fetching the paper…
Reading the bibliography…
Multispectral image pairs can provide the combined information, making object detection applications more reliable and robust in the open world.
Y. Zheng, I. H. Izzat, S. Ziaee, GFD-SSD: gated fusion double SSD for multispectral pedestrian detection, CoRR abs/1903.06999 · 1903
Earlier work this paper cites.
2004
Earlier work this paper cites.
2008
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, G. E. Hinton, Imagenet classification with deep convolutional neural networks, in: Proceedings of the Advances in Neural Information Processing Systems, NeurIPS 2012, Lake Tahoe, Nevada, United States, December 3-6, 2012, pp. 1106–1114
2012
Earlier work this paper cites.
R. B. Girshick, J. Donahue, T. Darrell, J. Malik, Rich feature hierarchies for accurate object detection and semantic segmentation, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2014, Columbus, OH, USA, June 23-28, 2014, pp. 580–587
2014
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, C. L. Zitnick, Microsoft coco: Common objects in context, in: Proceedings of the European Conference Computer Vision, ECCV 2014, Zurich, Switzerland, September 6-12, 2014, pp. 740–755
2014
Earlier work this paper cites.
J. Liu, S. Zhang, S. Wang, D. N. Metaxas, Multispectral deep neural networks for pedestrian detection, in: Proceedings of the British Machine Vision Conference, BMVC 2016, York, UK, September 19-22, 2016
2016
Earlier work this paper cites.
S. Razakarivony, F. Jurie, Vehicle detection in aerial imagery : A small target detection benchmark, J. Vis. Commun. Image Represent. 34 (2016) 187–203
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, J. Sun, Deep residual learning for image recognition, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2016, Las Vegas, NV, USA, June 27-30, 2016, pp. 770–778
2016
Earlier work this paper cites.
W. Liu, D. Anguelov, D. Erhan, C. Szegedy, S. E. Reed, C. Fu, A. C. Berg, SSD: single shot multibox detector, in: Proceedings of the European Conference Computer Vision, ECCV 2016, Amsterdam, The Netherlands, October 11-14, 2016, Vol. 9905, pp. 21–37
2016
Earlier work this paper cites.
J. Redmon, S. K. Divvala, R. B. Girshick, A. Farhadi, You only look once: Unified, real-time object detection, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2016, Las Vegas, NV, USA, June 27-30, 2016, pp. 779–788
2016
Earlier work this paper cites.
J. Wagner, V. Fischer, M. Herman, S. Behnke, Multispectral pedestrian detection using deep fusion convolutional neural networks, in: Proceedings of the European Symposium on Artificial Neural Networks, ESANN 2016, Bruges, Belgium, April 27-29, 2016
2016
Earlier work this paper cites.
S. Ren, K. He, R. B. Girshick, J. Sun, Faster R-CNN: towards real-time object detection with region proposal networks, IEEE Trans. Pattern Anal. Mach. Intell. 39 (6) (2017) 1137–1149
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, I. Polosukhin, Attention is all you need, in: Proceedings of the Advances in Neural Information Processing Systems, NeurIPS 2017, Long Beach, CA, USA, December 4-9, 2017, pp. 5998–6008
2017
Earlier work this paper cites.
K. Park, S. Kim, K. Sohn, Unified multi-spectral pedestrian detection based on probabilistic fusion networks, Pattern Recognit. 80 (2018) 143–155
2018
Cited alongside, same era.
C. Li, D. Song, R. Tong, M. Tang, Multispectral pedestrian detection via simultaneous detection and segmentation, in: Proceedings of the British Machine Vision Conference, BMVC 2018, Newcastle, UK, September 3-6, 2018, 2018, p. 225
2018
Cited alongside, same era.
N. Ma, X. Zhang, H. Zheng, J. Sun, Shufflenet V2: practical guidelines for efficient CNN architecture design, in: Proceedings of the European Conference Computer Vision, ECCV 2018, Munich, Germany, September 8-14, 2018, Vol. 11218, pp. 122–138
2018
Cited alongside, same era.
C. Li, D. Song, R. Tong, M. Tang, Illumination-aware faster R-CNN for robust multispectral pedestrian detection, Pattern Recognit. 85 (2019) 161–171
2019
Cited alongside, same era.
L. C. O. Tiong, S. T. Kim, Y. M. Ro, Multimodal facial biometrics recognition: Dual-stream convolutional neural networks with multi-feature fusion layers, Image and Vision Computing 102 (2020) 103977
2020
Later among the works it cites.
F. Qingyun, Z. Lin, W. Zhaokui, An efficient feature pyramid network for object detection in remote sensing imagery, IEEE Access 8 (2020) 93058–93068
2020
Later among the works it cites.
V. Gabeur, C. Sun, K. Alahari, C. Schmid, Multi-modal transformer for video retrieval, in: Proceedings of the European Conference Computer Vision, ECCV 2020, Glasgow, UK, August 23-28, 2020, Vol. 12349, pp. 214–229
2020
Later among the works it cites.
V. Iashin, E. Rahtu, Multi-modal dense video captioning, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, CVPRW 2020, Seattle, WA, USA, June 14-19, 2020, pp. 4117–4126
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
L. Zhang, Z. Liu, S. Zhang, X. Yang, H. Qiao, K. Huang, A. Hussain, Cross-modality interactive attention network for multispectral pedestrian detection, Inf. Fusion 50 (2019) 20–29
2019
Cited alongside, same era.
L. Ye, M. Rochan, Z. Liu, Y. Wang, Cross-modal self-attention network for referring image segmentation, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2019, Long Beach, CA, USA, June 16-20, 2019, pp. 10502–10511
2019
Cited alongside, same era.
J. Lu, D. Batra, D. Parikh, S. Lee, Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks, in: Proceedings of the Advances in Neural Information Processing Systems, NeurIPS 2019, Vancouver, BC, Canada, December 8-14, 2019,, pp. 13–23
2019
Cited alongside, same era.
C. Sun, A. Myers, C. Vondrick, K. Murphy, C. Schmid, Videobert: A joint model for video and language representation learning, in: Proceedings of the IEEE/CVF International Conference on Computer Vision, ICCV 2019, Seoul, Korea (South), October 27 - November 2, 2019, pp. 7463–7472
2019
Cited alongside, same era.
X. Chen, D. Liu, C. Lei, R. Li, Z. Zha, Z. Xiong, Bert4sessrec: Content-based video relevance prediction with bidirectional encoder representations from transformer, in: Proceedings of the ACM International Conference on Multimedia, MM 2019, Nice, France, October 21-25, 2019, pp. 2597–2601
2019
Cited alongside, same era.
G. Li, L. Zhu, P. Liu, Y. Yang, Entangled transformer for image captioning, in: Proceedings of the IEEE/CVF International Conference on Computer Vision, ICCV 2019, Seoul, Korea (South), October 27 - November 2, 2019, 2019, pp. 8927–8936
2019
Cited alongside, same era.
L. Zhang, X. Zhu, X. Chen, X. Yang, Z. Lei, Z. Liu, Weakly aligned cross-modal learning for multispectral pedestrian detection, in: Proceedings of the IEEE/CVF International Conference on Computer Vision, ICCV 2019, Seoul, Korea (South), October 27 - November 2, 2019, pp. 5126–5136
2019
Cited alongside, same era.
H. Rezatofighi, N. Tsoi, J. Gwak, A. Sadeghian, I. Reid, S. Savarese, Generalized intersection over union: A metric and a loss for bounding box regression, in: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, CVPR 2019,Long Beach, CA, USA, June 16-20, 2019, pp. 658–666
2019
Cited alongside, same era.
M. Pham, L. Courtrai, C. Friguet, S. Lefèvre, A. Baussard, Yolo-fine: One-stage detector of small objects under various backgrounds in remote sensing images, Remote Sens. 12 (15) (2020) 2501–2526
2020
Later among the works it cites.
M. Dhanaraj, M. Sharma, T. Sarkar, S. Karnam, D. G. Chachlakis, R. Ptucha, P. P. Markopoulos, E. Saber, Vehicle detection from multi-modal aerial imagery using YOLOv3 with mid-level fusion, in: Proceedings of the SPIE 11395, Big Data II: Learning, Analytics, and Applications, 2020, pp. 22 – 32
2020
Later among the works it cites.
H. Zhang, É. Fromont, S. Lefèvre, B. Avignon, Guided attentive feature fusion for multispectral pedestrian detection, in: Proceedings of the IEEE Winter Conference on Applications of Computer Vision, WACV 2021, Waikoloa, HI, USA, January 3-8, 2021, pp. 72–80
2021
Closest in time.
M. Sharma, M. Dhanaraj, S. Karnam, D. G. Chachlakis, R. Ptucha, P. P. Markopoulos, E. Saber, Yolors: Object detection in multimodal remote sensing imagery, IEEE J. Sel. Top. Appl. Earth Observ. Remote Sens. 14 (2021) 1497–1508
2021
Closest in time.
X. Jia, C. Zhu, M. Li, W. Tang, W. Zhou, Llvip: A visible-infrared paired dataset for low-light vision, in: Proceedings of the IEEE/CVF International Conference on Computer Vision Workshops, ICCVW 2021, Montreal, BC, Canada, October 11-17, 2021, pp. 3489–3497
2021
Closest in time.
E. Tzinis, S. Wisdom, A. Jansen, S. Hershey, T. Remez, D. Ellis, J. R. Hershey, Into the wild with audioscope: Unsupervised audio-visual separation of on-screen sounds, in: Proceedings of the International Conference on Learning Representations, ICLR 2021, Virtual Event, Austria, May 3-7, 2021
2021
Closest in time.
M. Dzabraev, M. Kalashnikov, S. Komkov, A. Petiushko, MDMMT: multidomain multimodal transformer for video retrieval, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops, CVPRW 2021, virtual, June 19-25, 2021, pp. 3354–3363
2021
Closest in time.
F. Qingyun, W. Zhaokui, Cross-modality attentive feature fusion for object detection in multispectral remote sensing imagery, Pattern Recognit. (2022) 108786
2022
Closest in time.
A. Amudhan, A. Sudheer, Lightweight and computationally faster hypermetropic convolutional neural network for small size object detection, Image and Vision Computing 119 (2022) 104396
2022
Closest in time.