Fetching the paper…
Reading the bibliography…
Data-fusion networks have shown significant promise for RGB-thermal scene parsing.
W. Zhou, et al., CACFNet: Cross-Modal Attention Cascaded Fusion Network for RGB-T Urban Scene Parsing, IEEE Transactions on Intelligent Vehicles 9 (1) (2023) 1919–1929
1929
Earlier work this paper cites.
A. Krizhevsky, et al., ImageNet Classification with Deep Convolutional Neural Networks, Advances in Neural Information Processing Systems (NeurIPS) 25 (2012) 1097–1105
2012
Earlier work this paper cites.
N. Silberman, et al., Indoor Segmentation and Support Inference from RGBD Images, in: European Conference on Computer Vision (ECCV), Springer, 2012, pp. 746–760
2012
Earlier work this paper cites.
B. Hariharan, et al., Simultaneous Detection and Segmentation, in: Proceedings of the European Conference on Computer Vision (ECCV), Springer, 2014, pp. 297–312
2014
Earlier work this paper cites.
J. Long, et al., Fully Convolutional Networks for Semantic Segmentation, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2015, pp. 3431–3440
2015
Earlier work this paper cites.
J. Dai, et al., Convolutional Feature Masking for Joint Object and Stuff Segmentation, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2015, pp. 3992–4000
2015
Earlier work this paper cites.
K. He, et al., Deep Residual Learning for Image Recognition, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2016, pp. 770–778
2016
Earlier work this paper cites.
M. Cordts, et al., The Cityscapes Dataset for Semantic Urban Scene Understanding, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2016, pp. 3213–3223
2016
Earlier work this paper cites.
Q. Ha, et al., MFNet: Towards Real-Time Semantic Segmentation for Autonomous Vehicles with Multi-Spectral Scenes, in: 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), IEEE, 2017, pp. 5108–5115
2017
Earlier work this paper cites.
L.-C. Chen, et al., DeepLab: Semantic Image Segmentation with Deep Convolutional Nets, Atrous Convolution, and Fully Connected CRFs, IEEE Transactions on Pattern Analysis and Machine Intelligence 40 (4) (2017) 834–848
2017
Earlier work this paper cites.
T.-Y. Lin, et al., Feature Pyramid Networks for Object Detection, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2017, pp. 2117–2125
2017
Earlier work this paper cites.
A. Vaswani, et al., Attention is All you Need, Advances in Neural Information Processing Systems (NeurIPS) 30 (2017) 5998–6008
2017
Earlier work this paper cites.
G. Huang, et al., Densely Connected Convolutional Networks, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2017, pp. 4700–4708
2017
Earlier work this paper cites.
S. Xie, et al., Aggregated Residual Transformations for Deep Neural Networks, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2017, pp. 1492–1500
2017
Earlier work this paper cites.
L. Chen et al
2018
Earlier work this paper cites.
T. Xiao, et al., Unified Perceptual Parsing for Scene Understanding, in: Proceedings of the European Conference on Computer Vision (ECCV), 2018, pp. 418–434
2018
Earlier work this paper cites.
I. Loshchilov, F. Hutter, Decoupled Weight Decay Regularization, in: International Conference on Learning Representations (ICLR), 2018
2018
Earlier work this paper cites.
Y. Sun, et al., RTFNet: RGB-Thermal Fusion Network for Semantic Segmentation of Urban Scenes, IEEE Robotics and Automation Letters 4 (3) (2019) 2576–2583
2019
Earlier work this paper cites.
Q. Zhang, et al., RGB-T Salient Object Detection via Fusing Multi-Level CNN Features, IEEE Transactions on Image Processing 29 (2019) 3321–3335
2019
Earlier work this paper cites.
S. S. Shivakumar, et al., PST900: RGB-Thermal Calibration, Dataset and Segmentation Network, in: 2020 IEEE International Conference on Robotics and Automation (ICRA), IEEE, 2020, pp. 9441–9447
2020
Earlier work this paper cites.
Y, Sun et al
2020
Earlier work this paper cites.
X. Zhu, et al., Deformable DETR: Deformable Transformers for End-to-End Object Detection, in: International Conference on Learning Representations (ICLR), 2020
2020
Earlier work this paper cites.
N. Carion, et al., End-to-End Object Detection with Transformers, in: European Conference on Computer Vision (ECCV), Springer, 2020, pp. 213–229
2020
Earlier work this paper cites.
X. Chen, et al., Bi-directional Cross-Modality Feature Propagation with Separation-and-Aggregation Gate for RGB-D Semantic Segmentation, in: European Conference on Computer Vision (ECCV), Springer, 2020, pp. 561–577
2020
Earlier work this paper cites.
A. Dosovitskiy, et al., An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale, International Conference on Learning Representations (ICLR) (2020)
2020
Earlier work this paper cites.
Q. Zhang, et al., ABMDRNet: Adaptive-weighted Bi-directional Modality Difference Reduction Network for RGB-T Semantic Segmentation, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2021, pp. 2633–2642
2021
Cited alongside, same era.
W. Zhou et al
2021
Cited alongside, same era.
Z. Liu, et al., Swin Transformer: Hierarchical Vision Transformer using Shifted Windows, in: Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), 2021, pp. 10012–10022
2021
Cited alongside, same era.
F. Deng, et al., FEANet: Feature-Enhanced Attention Network for RGB-Thermal Real-time Semantic Segmentation, in: 2021 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), IEEE, 2021, pp. 4467–4473
2021
Cited alongside, same era.
M. Oquab, et al., DINOv2: Learning Robust Visual Features Without Supervision, Transactions on Machine Learning Research (2023)
2023
Later among the works it cites.
S. Zhao, et al., Mitigating Modality Discrepancies for RGB-T Semantic Segmentation, IEEE Transactions on Neural Networks and Learning SystemsDOI:10.1109/TNNLS.2022.3233089 (2023)
2023
Later among the works it cites.
W. Zhou, et al., DBCNet: Dynamic Bilateral Cross-Fusion Network for RGB-T Urban Scene Understanding in Intelligent Vehicles, IEEE Transactions on Systems, Man, and Cybernetics: Systems 53 (12) (2023) 7631–7641
2023
Later among the works it cites.
J, Zhang et al
2023
Later among the works it cites.
K. Li, et al., UniFormer: Unifying Convolution and Self-Attention for Visual Recognition, IEEE Transactions on Pattern Analysis and Machine Intelligence 45 (10) (2023) 12581–12600
2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2021
Cited alongside, same era.
Y.-H. Kim, et al., MS-UDA: Multi-Spectral Unsupervised Domain Adaptation for Thermal Image Semantic Segmentation, IEEE Robotics and Automation Letters 6 (4) (2021) 6497–6504
2021
Cited alongside, same era.
S. Zheng, et al., Rethinking Semantic Segmentation From a Sequence-to-Sequence Perspective With Transformers, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2021, pp. 6881–6890
2021
Cited alongside, same era.
E. Xie, et al., SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers, Advances in Neural Information Processing Systems (NeurIPS) 34 (2021) 12077–12090
2021
Cited alongside, same era.
B. Cheng, et al., Per-Pixel Classification is Not All You Need for Semantic Segmentation, Advances in Neural Information Processing Systems (NeurIPS) 34 (2021) 17864–17875
2021
Cited alongside, same era.
S. Huang, et al., FaPN: Feature-Aligned Pyramid Network for Dense Image Prediction, in: Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), 2021, pp. 864–873
2021
Cited alongside, same era.
H. Touvron, et al., Training data-efficient image transformers & distillation through attention, in: International Conference on Machine Learning (ICML), PMLR, 2021, pp. 10347–10357
2021
Cited alongside, same era.
Z. Liu et al
2022
Cited alongside, same era.
Later among the works it cites.
Z. Chen, et al., Vision Transformer Adapter for Dense Predictions, in: The Eleventh International Conference on Learning Representations (ICLR), 2023
2023
Later among the works it cites.
W. Zhou et al
2023
Later among the works it cites.
X. He, et al., SFAF-MA: Spatial Feature Aggregation and Fusion With Modality Adaptation for RGB-Thermal Semantic Segmentation, IEEE Transactions on Instrumentation and Measurement 72 (2023) 1–10
2023
Later among the works it cites.
U. Shin, et al., Complementary Random Masking for RGB-Thermal Semantic Segmentation, in: 2024 IEEE International Conference on Robotics and Automation (ICRA), IEEE, 2024, pp. 11110–11117
2024
Closest in time.
Y. Lv, et al., Context-Aware Interaction Network for RGB-T Semantic Segmentation, IEEE Transactions on MultimediaDOI:10.1109/TMM.2023.3349072 (2024)
2024
Closest in time.
doi:10.1109/TIV.2024.3357056
Z. Wu, et al., S 3 · 2024
Closest in time.
S. Sun, et al., ReMaX: Relaxing for Better Training on Efficient Panoptic Segmentation, Advances in Neural Information Processing Systems (NeurIPS) 36 (2024)
2024
Closest in time.
B. Yin, et al., DFormer: Rethinking RGBD Representation Learning for Semantic Segmentation, International Conference on Learning Representations (ICLR) (2024)
2024
Closest in time.
J. Li, et al., RoadFormer: Duplex Transformer for RGB-Normal Semantic Road Scene Parsing, IEEE Transactions on Intelligent Vehicles 9 (7) (2024) 5163–5172
2024
Closest in time.
X. Yang, et al., PolyMaX: General Dense Prediction with Mask Transformer, in: Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV), 2024, pp. 1050–1061
2024
Closest in time.
S. Srivastava, G. Sharma, OmniVec: Learning Robust Representations With Cross Modal Sharing, in: Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV), 2024, pp. 1236–1248
2024
Closest in time.
S. Du, et al., AsymFormer: Asymmetrical Cross-Modal Representation Learning for Mobile Platform Real-Time RGB-D Semantic Segmentation, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2024, pp. 7608–7615
2024
Closest in time.
doi:10.1109/TIP.2025.3618378
S. Guo, et al., LIX: Implicitly infusing spatial geometric prior knowledge into visual semantic segmentation for autonomous driving, IEEE Transactions on Image Processing 34 (2025) 7250–7263 · 2025
Closest in time.
doi:10.1109/LSP.2025.3575640
J. Huang, et al., DepthMatch: Semi-supervised RGB-D scene parsing through depth-guided regularization, IEEE Signal Processing Letters 32 (2025) 2549–2553 · 2025
Closest in time.
doi:10.1109/TASE.2025.3586286
G. Tang, et al., TiCoSS: Tightening the coupling between semantic segmentation and stereo matching within a joint learning framework, IEEE Transactions on Automation Science and Engineering 22 (2025) 18646–18658 · 2025
Closest in time.
doi:10.1109/TIM.2025.3579733
M.-J. Lee, et al., SG-RoadSeg+: End-to-end freespace detection upgraded at data, feature, and loss levels, IEEE Transactions on Instrumentation and Measurement 74 (2025) 1–9 · 2025
Closest in time.
Y. Feng, et al., SNE-RoadSegV2: Advancing Heterogeneous Feature Fusion and Fallibility Awareness for Freespace Detection, IEEE Transactions on Instrumentation and Measurement 74 (2025) 1–9
2025
Closest in time.
J. Huang, et al., RoadFormer+: Delivering RGB-X Scene Parsing through Scale-Aware Information Decoupling and Advanced Heterogeneous Feature Fusion, IEEE Transactions on Intelligent Vehicles 10 (5) (2025) 3156–3165
2025
Closest in time.
X. Guo, et al., Low-Light Enhancement and Global-Local Feature Interaction for RGB-T Semantic Segmentation, IEEE Transactions on Instrumentation and Measurement (2025)
2025
Closest in time.