Fetching the paper…
Reading the bibliography…
Integrating information from multiple modalities enhances the robustness of scene perception systems in autonomous vehicles, providing a more comprehensive and reliable sensory framework.
M. Cordts et al. , “The cityscapes dataset for semantic urban scene understanding,” in CVPR , 2016
2016
Earlier work this paper cites.
X. Hu, K. Yang, L. Fei, and K. Wang, “ACNet: Attention based network to exploit complementary features for RGBD semantic segmentation,” in ICIP , 2019
2019
Earlier work this paper cites.
Y. Choukroun, E. Kravchik, F. Yang, and P. Kisilev, “Low-bit quantization of neural networks for efficient inference,” in ICCVW , 2019
2019
Earlier work this paper cites.
X. Chen et al. , “Bi-directional cross-modality feature propagation with separation-and-aggregation gate for RGB-D semantic segmentation,” in ECCV , 2020
2020
Earlier work this paper cites.
M. Ma, J. Ren, L. Zhao, S. Tulyakov, C. Wu, and X. Peng, “SMIL: Multimodal earning with severely missing modality,” in AAAI , 2021
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
W. Zhou, J. Liu, J. Lei, L. Yu, and J.-N. Hwang, “GMNet: Graded-feature multilabel-learning network for RGB-thermal urban scene semantic segmentation,” TIP , 2021
2021
Earlier work this paper cites.
Q. Zhang, S. Zhao, Y. Luo, D. Zhang, N. Huang, and J. Han, “ABMDRNet: Adaptive-weighted bi-directional modality difference reduction network for RGB-T semantic segmentation,” in CVPR , 2021
2021
Earlier work this paper cites.
K. Xiang, K. Yang, and K. Wang, “Polarization-driven semantic segmentation via efficient attention-bridged fusion,” OE , 2021
2021
Earlier work this paper cites.
R. Yan, K. Yang, and K. Wang, “NLFNet: Non-local fusion towards generalized multimodal semantic segmentation across RGB-depth, polarization, and thermal images,” in ROBIO , 2021
2021
Earlier work this paper cites.
J. Zhang, K. Yang, and R. Stiefelhagen, “ISSAFE: Improving semantic segmentation in accidents by fusing event-based data,” in IROS , 2021
2021
Earlier work this paper cites.
A. Dosovitskiy et al. , “An image is worth 16x16 words: Transformers for image recognition at scale,” in ICLR , 2021
2021
Earlier work this paper cites.
R. Bachmann, D. Mizrahi, A. Atanov, and A. Zamir, “MultiMAE: Multi-modal multi-task masked autoencoders,” in ECCV , 2022
2022
Earlier work this paper cites.
M. Jia et al. , “Visual prompt tuning,” in ECCV , 2022
2022
Earlier work this paper cites.
J. Lee-Thorp, J. Ainslie, I. Eckstein, and S. Ontanon, “FNet: Mixing tokens with fourier transforms,” in NAACL-HLT , 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
S. Chen et al. , “AdaptFormer: Adapting vision transformers for scalable visual recognition,” in NeurIPS , 2022
2022
Cited alongside, same era.
J. Zhang, K. Yang, and R. Stiefelhagen, “Exploring event-driven dynamic context for accident scene segmentation,” T-ITS , 2022
2022
Cited alongside, same era.
R. Girdhar, M. Singh, N. Ravi, L. van der Maaten, A. Joulin, and I. Misra, “Omnivore: A single model for many visual modalities,” in CVPR , 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
D. Lian, D. Zhou, J. Feng, and X. Wang, “Scaling & shifting your features: A new baseline for efficient model tuning,” NeurIPS , 2022
2022
Cited alongside, same era.
H. Wang et al. , “Learnable cross-modal knowledge distillation for multi-modal learning with missing modality,” in MICCAI , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2022
Cited alongside, same era.
J. Zhang, H. Liu, K. Yang, X. Hu, R. Liu, and R. Stiefelhagen, “CMX: Cross-modal fusion for RGB-X semantic segmentation with transformers,” T-ITS , 2023
2023
Cited alongside, same era.
J. Zhang, R. Liu, H. Shi, K. Yang, S. Reiß, K. Peng, H. Fu, K. Wang, and R. Stiefelhagen, “Delivering arbitrary-modal semantic segmentation,” in CVPR , 2023
2023
Cited alongside, same era.
T. Broedermann, C. Sakaridis, D. Dai, and L. Van Gool, “HRFuser: A multi-resolution sensor fusion architecture for 2D object detection,” in ITSC , 2023
2023
Cited alongside, same era.
Y.-L. Lee, Y.-H. Tsai, W.-C. Chiu, and C.-Y. Lee, “Multimodal prompting with missing modalities for visual recognition,” in CVPR , 2023
2023
Cited alongside, same era.
H. Wang, Y. Chen, C. Ma, J. Avery, L. Hull, and G. Carneiro, “Multi-modal learning with missing modality via shared-specific feature modelling,” in CVPR , 2023
2023
Cited alongside, same era.
C. Ge et al. , “MetaBEV: Solving sensor failures for 3D detection and map segmentation,” in ICCV , 2023
2023
Cited alongside, same era.
2023
Later among the works it cites.
Z. Jiawen, l. Simiao, C. Xin, D. Wang, and H. Lu, “Visual prompt multi-modal tracking,” in CVPR , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
Q. Xu, Y. Li, J. Shen, J. K. Liu, H. Tang, and G. Pan, “Constructing deep spiking neural networks from artificial neural networks with knowledge distillation,” in CVPR , 2023
2023
Later among the works it cites.
B. Rokh, A. Azarpeyvand, and A. Khanteymoori, “A comprehensive survey on model quantization for deep neural networks in image classification,” TIST , 2023
2023
Later among the works it cites.
A. Aich, S. Schulter, A. K. Roy-Chowdhury, M. Chandraker, and Y. Suh, “Efficient controllable multi-task architectures,” in CVPR , 2023
2023
Later among the works it cites.
X. Jia et al. , “Fourier-Net: Fast image registration with band-limited deformation,” in AAAI , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
X. Nie et al. , “Pro-tuning: Unified prompt tuning for vision tasks,” TCVST , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
S. Srivastava and G. Sharma, “OmniVec: Learning robust representations with cross modal sharing,” in WACV , 2024
2024
Closest in time.