Fetching the paper…
Reading the bibliography…
Fusing and balancing multi-modal inputs from novel sensors for dense prediction tasks, particularly semantic segmentation, is critically important yet remains a significant challenge.
I. Alonso and A. C. Murillo, “Ev-segnet: Semantic segmentation for event-based cameras,” in
2019
Earlier work this paper cites.
I. Gat, I. Schwartz, A. Schwing, and T. Hazan, “Removing bias in multi-modal classifiers: Regularization by maximizing functional entropies,”
2020
Earlier work this paper cites.
Y. Wang, F. Sun, M. Lu, and A. Yao, “Learning deep multimodal feature representation with asymmetric multi-layer fusion,” in
2020
Earlier work this paper cites.
H. Zhou, L. Qi, Z. Wan, H. Huang, and X. Yang, “Rgb-d co-attention network for semantic segmentation,” in
2020
Earlier work this paper cites.
Y. Wang, W. Huang, F. Sun, T. Xu, Y. Rong, and J. Huang, “Deep multimodal fusion by channel exchanging,”
2020
Earlier work this paper cites.
S. S. Shivakumar, N. Rodrigues, A. Zhou, I. D. Miller, V. Kumar, and C. J. Taylor, “Pst900: Rgb-thermal calibration, dataset and segmentation network,” in
2020
Earlier work this paper cites.
E. Xie, W. Wang, Z. Yu, A. Anandkumar, J. M. Alvarez, and P. Luo, “Segformer: Simple and efficient design for semantic segmentation with transformers,”
2021
Earlier work this paper cites.
J. Cao, H. Leng, D. Lischinski, D. Cohen-Or, C. Tu, and Y. Li, “Shapeconv: Shape-aware convolutional layer for indoor rgb-d semantic segmentation,” in
2021
Earlier work this paper cites.
L.-Z. Chen, Z. Lin, Z. Wang, Y.-L. Yang, and M.-M. Cheng, “Spatial information guided convolution for real-time rgbd semantic segmentation,”
2021
Earlier work this paper cites.
Q. Zhang, S. Zhao, Y. Luo, D. Zhang, N. Huang, and J. Han, “Abmdrnet: Adaptive-weighted bi-directional modality difference reduction network for rgb-t semantic segmentation,” in
2021
Earlier work this paper cites.
J. Zhang, K. Yang, and R. Stiefelhagen, “Issafe: Improving semantic segmentation in accidents by fusing event-based data,” in
2021
Earlier work this paper cites.
Z. Zhuang, R. Li, K. Jia, Q. Wang, Y. Li, and M. Tan, “Perception-aware multi-sensor fusion for 3d lidar semantic segmentation,” in
2021
Earlier work this paper cites.
X. Zheng, C. Fu, H. Xie, J. Chen, X. Wang, and C. Sham, “Uncertainty-aware deep co-training for semi-supervised medical image segmentation,”
2022
Earlier work this paper cites.
J. Chen, C. Fu, H. Xie, X. Zheng, R. Geng, and C. Sham, “Uncertainty teacher with dense focal loss for semi-supervised medical image segmentation,”
2022
Earlier work this paper cites.
X. Ying and M. C. Chuah, “Uctnet: Uncertainty-aware cross-modal transformer network for indoor rgb-d semantic segmentation,” in
2022
Earlier work this paper cites.
M. Lee, C. Park, S. Cho, and S. Lee, “Spsn: Superpixel prototype sampling network for rgb-d salient object detection,” in
2022
Earlier work this paper cites.
R. Cong, Q. Lin, C. Zhang, C. Li, X. Cao, Q. Huang, and Y. Zhao, “Cir-net: Cross-modality interaction and refinement for rgb-d salient object detection,”
2022
Earlier work this paper cites.
W. Ji, G. Yan, J. Li, Y. Piao, S. Yao, M. Zhang, L. Cheng, and H. Lu, “Dmra: Depth-induced multi-scale recurrent attention network for rgb-d saliency detection,”
2022
Earlier work this paper cites.
F. Wang, J. Pan, S. Xu, and J. Tang, “Learning discriminative cross-modality features for rgb-d saliency detection,”
2022
Earlier work this paper cites.
M. Song, W. Song, G. Yang, and C. Chen, “Improving rgb-d salient object detection via modality-aware decoder,”
2022
Earlier work this paper cites.
W. Wu, T. Chu, and Q. Liu, “Complementarity-aware cross-modal feature fusion network for rgb-t semantic segmentation,”
2022
Earlier work this paper cites.
G. Liao, W. Gao, G. Li, J. Wang, and S. Kwong, “Cross-collaborative fusion-encoder network for robust rgb-thermal salient object detection,”
2022
Earlier work this paper cites.
G. Chen, F. Shao, X. Chai, H. Chen, Q. Jiang, X. Meng, and Y.-S. Ho, “Modality-induced transfer-fusion network for rgb-d and rgb-t salient object detection,”
2022
Earlier work this paper cites.
X. Yan, J. Gao, C. Zheng, C. Zheng, R. Zhang, S. Cui, and Z. Li, “2dpass: 2d priors assisted semantic segmentation on lidar point clouds,” in
2022
Earlier work this paper cites.
Y. Wang, X. Chen, L. Cao, W. Huang, F. Sun, and Y. Wang, “Multimodal token fusion for vision transformers,” in
2022
Earlier work this paper cites.
Y. Li, A. W. Yu, T. Meng, B. Caine, J. Ngiam, D. Peng, J. Shen, Y. Lu, D. Zhou, Q. V. Le,
2022
Cited alongside, same era.
H. Liu, T. Lu, Y. Xu, J. Liu, W. Li, and L. Chen, “Camliflow: Bidirectional camera-lidar fusion for joint optical flow and scene flow estimation,” in
2022
Cited alongside, same era.
Y. Liang, R. Wakaki, S. Nobuhara, and K. Nishino, “Multimodal material segmentation,” in
2022
Cited alongside, same era.
2022
Cited alongside, same era.
R. Bachmann, D. Mizrahi, A. Atanov, and A. Zamir, “Multimae: Multi-modal multi-task masked autoencoders,” in
2022
Cited alongside, same era.
2024
Later among the works it cites.
X. Zheng, P. Zhou, A. V. Vasilakos, and L. Wang, “Semantics, distortion, and style matter: Towards source-free UDA for panoramic segmentation,” in
2024
Later among the works it cites.
B. Ren, Y. Li, J. Liang, R. Ranjan, M. Liu, R. Cucchiara, L. V. Gool, M.-H. Yang, and N. Sebe, “Sharing key semantics in transformer makes efficient image restoration,”
2024
Later among the works it cites.
B. Ren, G. Mei, D. P. Paudel, W. Wang, Y. Li, M. Liu, R. Cucchiara, L. Van Gool, and N. Sebe, “Bringing masked autoencoders explicit contrastive properties for point cloud self-supervised learning,” in
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Zhang, R. Liu, H. Shi, K. Yang, S. Reiß, K. Peng, H. Fu, K. Wang, and R. Stiefelhagen, “Delivering arbitrary-modal semantic segmentation,” in
2023
Cited alongside, same era.
2023
Cited alongside, same era.
H. Xie, C. Fu, X. Zheng, Y. Zheng, C. Sham, and X. Wang, “Adversarial co-training for semantic segmentation over medical images,”
2023
Cited alongside, same era.
B. Ren, Y. Liu, Y. Song, W. Bi, R. Cucchiara, N. Sebe, and W. Wang, “Masked jigsaw puzzle: A versatile position embedding for vision transformers,” in
2023
Cited alongside, same era.
2023
Cited alongside, same era.
W. Zhou, H. Zhang, W. Yan, and W. Lin, “Mmsmcnet: Modal memory sharing and morphological complementary networks for rgb-t urban scene semantic segmentation,”
2023
Cited alongside, same era.
Z. Xie, F. Shao, G. Chen, H. Chen, Q. Jiang, X. Meng, and Y.-S. Ho, “Cross-modality double bidirectional interaction and fusion network for rgb-t salient object detection,”
2023
Cited alongside, same era.
2024
Later among the works it cites.
Y. Zhang, P. E. Latham, and A. M. Saxe, “Understanding unimodal bias in multimodal deep linear networks,” in
2024
Later among the works it cites.
2024
Later among the works it cites.
P. Wang, S. Bai, S. Tan, S. Wang, Z. Fan, J. Bai, K. Chen, X. Liu, J. Wang, W. Ge,
2024
Later among the works it cites.
2024
Later among the works it cites.
Y. Lyu, X. Zheng, J. Zhou, and L. Wang, “Unibind: Llm-augmented unified and balanced representation space to bind them all,” in
2024
Later among the works it cites.
2024
Later among the works it cites.
X. Zheng and L. Wang, “Eventdance: Unsupervised source-free cross-modal adaptation for event-based object recognition,” in
2024
Later among the works it cites.
J. Zhou, X. Zheng, Y. Lyu, and L. Wang, “Exact: Language-guided conceptual reasoning and uncertainty estimation for event-based action recognition and more,” in
2024
Later among the works it cites.
W. Zhang, Y. Liu, X. Zheng, and L. Wang, “Goodsam: Bridging domain and capacity gaps via segment anything model for distortion-aware panoramic semantic segmentation,” in
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
R. Liu, J. Zhang, K. Peng, Y. Chen, K. Cao, J. Zheng, M. S. Sarfraz, K. Yang, and R. Stiefelhagen, “Fourier prompt tuning for modality-incomplete scene segmentation,” in
2024
Later among the works it cites.
T. Brödermann, D. Bruggemann, C. Sakaridis, K. Ta, O. Liagouris, J. Corkill, and L. Van Gool, “Muses: The multi-sensor semantic perception dataset for driving under uncertainty,” in
2024
Later among the works it cites.
X. Zheng, Y. Lyu, J. Zhou, and L. Wang, “Centering the value of every modality: Towards efficient and resilient modality-agnostic semantic segmentation,” in
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
T. Brödermann, D. Bruggemann, C. Sakaridis, K. Ta, O. Liagouris, J. Corkill, and L. Van Gool, “Muses: The multi-sensor semantic perception dataset for driving under uncertainty,” in
2025
Closest in time.