Fetching the paper…
Reading the bibliography…
Research has focused on Multi-Modal Semantic Segmentation (MMSS), where pixel-wise predictions are derived from multiple visual modalities captured by diverse sensors.
Visualizing data using t-SNE
Laurens Van der Maaten and Geoffrey Hinton · 2008
Earlier work this paper cites.
Scikit-learn: Machine learning in python
Fabian Pedregosa, Gaël Varoquaux, Alexandre Gramfort, Vincent Michel, Bertrand Thirion, Olivier Grisel, Mathieu Blondel, Peter Prettenhofer, Ron Weiss, Vincent Dubourg, Jake Vanderplas, Alexandre Passos, David Cournapeau, Matthieu Brucher, Matthieu Perrot, and Édouard Duchesnay · 2011
Earlier work this paper cites.
ENet: A deep neural network architecture for real-time semantic segmentation
Adam Paszke, Abhishek Chaurasia, Sangpil Kim, and Eugenio Culurciello · 2016
Earlier work this paper cites.
Training region-based object detectors with online hard example mining
Abhinav Shrivastava, Abhinav Gupta, and Ross Girshick · 2016
Earlier work this paper cites.
CARLA: An open urban driving simulator
Alexey Dosovitskiy, German Ros, Felipe Codevilla, Antonio Lopez, and Vladlen Koltun · 2017
Earlier work this paper cites.
Robust classification with convolutional prototype learning
Hong-Ming Yang, Xu-Yao Zhang, Fei Yin, and Cheng-Lin Liu · 2018
Earlier work this paper cites.
BiSeNet: Bilateral segmentation network for real-time semantic segmentation
Changqian Yu, Jingbo Wang, Chao Peng, Changxin Gao, Gang Yu, and Nong Sang · 2018
Earlier work this paper cites.
RGB and LiDAR fusion based 3D semantic segmentation for autonomous driving
Khaled El Madawi, Hazem Rashed, Ahmad El Sallab, Omar Nasr, Hanan Kamel, and Senthil Yogamani · 2019
Earlier work this paper cites.
Parameter-efficient transfer learning for NLP
Neil Houlsby, Andrei Giurgiu, Stanislaw Jastrzebski, Bruna Morrone, Quentin De Laroussilhe, Andrea Gesmundo, Mona Attariyan, and Sylvain Gelly · 2019
Earlier work this paper cites.
Fast-SCNN: Fast semantic segmentation network
Rudra P. K. Poudel, Stephan Liwicki, and Roberto Cipolla · 2019
Earlier work this paper cites.
Bi-directional cross-modality feature propagation with separation-and-aggregation gate for RGB-D semantic segmentation
Xiaokang Chen, Kwan-Yee Lin, Jingbo Wang, Wayne Wu, Chen Qian, Hongsheng Li, and Gang Zeng · 2020
Earlier work this paper cites.
ShapeConv: Shape-aware convolutional layer for indoor RGB-D semantic segmentation
Jinming Cao, Hanchao Leng, Dani Lischinski, Daniel Cohen-Or, Changhe Tu, and Yangyan Li · 2021
Earlier work this paper cites.
Spatial information guided convolution for real-time RGBD semantic segmentation
Lin-Zhuo Chen, Zheng Lin, Ziqin Wang, Yong-Liang Yang, and Ming-Ming Cheng · 2021
Earlier work this paper cites.
FEANet: Feature-enhanced attention network for RGB-thermal real-time semantic segmentation
Fuqin Deng, Hua Feng, Mingjian Liang, Hongmin Wang, Yong Yang, Yuan Gao, Junfeng Chen, Junjie Hu, Xiyue Guo, and Tin Lun Lam · 2021
Earlier work this paper cites.
Rethinking BiSeNet for real-time semantic segmentation
Mingyuan Fan, Shenqi Lai, Junshi Huang, Xiaoming Wei, Zhenhua Chai, Junfeng Luo, and Xiaolin Wei · 2021
Earlier work this paper cites.
ISNet: Integrate image-level and semantic-level context for semantic segmentation
Zhenchao Jin, Bin Liu, Qi Chu, and Nenghai Yu · 2021
Earlier work this paper cites.
Benchmarking low-light image enhancement and beyond
Jiaying Liu, Dejia Xu, Wenhan Yang, Minhao Fan, and Haofeng Huang · 2021
Earlier work this paper cites.
Efficient RGB-D semantic segmentation for indoor scene analysis
Daniel Seichter, Mona Köhler, Benjamin Lewandowski, Tim Wengefeld, and Horst-Michael Gross · 2021
Earlier work this paper cites.
SegFormer: Simple and efficient design for semantic segmentation with transformers
Enze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar, Jose M. Alvarez, and Ping Luo · 2021
Earlier work this paper cites.
BiSeNet V2: Bilateral network with guided aggregation for real-time semantic segmentation
Changqian Yu, Changxin Gao, Jingbo Wang, Gang Yu, Chunhua Shen, and Nong Sang · 2021
Earlier work this paper cites.
LoRA: Low-rank adaptation of large language models
Edward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen · 2022
Earlier work this paper cites.
Multimodal material segmentation
Yupeng Liang, Ryosuke Wakaki, Shohei Nobuhara, and Ko Nishino · 2022
Cited alongside, same era.
Review the state-of-the-art technologies of semantic segmentation based on deep learning
Yujian Mo, Yan Wu, Xinneng Yang, Feilin Liu, and Yujun Liao · 2022
Cited alongside, same era.
PP-LiteSeg: A superior real-time semantic segmentation model
Juncai Peng, Yi Liu, Shiyu Tang, Yuying Hao, Lutao Chu, Guowei Chen, Zewu Wu, Zeyu Chen, Zhiliang Yu, Yuning Du, Qingqing Dang, Baohua Lai, Qiwen Liu, Xiaoguang Hu, Dianhai Yu, and Yanjun Ma · 2022
Cited alongside, same era.
RTFormer: Efficient design for real-time semantic segmentation with transformer
Jian Wang, Chenhui Gou, Qiman Wu, Haocheng Feng, Junyu Han, Errui Ding, and Jingdong Wang · 2022
Cited alongside, same era.
Complementarity-aware cross-modal feature fusion network for RGB-T semantic segmentation
Wei Wu, Tao Chu, and Qiong Liu · 2022
Cited alongside, same era.
Chasing day and night: Towards robust and efficient all-day object detection guided by an event camera
Jiahang Cao, Xu Zheng, Yuanhuiyi Lyu, Jiaxu Wang, Renjing Xu, and Lin Wang · 2024
Later among the works it cites.
SAM2Long: Enhancing SAM 2 for long video segmentation with a training-free memory tree
Shuangrui Ding, Rui Qian, Xiaoyi Dong, Pan Zhang, Yuhang Zang, Yuhang Cao, Yuwei Guo, Dahua Lin, and Jiaqi Wang · 2024
Later among the works it cites.
CoSAM: Self-correcting SAM for domain generalization in 2D medical image segmentation
Yihang Fu, Ziyang Chen, Yiwen Ye, Xingliang Lei, Zhisong Wang, and Yong Xia · 2024
Later among the works it cites.
Relax image-specific prompt requirement in SAM: A single generic prompt for segmenting camouflaged objects
Jian Hu, Jiayi Lin, Shaogang Gong, and Weitong Cai · 2024
Later among the works it cites.
SAM-I2I: Unleash the power of segment anything model for medical image translation
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Attribute-based progressive fusion network for RGBT tracking
Yun Xiao, Mengmeng Yang, Chenglong Li, Lei Liu, and Jin Tang · 2022
Cited alongside, same era.
Window attention is bugged: How not to interpolate position embeddings
Daniel Bolya, Chaitanya Ryali, Judy Hoffman, and Christoph Feichtenhofer · 2023
Cited alongside, same era.
X-Align: Cross-modal cross-view alignment for bird’s-eye-view segmentation
Shubhankar Borse, Marvin Klingner, Varun Ravi Kumar, Hong Cai, Abdulaziz Almuzairee, Senthil Yogamani, and Fatih Porikli · 2023
Cited alongside, same era.
SAM-Adapter: Adapting segment anything in underperformed scenes
Tianrun Chen, Lanyun Zhu, Chaotao Deng, Runlong Cao, Yan Wang, Shangzhan Zhang, Zejian Li, Lingyun Sun, Ying Zang, and Papa Mao · 2023
Cited alongside, same era.
Prototype and context-enhanced learning for unsupervised domain adaptation semantic segmentation of remote sensing images
Kuiliang Gao, Anzhu Yu, Xiong You, Chunping Qiu, and Bing Liu · 2023
Cited alongside, same era.
Bimodal SegNet: Instance segmentation fusing events and RGB frames for robotic grasping
Sanket Kachole, Xiaoqian Huang, Fariborz Baghaei Naeini, Rajkumar Muthusamy, Dimitrios Makris, and Yahya Zweiri · 2023
Cited alongside, same era.
Segment anything
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloé Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C. Berg, Wan-Yen Lo, Piotr Dollár, and Ross B. Girshick · 2023
Cited alongside, same era.
Jiayu Huo, Sebastien Ourselin, and Rachel Sparks · 2024
Later among the works it cites.
Xinyang Pu, Hecheng Jia, Linghao Zheng, Feng Wang, and Feng Xu · 2024
Later among the works it cites.
DB-SAM: Delving into high quality universal medical image segmentation
Chao Qin, Jiale Cao, Huazhu Fu, Fahad Shahbaz Khan, and Rao Muhammad Anwer · 2024
Later among the works it cites.
SAM 2: Segment anything in images and videos
Nikhila Ravi, Valentin Gabeur, Yuan-Ting Hu, Ronghang Hu, Chaitanya Ryali, Tengyu Ma, Haitham Khedr, Roman Rädle, Chloé Rolland, Laura Gustafson, Eric Mintun, Junting Pan, Kalyan Vasudev Alwala, Nicolas Carion, Chao-Yuan Wu, Ross B. Girshick, Piotr Dollár, and Christoph Feichtenhofer · 2024
Later among the works it cites.
Real-time semantic segmentation for underground mine tunnel
Jiawen Wang, Dewei Li, Qihang Long, Zhongqi Zhao, Xuan Gao, Jingchuan Chen, and Kehu Yang · 2024
Later among the works it cites.
Segment anything with multiple modalities
Aoran Xiao, Weihao Xuan, Heli Qi, Yun Xing, Naoto Yokoya, and Shijian Lu · 2024
Later among the works it cites.
EISNet: A multi-modal fusion network for semantic segmentation with events and images
Bochen Xie, Yongjian Deng, Zhanpeng Shao, and Youfu Li · 2024
Later among the works it cites.
SAMURAI: Adapting segment anything model for zero-shot visual tracking with motion-aware memory
Cheng-Yen Yang, Hsiang-Wei Huang, Wenhao Chai, Zhongyu Jiang, and Jenq-Neng Hwang · 2024
Later among the works it cites.
SAM-Event-Adapter: Adapting segment anything model for event-RGB semantic segmentation
Bowen Yao, Yongjian Deng, Yuhan Liu, Hao Chen, Youfu Li, and Zhen Yang · 2024
Later among the works it cites.
DFormer: Rethinking RGBD representation learning for semantic segmentation
Bowen Yin, Xuying Zhang, Zhongyu Li, Li Liu, Ming-Ming Cheng, and Qibin Hou · 2024
Later among the works it cites.
A generative adversarial network approach for removing motion blur in the automatic detection of pavement cracks
Yu Zhang and Lin Zhang · 2024
Later among the works it cites.
Mitigating modality discrepancies for RGB-T semantic segmentation
Shenlu Zhao, Yichen Liu, Qiang Jiao, Qiang Zhang, and Jungong Han · 2024
Later among the works it cites.
ExACT: Language-guided conceptual reasoning and uncertainty estimation for event-based action recognition and more
Jiazhou Zhou, Xu Zheng, Yuanhuiyi Lyu, and Lin Wang · 2024
Later among the works it cites.
Customize segment anything model for multi-modal semantic segmentation with mixture of LoRA experts
Chenyang Zhu, Bin Xiao, Lin Shi, Shoukun Xu, and Xu Zheng · 2024
Later among the works it cites.
Class incremental learning with self-supervised pre-training and prototype learning
Wenzhuo Liu, Xin-Jian Wu, Fei Zhu, Ming-Ming Yu, Chuang Wang, and Cheng-Lin Liu · 2025
Closest in time.
Jiayi Zhao, Fei Teng, Kai Luo, Guoqiang Zhao, Zhiyong Li, Xu Zheng, and Kailun Yang · 2025
Closest in time.
360SFUDA++: Towards source-free UDA for panoramic segmentation by learning reliable category prototypes
Xu Zheng, Peng Yuan Zhou, Athanasios V. Vasilakos, and Lin Wang · 2025
Closest in time.