Fetching the paper…
Reading the bibliography…
Mamba, a recent selective structured state space model, excels in long sequence modeling, which is vital in the large model era.
A method of analysing the behaviour of linear systems in terms of time series
Arnold Tustin. 1947 · 1947
Earlier work this paper cites.
A New Approach to Linear Filtering and Prediction Problems
Rudolph Emil Kalman. 1960 · 1960
Earlier work this paper cites.
Dynamic causal modelling
Karl J. Friston, Lee M. Harrison, and William D. Penny. 2003 · 2003
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database. In CVPR . 248–255
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei. 2009 · 2009
Earlier work this paper cites.
ImageNet Classification with Deep Convolutional Neural Networks. In NeurIPS . 1106–1114
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E. Hinton. 2012 · 2012
Earlier work this paper cites.
Microsoft COCO: Common Objects in Context. In ECCV (Lecture Notes in Computer Science, Vol. 8693) . Springer, 740–755
Tsung-Yi Lin, Michael Maire, Serge J. Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C. Lawrence Zitnick. 2014 · 2014
Earlier work this paper cites.
Very Deep Convolutional Networks for Large-Scale Image Recognition. In ICLR
Karen Simonyan and Andrew Zisserman. 2015 · 2015
Earlier work this paper cites.
Deep Residual Learning for Image Recognition. In CVPR . 770–778
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Earlier work this paper cites.
Mask R-CNN. In ICCV . 2980–2988
Kaiming He, Georgia Gkioxari, Piotr Dollár, and Ross B. Girshick. 2017 · 2017
Earlier work this paper cites.
Neural Discrete Representation Learning. In NeurIPS . 6306–6315
Aäron van den Oord, Oriol Vinyals, and Koray Kavukcuoglu. 2017 · 2017
Earlier work this paper cites.
Attention is All you Need. In NeurIPS . 5998–6008
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Unified Perceptual Parsing for Scene Understanding. In ECCV , Vol. 11209. Springer, 432–448
Tete Xiao, Yingcheng Liu, Bolei Zhou, Yuning Jiang, and Jian Sun. 2018 · 2018
Earlier work this paper cites.
Semantic Understanding of Scenes Through the ADE20K Dataset
Bolei Zhou, Hang Zhao, Xavier Puig, Tete Xiao, Sanja Fidler, Adela Barriuso, and Antonio Torralba. 2019 · 2019
Earlier work this paper cites.
HiPPO: Recurrent Memory with Optimal Polynomial Projections. In NeurIPS
Albert Gu, Tri Dao, Stefano Ermon, Atri Rudra, and Christopher Ré. 2020 · 2020
Earlier work this paper cites.
Dream to Control: Learning Behaviors by Latent Imagination. In ICLR
Danijar Hafner, Timothy P. Lillicrap, Jimmy Ba, and Mohammad Norouzi. 2020 · 2020
Earlier work this paper cites.
Denoising Diffusion Probabilistic Models. In NeurIPS
Jonathan Ho, Ajay Jain, and Pieter Abbeel. 2020 · 2020
Earlier work this paper cites.
Transformers are RNNs: Fast Autoregressive Transformers with Linear Attention. In ICML , Vol. 119. 5156–5165
Angelos Katharopoulos, Apoorv Vyas, Nikolaos Pappas, and François Fleuret. 2020 · 2020
Earlier work this paper cites.
Designing Network Design Spaces. In CVPR . 10425–10433
Ilija Radosavovic, Raj Prateek Kosaraju, Ross B. Girshick, Kaiming He, and Piotr Dollár. 2020 · 2020
Earlier work this paper cites.
Emerging Properties in Self-Supervised Vision Transformers. In ICCV . 9630–9640
Mathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou, Julien Mairal, Piotr Bojanowski, and Armand Joulin. 2021 · 2021
Earlier work this paper cites.
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale. In ICLR
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, et al · 2021
Earlier work this paper cites.
Combining Recurrent, Convolutional, and Continuous-time Models with Linear State Space Layers. In NeurIPS . 572–585
Albert Gu, Isys Johnson, Karan Goel, Khaled Saab, Tri Dao, Atri Rudra, and Christopher Ré. 2021 · 2021
Earlier work this paper cites.
Swin Transformer: Hierarchical Vision Transformer using Shifted Windows. In ICCV . IEEE, 9992–10002
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo. 2021 · 2021
Earlier work this paper cites.
NormFormer: Improved Transformer Pretraining with Extra Normalization
Sam Shleifer, Jason Weston, and Myle Ott. 2021 · 2021
Earlier work this paper cites.
Training data-efficient image transformers & distillation through attention. In ICML , Vol. 139. 10347–10357
Hugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa, Alexandre Sablayrolles, and Hervé Jégou. 2021 · 2021
Earlier work this paper cites.
Multi-Scale Vision Longformer: A New Vision Transformer for High-Resolution Image Encoding. In iccv21/vil . 2978–2988
Pengchuan Zhang, Xiyang Dai, Jianwei Yang, Bin Xiao, Lu Yuan, Lei Zhang, and Jianfeng Gao. 2021 · 2021
Earlier work this paper cites.
Sharpness-Aware Training for Free. In NeurIPS
Jiawei Du, Daquan Zhou, Jiashi Feng, Vincent Y. F. Tan, and Joey Tianyi Zhou. 2022 · 2022
Earlier work this paper cites.
Diagonal State Spaces are as Effective as Structured State Spaces. In NeurIPS
Ankit Gupta, Albert Gu, and Jonathan Berant. 2022 · 2022
Earlier work this paper cites.
A ConvNet for the 2020s. In CVPR . 11966–11976
Zhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer, Trevor Darrell, and Saining Xie. 2022 · 2022
Earlier work this paper cites.
High-Resolution Image Synthesis with Latent Diffusion Models. In CVPR . 10674–10685
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer. 2022 · 2022
Earlier work this paper cites.
DeiT III: Revenge of the ViT. In ECCV , Vol. 13684. Springer, 516–533
Hugo Touvron, Matthieu Cord, and Hervé Jégou. 2022 · 2022
Earlier work this paper cites.
PVT v2: Improved baselines with Pyramid Vision Transformer
Wenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan, Kaitao Song, Ding Liang, Tong Lu, Ping Luo, and Ling Shao. 2022b · 2022
Earlier work this paper cites.
Wave-ViT: Unifying Wavelet and Transformers for Visual Representation Learning. In ECCV , Vol. 13685. 328–345
Ting Yao, Yingwei Pan, Yehao Li, Chong-Wah Ngo, and Tao Mei. 2022 · 2022
Earlier work this paper cites.
Vision Transformer Adapter for Dense Predictions. In ICLR
Zhe Chen, Yuchen Duan, Wenhai Wang, Junjun He, Tong Lu, Jifeng Dai, and Yu Qiao. 2023 · 2023
Earlier work this paper cites.
Hungry Hungry Hippos: Towards Language Modeling with State Space Models. In ICLR
Daniel Y. Fu, Tri Dao, Khaled Kamal Saab, Armin W. Thomas, Atri Rudra, and Christopher Ré. 2023 · 2023
Earlier work this paper cites.
How to Train your HIPPO: State Space Models with Generalized Orthogonal Basis Projections. In ICLR
Albert Gu, Isys Johnson, Aman Timalsina, Atri Rudra, and Christopher Ré. 2023 · 2023
Earlier work this paper cites.
Liquid Structural State-Space Models. In ICLR
Ramin M. Hasani, Mathias Lechner, Tsun-Hsuan Wang, Makram Chahine, Alexander Amini, and Daniela Rus. 2023 · 2023
Earlier work this paper cites.
Segment anything. In ICCV . 4015–4026
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C Berg, Wan-Yen Lo, et al · 2023
Earlier work this paper cites.
Resurrecting Recurrent Neural Networks for Long Sequences. In ICML , Vol. 202. 26670–26698
Antonio Orvieto, Samuel L. Smith, Albert Gu, Anushan Fernando, Çaglar Gülçehre, Razvan Pascanu, and Soham De. 2023 · 2023
Earlier work this paper cites.
Scattering Vision Transformer: Spectral Mixing Matters. In NeurIPS
Badri N. Patro and Vijay Srinivas Agneeswaran. 2023 · 2023
Earlier work this paper cites.
Scalable Diffusion Models with Transformers. In ICCV . 4172–4182
William Peebles and Saining Xie. 2023 · 2023
Earlier work this paper cites.
SG-Former: Self-guided Transformer with Evolving Token Reallocation. In ICCV . 5980–5991
Sucheng Ren, Xingyi Yang, Songhua Liu, and Xinchao Wang. 2023 · 2023
Earlier work this paper cites.
Large Language Models Can Be Easily Distracted by Irrelevant Context. In ICML , Vol. 202. 31210–31227
Freda Shi, Xinyun Chen, Kanishka Misra, Nathan Scales, David Dohan, Ed H. Chi, Nathanael Schärli, and Denny Zhou. 2023 · 2023
Earlier work this paper cites.
Simplified State Space Layers for Sequence Modeling. In ICLR
Jimmy T. H. Smith, Andrew Warrington, and Scott W. Linderman. 2023 · 2023
Earlier work this paper cites.
VOLO: Vision Outlooker for Visual Recognition
Li Yuan, Qibin Hou, Zihang Jiang, Jiashi Feng, and Shuicheng Yan. 2023 · 2023
Earlier work this paper cites.
The Hidden Attention of Mamba Models
Ameen Ali, Itamar Zimerman, and Lior Wolf. 2024 · 2024
Earlier work this paper cites.
ViM-UNet: Vision Mamba for Biomedical Segmentation
Anwai Archit and Constantin Pape. 2024 · 2024
Earlier work this paper cites.
I2I-Mamba: Multi-modal medical image synthesis via selective state space modeling
Omer F. Atli, Bilal Kabas, Fuat Arslan, Mahmut Yurt, Onat Dalmaz, and Tolga Çukur. 2024 · 2024
Earlier work this paper cites.
The Pitfalls of Next-Token Prediction. In ICML , Vol. 235. 2296–2318
Gregor Bachmann and Vaishnavh Nagarajan. 2024 · 2024
Earlier work this paper cites.
Retinexmamba: Retinex-based Mamba for Low-light Image Enhancement
Jiesong Bai, Yuhao Yin, and Qiyuan He. 2024 · 2024
Earlier work this paper cites.
A Novel State Space Model with Local Enhancement and State Sharing for Image Fusion
Zihan Cao, Xiao Wu, Liang-Jian Deng, and Yu Zhong. 2024 · 2024
Earlier work this paper cites.
EM-Net: Efficient Channel and Frequency Learning with Mamba for 3D Medical Image Segmentation. In MICCAI
Ao Chang, Jiajun Zeng, Ruobing Huang, and Dong Ni. 2024 · 2024
Earlier work this paper cites.
Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding
Guo Chen, Yifei Huang, Jilan Xu, Baoqi Pei, Zhe Chen, Zhiqi Li, Jiahao Wang, Kunchang Li, Tong Lu, and Limin Wang. 2024c · 2024
Earlier work this paper cites.
DeMamba: AI-Generated Video Detection on Million-Scale GenVideo Benchmark
Haoxing Chen, Yan Hong, Zizheng Huang, Zhuoer Xu, Zhangxuan Gu, et al · 2024
Earlier work this paper cites.
ChangeMamba: Remote Sensing Change Detection with Spatio-Temporal State Space Model
Hongruixuan Chen, Jian Song, Chengxi Han, Junshi Xia, and Naoto Yokoya. 2024e · 2024
Earlier work this paper cites.
Rsmamba: Remote sensing image classification with state space model
Keyan Chen, Bowen Chen, Chenyang Liu, Wenyuan Li, Zhengxia Zou, and Zhenwei Shi. 2024a · 2024
Earlier work this paper cites.
MiM-ISTD: Mamba-in-Mamba for Efficient Infrared Small Target Detection
Tianxiang Chen, Zhentao Tan, Tao Gong, Qi Chu, Yue Wu, Bin Liu, Jieping Ye, and Nenghai Yu. 2024f · 2024
Earlier work this paper cites.
TokenUnify: Scalable Autoregressive Visual Pre-training with Mixture Token Prediction
Yinda Chen, Haoyuan Shi, Xiaoyu Liu, Te Shi, Ruobing Zhang, Dong Liu, Zhiwei Xiong, and Feng Wu. 2024d · 2024
Earlier work this paper cites.
SurvMamba: State Space Model with Multi-grained Multi-modal Interaction for Survival Prediction
Ying Chen, Jiajing Xie, Yuxiang Lin, Yuhang Song, Wenxian Yang, and Rongshan Yu. 2024g · 2024
Earlier work this paper cites.
MambaUIE: Unraveling the Ocean’s Secrets with Only 2.8 FLOPs
Zhihao Chen and Yiyuan Ge. 2024 · 2024
Earlier work this paper cites.
Activating Wider Areas in Image Super-Resolution
Cheng Cheng, Hang Wang, and Hongbin Sun. 2024 · 2024
Earlier work this paper cites.
Theoretical Foundations of Deep Selective State-Space Models
Nicola Muca Cirone, Antonio Orvieto, Benjamin Walker, Cristopher Salvi, and Terry J. Lyons. 2024 · 2024
Earlier work this paper cites.
Vision Transformers Need Registers. In ICLR
Timothée Darcet, Maxime Oquab, Julien Mairal, and Piotr Bojanowski. 2024 · 2024
Earlier work this paper cites.
CU-Mamba: Selective State Space Models with Channel Learning for Image Restoration
Rui Deng and Tianpei Gu. 2024 · 2024
Earlier work this paper cites.
Fusion-Mamba for Cross-modality Object Detection
Wenhao Dong, Haodong Zhu, Shaohui Lin, Xiaoyan Luo, Yunhang Shen, Xuhui Liu, Juan Zhang, Guodong Guo, and Baochang Zhang. 2024d · 2024
Earlier work this paper cites.
Understanding Robustness of Visual State Space Models for Image Classification
Chengbin Du, Yanxi Li, and Chang Xu. 2024a · 2024
Earlier work this paper cites.
A mixed Mamba U-net for prostate segmentation in MR images
Qiu Du, Luowu Wang, and Hao Chen. 2024b · 2024
Earlier work this paper cites.
PathMamba: Weakly Supervised State Space Model for Multi-class Segmentation of Pathology Images . In MICCAI
Jiansong Fan, Tianxu Lv, Yicheng Di, Lihua Li, and Xiang Pan. 2024 · 2024
Earlier work this paper cites.
MamMIL: Multiple Instance Learning for Whole Slide Images with State Space Models. In BIBM
Zijie Fang, Yifeng Wang, Zhi Wang, Jian Zhang, Xiangyang Ji, and Yongbing Zhang. 2024 · 2024
Earlier work this paper cites.
Scalable Diffusion Models with State Space Backbone
Zhengcong Fei, Mingyuan Fan, Changqian Yu, and Junshi Huang. 2024 · 2024
Earlier work this paper cites.
SSUMamba: Spatial-Spectral Selective State Space Model for Hyperspectral Image Denoising
Guanyiman Fu, Fengchao Xiong, Jianfeng Lu, Jun Zhou, and Yuntao Qian. 2024c · 2024
Earlier work this paper cites.
MD-Dose: A Diffusion Model based on the Mamba for Radiotherapy Dose Prediction
Linjie Fu, Xia Li, Xiuding Cai, Yingkai Wang, Xueyao Wang, Yali Shen, and Yu Yao. 2024a · 2024
Earlier work this paper cites.
Learning Enriched Features via Selective State Spaces Model for Efficient Image Deblurring
Hu Gao and Depeng Dang. 2024 · 2024
Earlier work this paper cites.
Matten: Video Generation with Mamba-Attention
Yu Gao, Jiancheng Huang, Xiaopeng Sun, Zequn Jie, Yujie Zhong, and Lin Ma. 2024 · 2024
Earlier work this paper cites.
MambaTSR: You only need 90k parameters for traffic sign recognition
Yiyuan Ge, Zhihao Chen, Mingxin Yu, Qing Yue, Rui You, and Lianqing Zhu. 2024 · 2024
Earlier work this paper cites.
Haifan Gong, Luoyao Kang, Yitao Wang, Xiang Wan, and Haofeng Li. 2024 · 2024
Cited alongside, same era.
Is Mamba Capable of In-Context Learning?. In AutoML
Riccardo Grazzi, Julien Siems, Simon Schrodi, Thomas Brox, and Frank Hutter. 2024 · 2024
Cited alongside, same era.
Mamba: Linear-Time Sequence Modeling with Selective State Spaces. In COLM
Albert Gu and Tri Dao. 2024 · 2024
Cited alongside, same era.
WaterMamba: Visual State Space Model for Underwater Image Enhancement
Meisheng Guan, Haiyong Xu, Gangyi Jiang, Mei Yu, Yeyao Chen, Ting Luo, and Yang Song. 2024 · 2024
Cited alongside, same era.
MambaMorph: a Mamba-based Framework for Medical MR-CT Deformable Registration
Tao Guo, Yinuo Wang, Shihao Shu, Diansheng Chen, Zhouping Tang, Cai Meng, and Xiangzhi Bai. 2024c · 2024
Kazi Shahriar Sanjid, Md Tanzim Hossain, Md Shakib Shahariar Junayed, and Dr Mohammad Monir Uddin. 2024a · 2024
Closest in time.
Kazi Shahriar Sanjid, Md. Tanzim Hossain, Md. Shakib Shahariar Junayed, and M. Monir Uddin. 2024b · 2024
Closest in time.
GroupMamba: Parameter-Efficient and Accurate Group Visual State Space Model
Abdelrahman Shaker, Syed Talal Wasim, Salman Khan, Juergen Gall, and Fahad Shahbaz Khan. 2024 · 2024
Closest in time.
Locating and Editing Factual Associations in Mamba. In COLM
Arnab Sen Sharma, David Atkinson, and David Bau. 2024 · 2024
Closest in time.
Gamba: Marry Gaussian Splatting with Mamba for single view 3D reconstruction
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Mamba3D: Enhancing Local Features for 3D Point Cloud Analysis via State Space Model. In ACM MM
Xu Han, Yuan Tang, Zhaoxuan Wang, and Xianzhi Li. 2024 · 2024
Cited alongside, same era.
T-Mamba: Frequency-Enhanced Gated Long-Range Dependency for Tooth 3D CBCT Segmentation
Jing Hao, Lei He, and Kuo Feng Hung. 2024 · 2024
Cited alongside, same era.
MambaVision: A Hybrid Mamba-Transformer Vision Backbone
Ali Hatamizadeh and Jan Kautz. 2024 · 2024
Cited alongside, same era.
Pan-Mamba: Effective pan-sharpening with State Space Model
Xuanhua He, Ke Cao, Keyu Yan, Rui Li, Chengjun Xie, Jie Zhang, and Man Zhou. 2024b · 2024
Cited alongside, same era.
3DSS-Mamba: 3D-Spectral-Spatial Mamba for Hyperspectral Image Classification
Yan He, Bing Tu, Bo Liu, Jun Li, and Antonio Plaza. 2024c · 2024
Cited alongside, same era.
Computation-Efficient Era: A Comprehensive Survey of State Space Models in Medical Image Analysis
Moein Heidari, Sina Ghorbani Kolahi, Sanaz Karimijafarbigloo, Bobby Azad, Afshin Bozorgpour, et al · 2024
Cited alongside, same era.
Zigma: Zigzag mamba diffusion model. In ECCV
Vincent Tao Hu, Stefan Andreas Baumann, Ming Gui, Olga Grebenkova, Pingchuan Ma, Johannes Fischer, and Bjorn Ommer. 2024 · 2024
Cited alongside, same era.
Qiuhong Shen, Xuanyu Yi, Zike Wu, Pan Zhou, Hanwang Zhang, Shuicheng Yan, and Xinchao Wang. 2024 · 2024
Closest in time.
VSSD: Vision Mamba with Non-Causal State Space Duality
Yuheng Shi, Minjing Dong, Mingjia Li, and Chang Xu. 2024b · 2024
Closest in time.
Vmambair: Visual state space model for image restoration
Yuan Shi, Bin Xia, Xiaoyu Jin, Xing Wang, Tianyu Zhao, Xin Xia, Xuefeng Xiao, and Wenming Yang. 2024d · 2024
Closest in time.
Yuntao Shou, Tao Meng, Fuchen Zhang, Nan Yin, and Keqin Li. 2024 · 2024
Closest in time.
Distillation-free Scaling of Large SSMs for Images and Videos
Hamid Suleman, Syed Talal Wasim, Muzammal Naseer, and Juergen Gall. 2024 · 2024
Closest in time.
R2Gen-Mamba: A Selective State Space Model for Radiology Report Generation. In ISBI
Yongheng Sun, Yueh Z Lee, Genevieve A Woodard, Hongtu Zhu, Chunfeng Lian, and Mingxia Liu. 2024 · 2024
Closest in time.
Wavelet-based Mamba with Fourier Adjustment for Low-light Image Enhancement. In ACCV
Junhao Tan, Songwen Pei, Wei Qin, Bo Fu, Ximing Li, and Libo Huang. 2024 · 2024
Closest in time.
Rotate to Scan: UNet-like Mamba with Triplet SSM Module for Medical Image Segmentation
Hao Tang, Lianglun Cheng, Guoheng Huang, Zhengguang Tan, Junhao Lu, and Kaihong Wu. 2024a · 2024
Closest in time.
Scalable Visual State Space Model with Fractal Scanning
Lv Tang, HaoKe Xiao, Peng-Tao Jiang, Hao Zhang, Jinwei Chen, and Bo Li. 2024c · 2024
Closest in time.
DiM: Diffusion Mamba for Efficient High-Resolution Image Synthesis
Yao Teng, Yue Wu, Han Shi, Xuefei Ning, Guohao Dai, Yu Wang, Zhenguo Li, and Xihui Liu. 2024 · 2024
Closest in time.
UU-Mamba: Uncertainty-aware U-Mamba for Cardiac Image Segmentation
Ting Yu Tsai, Li Lin, Shu Hu, Hongtu Zhu, Xin Wang, et al · 2024
Closest in time.
Tushar Verma, Jyotsna Singh, Yash Bhartari, Rishi Jarwal, Suraj Singh, and Shubhkarman Singh. 2024 · 2024
Closest in time.
V2M: Visual 2-Dimensional Mamba for Image Representation Learning
Chengkun Wang, Wenzhao Zheng, Yuanhui Huang, Jie Zhou, and Jiwen Lu. 2024m · 2024
Closest in time.
GlobalMamba: Global Image Serialization for Vision Mamba
Chengkun Wang, Wenzhao Zheng, Jie Zhou, and Jiwen Lu. 2024p · 2024
Closest in time.
Mamba-R: Vision Mamba ALSO Needs Registers
Feng Wang, Jiahao Wang, Sucheng Ren, Guoyizhe Wei, Jieru Mei, Wei Shao, Yuyin Zhou, Alan Yuille, and Cihang Xie. 2024j · 2024
Closest in time.
S2Mamba: A Spatial-spectral State Space Model for Hyperspectral Image Classification
Guanchun Wang, Xiangrong Zhang, Zelin Peng, Tianyang Zhang, Xiuping Jia, and Licheng Jiao. 2024l · 2024
Closest in time.
SAM-Med3D: Towards General-purpose Segmentation Models for Volumetric Medical Images
Haoyu Wang, Sizheng Guo, Jin Ye, Zhongying Deng, Junlong Cheng, Tianbin Li, Jianpin Chen, Yanzhou Su, Ziyan Huang, Yiqing Shen, Bin Fu, Shaoting Zhang, Junjun He, and Yu Qiao. 2024d · 2024
Closest in time.
Large Window-based Mamba UNet for Medical Image Segmentation: Beyond Convolution and Self-attention
Jinhong Wang, Jintai Chen, Danny Chen, and Jian Wu. 2024a · 2024
Closest in time.
Insectmamba: Insect pest classification with state space model
Qianning Wang, Chenglin Wang, Zhixin Lai, and Yucheng Zhou. 2024i · 2024
Closest in time.
StableSSM: Alleviating the Curse of Memory in State-space Models through Stable Reparameterization
Shida Wang and Qianxiao Li. 2024 · 2024
Closest in time.
Mask-Guided Mamba Fusion for Drone-Based Visible-Infrared Vehicle Detection
Simiao Wang, Chunpeng Wang, Chaoyi Shi, Yunan Liu, and Mingyu Lu. 2024k · 2024
Closest in time.
GMSR:Gradient-Guided Mamba for Spectral Reconstruction from RGB Images
Xinying Wang, Zhixiong Huang, Sifan Zhang, Jiawen Zhu, and Lin Feng. 2024e · 2024
Closest in time.
Text-controlled Motion Mamba: Text-Instructed Temporal Grounding of Human Motion
Xinghan Wang, Zixi Kang, and Yadong Mu. 2024f · 2024
Closest in time.
State Space Model for New-Generation Network Alternative to Transformers: A Survey
Xiao Wang, Shiao Wang, Yuhe Ding, Yuehang Li, Wentao Wu, Yao Rong, Weizhe Kong, Ju Huang, Shihao Li, Haoxiang Yang, Ziwen Wang, Bo Jiang, Chenglong Li, Yaowei Wang, Yonghong Tian, and Jin Tang. 2024h · 2024
Closest in time.
PoinTramba: A Hybrid Transformer-Mamba Framework for Point Cloud Analysis
Zicheng Wang, Zhenghao Chen, Yiming Wu, Zhen Zhao, Luping Zhou, and Dong Xu. 2024c · 2024
Closest in time.
Ziyang Wang and Chao Ma. 2024 · 2024
Closest in time.
Ziyang Wang, Jian-Qing Zheng, Chao Ma, and Tao Guo. 2024n · 2024
Closest in time.
Mamba-unet: Unet-like pure visual mamba for medical image segmentation
Ziyang Wang, Jian-Qing Zheng, Yichi Zhang, Ge Cui, and Lei Li. 2024o · 2024
Closest in time.
OneBEV: Using One Panoramic Image for Bird’s-Eye-View Semantic Mapping. In ACCV
Jiale Wei, Junwei Zheng, Ruiping Liu, Jie Hu, Jiaming Zhang, and Rainer Stiefelhagen. 2024 · 2024
Closest in time.
MambaLLIE: Implicit Retinex-Aware Low Light Enhancement with Global-then-Local State Space. In NeurIPS
Jiangwei Weng, Zhiqiang Yan, Ying Tai, Jianjun Qian, Jian Yang, and Jun Li. 2024 · 2024
Closest in time.
H-vmunet: High-order vision mamba unet for medical image segmentation
Renkai Wu, Yinghao Liu, Pengchen Liang, and Qing Chang. 2024a · 2024
Closest in time.
Renkai Wu, Yinghao Liu, Pengchen Liang, and Qing Chang. 2024b · 2024
Closest in time.
OverlapMamba: Novel Shift State Space Model for LiDAR-based Place Recognition
Qiuchi Xiang, Jintao Cheng, Jiehao Luo, Jin Wu, Rui Fan, Xieyuanli Chen, and Xiaoyu Tang. 2024 · 2024
Closest in time.
Spatial-Mamba: Effective Visual State Space Models via Structure-Aware State Fusion
Chaodong Xiao, Minghan Li, Zhengqiang Zhang, Deyu Meng, and Lei Zhang. 2024b · 2024
Closest in time.
Frequency-Assisted Mamba for Remote Sensing Image Super-Resolution
Yi Xiao, Qiangqiang Yuan, Kui Jiang, Yuzeng Chen, Qiang Zhang, and Chia-Wen Lin. 2024c · 2024
Closest in time.
FusionMamba: Dynamic Feature Enhancement for Multimodal Image Fusion with Mamba
Xinyu Xie, Yawen Cui, Chio-In Ieong, Tao Tan, Xiaozhi Zhang, Xubin Zheng, and Zitong Yu. 2024a · 2024
Closest in time.
Segmamba: Long-range sequential modeling mamba for 3d medical image segmentation. In MICCAI
Zhaohu Xing, Tian Ye, Yijun Yang, Guang Liu, and Lei Zhu. 2024 · 2024
Closest in time.
HC-Mamba: Vision MAMBA with Hybrid Convolutional Techniques for Medical Image Segmentation
Jiashu Xu. 2024 · 2024
Closest in time.
Image Deraining with Frequency-Enhanced State Space Model
Shugo Yamashita and Masaaki Ikehara. 2024 · 2024
Closest in time.
Guangqian Yang, Kangrui Du, Zhihan Yang, Ye Du, Yongping Zheng, and Shujun Wang. 2024b · 2024
Closest in time.
Vivim: a Video Vision Mamba for Medical Video Object Segmentation
Yijun Yang, Zhaohu Xing, Chunwang Huang, and Lei Zhu. 2024e · 2024
Closest in time.
Spectralmamba: Efficient mamba for hyperspectral image classification
Jing Yao, Danfeng Hong, Chenyu Li, and Jocelyn Chanussot. 2024 · 2024
Closest in time.
Zi Ye and Tianxiang Chen. 2024 · 2024
Closest in time.
MambaOut: Do We Really Need Mamba for Vision?
Weihao Yu and Xinchao Wang. 2024 · 2024
Closest in time.
MUCM-Net: A Mamba Powered UCM-Net for Skin Lesion Segmentation
Chunyu Yuan, Dongfang Zhao, and Sos S Agaian. 2024 · 2024
Closest in time.
Medmamba: Vision mamba for medical image classification
Yubiao Yue and Zhenzhang Li. 2024 · 2024
Closest in time.
MambaMOS: LiDAR-based 3D Moving Object Segmentation with Motion-aware State Space Model. In ACM MM
Kang Zeng, Hao Shi, Jiacheng Lin, Siyu Li, Jintao Cheng, Kaiwei Wang, Zhiyong Li, and Kailun Yang. 2024 · 2024
Closest in time.
LCM: Locally Constrained Compact Point Cloud Model for Masked Point Modeling. In NeurIPS
Yaohua Zha, Naiqi Li, Yanzi Wang, Tao Dai, Hang Guo, Bin Chen, Zhi Wang, Zhihao Ouyang, and Shu-Tao Xia. 2024 · 2024
Closest in time.
Exploring Token Pruning in Vision State Space Models. In NeurIPS
Zheng Zhan, Zhenglun Kong, Yifan Gong, Yushu Wu, Zichong Meng, Hangyu Zheng, Xuan Shen, Stratis Ioannidis, Wei Niu, Pu Zhao, et al · 2024
Closest in time.
HRVMamba: High-Resolution Visual State Space Model for Dense Prediction
Hao Zhang, Yongqiang Ma, Wenqi Shao, Ping Luo, Nanning Zheng, and Kaipeng Zhang. 2024g · 2024
Closest in time.
Hanwei Zhang, Ying Zhu, Dan Wang, Lijun Zhang, Tianxiang Chen, and Zi Ye. 2024i · 2024
Closest in time.
Vim-F: Visual State Space Model Benefiting from Learning in the Frequency Domain
Juntao Zhang, Kun Bian, Peng Cheng, Wenbo An, Jianning Liu, and Jun Zhou. 2024a · 2024
Closest in time.
Point Could Mamba: Point Cloud Learning via State Space Model
Tao Zhang, Xiangtai Li, Haobo Yuan, Shunping Ji, and Shuicheng Yan. 2024d · 2024
Closest in time.
Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient Inference
Han Zhao, Min Zhang, Wei Zhao, Pengxiang Ding, Siteng Huang, and Donglin Wang. 2024b · 2024
Closest in time.
RS-Mamba for Large Remote Sensing Image Dense Prediction
Sijie Zhao, Hao Chen, Xueliang Zhang, Pengfeng Xiao, Lei Bai, and Wanli Ouyang. 2024a · 2024
Closest in time.
FreqMamba: Viewing Mamba from a Frequency Perspective for Image Deraining. In ACM MM
Zou Zhen, Yu Hu, and Zhao Feng. 2024 · 2024
Closest in time.
U-shaped Vision Mamba for Single Image Dehazing
Zhuoran Zheng and Chen Wu. 2024 · 2024
Closest in time.
FD-Vision Mamba for Endoscopic Exposure Correction
Zhuoran Zheng and Jun Zhang. 2024 · 2024
Closest in time.
MambaFormerSR: A Lightweight model for Remote-Sensing Image Super-Resolution
Ruicong Zhi, Xiaopei Fan, and Jingye Shi. 2024 · 2024
Closest in time.
RSDehamba: Lightweight Vision Mamba for Remote Sensing Satellite Image Dehazing
Huiling Zhou, Xianhao Wu, Hongming Chen, Xiang Chen, and Xin He. 2024b · 2024
Closest in time.
3DMambaIPF: A State Space Model for Iterative Point Cloud Filtering via Differentiable Rendering
Qingyuan Zhou, Weidong Yang, Ben Fei, Jingyi Xu, Rui Zhang, Keyi Liu, Yeqi Luo, and Ying He. 2024c · 2024
Closest in time.
Weilian Zhou, Sei-Ichiro Kamata, Haipeng Wang, Man-Sing Wong, et al · 2024
Closest in time.
Samba: Semantic Segmentation of Remotely Sensed Images with State Space Model
Qinfeng Zhu, Yuanzhi Cai, Yuan Fang, Yihan Yang, Cheng Chen, Lei Fan, and Anh Nguyen. 2024a · 2024
Closest in time.
Qinfeng Zhu, Yuan Fang, Yuanzhi Cai, Cheng Chen, and Lei Fan. 2024b · 2024
Closest in time.
RhythmMamba: Fast Remote Physiological Measurement with Arbitrary Length Videos
Bochao Zou, Zizheng Guo, Xiaocheng Hu, and Huimin Ma. 2024b · 2024
Closest in time.
Venturing into Uncharted Waters: The Navigation Compass from Transformer to Mamba
Yuchen Zou, Yineng Chen, Zuchao Li, Lefei Zhang, and Hai Zhao. 2024a · 2024
Closest in time.
State Space Models for Event Cameras. In CVPR
Nikola Zubic, Mathias Gehrig, and Davide Scaramuzza. 2024 · 2024
Closest in time.
SUM: Saliency Unification through Mamba for Visual Attention Modeling. In WACV
Alireza Hosseini, Amirhossein Kazerouni, Saeed Akhavan, Michael Brudno, and Babak Taati. 2025 · 2025
Closest in time.
Convolution and Attention-Free Mamba-based Cardiac Image Segmentation. In WACV
Abbas Khan, Muhammad Asad, Martin Benning, Caroline Roney, and Gregory Slabaugh. 2025 · 2025
Closest in time.
Sigma: Siamese Mamba Network for Multi-Modal Semantic Segmentation. In WACV
Zifu Wan, Yuhao Wang, Silong Yong, Pingping Zhang, Simon Stepputtis, Katia Sycara, and Yaqi Xie. 2025 · 2025
Closest in time.