Fetching the paper…
Reading the bibliography…
State space models (SSMs) with selection mechanisms and hardware-aware architectures, namely Mamba, have recently demonstrated significant promise in long-sequence modeling.
The perceptron, a perceiving and recognizing automaton Project Para
Frank Rosenblatt · 1957
Earlier work this paper cites.
Principles of neurodynamics: Perceptrons and the theory of brain mechanisms
Frank Rosenblatt et al · 1962
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, and Patrick Haffner · 1998
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton · 2012
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2014
Earlier work this paper cites.
Convolutional lstm network: A machine learning approach for precipitation nowcasting
Xingjian Shi, Zhourong Chen, Hao Wang, Dit-Yan Yeung, Wai-Kin Wong, and Wang-chun Woo · 2015
Earlier work this paper cites.
A decomposable attention model for natural language inference
Ankur P Parikh, Oscar Täckström, Dipanjan Das, and Jakob Uszkoreit · 2016
Earlier work this paper cites.
Gaussian error linear units (gelus)
Dan Hendrycks and Kevin Gimpel · 2016
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Swish: a self-gated activation function
Prajit Ramachandran, Barret Zoph, and Quoc V. Le · 2017
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Earlier work this paper cites.
Can recurrent neural networks warp time?
Corentin Tallec and Yann Ollivier · 2018
Earlier work this paper cites.
Digital radiography
Euclid Seeram and Euclid Seeram · 2019
Earlier work this paper cites.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Earlier work this paper cites.
Transformers are rnns: Fast autoregressive transformers with linear attention
Angelos Katharopoulos, Apoorv Vyas, Nikolaos Pappas, and François Fleuret · 2020
Earlier work this paper cites.
Overview of guidance for endoscopy during the coronavirus disease 2019 pandemic
Rashid N Lui, Sunny H Wong, Sergio A Sánchez-Luna, Gianluca Pellino, Steven Bollipo, Mei-Yin Wong, Philip WY Chiu, and Joseph JY Sung · 2020
Earlier work this paper cites.
Super-resolution ultrasound imaging
Kirsten Christensen-Jeffries, Olivier Couture, Paul A Dayton, Yonina C Eldar, Kullervo Hynynen, Fabian Kiessling, Meaghan O’Reilly, Gianmarco F Pinton, Georg Schmitz, Meng-Xing Tang, et al · 2020
Earlier work this paper cites.
Brain tumor segmentation and classification from magnetic resonance images: Review of selected methods from 2014 to 2019
Arti Tiwari, Shilpa Srivastava, and Millie Pant · 2020
Earlier work this paper cites.
Ckconv: Continuous kernel convolution for sequential data
David W Romero, Anna Kuzina, Erik J Bekkers, Jakub M Tomczak, and Mark Hoogendoorn · 2021
Earlier work this paper cites.
Shuangfei Zhai, Walter Talbott, Nitish Srivastava, Chen Huang, Hanlin Goh, Ruixiang Zhang, and Josh Susskind · 2021
Earlier work this paper cites.
X-ray computed tomography
Philip J Withers, Charles Bouman, Simone Carmignato, Veerle Cnudde, David Grimaldi, Charlotte K Hagen, Eric Maire, Marena Manley, Anton Du Plessis, and Stuart R Stock · 2021
Earlier work this paper cites.
Hungry hungry hippos: Towards language modeling with state space models
Daniel Y Fu, Tri Dao, Khaled K Saab, Armin W Thomas, Atri Rudra, and Christopher Ré · 2022
Earlier work this paper cites.
Long movie clip classification with state-space video models
Md Mohaiminul Islam and Gedas Bertasius · 2022
Earlier work this paper cites.
Mamba: Linear-time sequence modeling with selective state spaces
Albert Gu and Tri Dao · 2023
Earlier work this paper cites.
Retentive network: A successor to transformer for large language models
Yutao Sun, Li Dong, Shaohan Huang, Shuming Ma, Yuqing Xia, Jilong Xue, Jianyong Wang, and Furu Wei · 2023
Earlier work this paper cites.
Hyena hierarchy: Towards larger convolutional language models
Michael Poli, Stefano Massaroli, Eric Nguyen, Daniel Y Fu, Tri Dao, Stephen Baccus, Yoshua Bengio, Stefano Ermon, and Christopher Ré · 2023
Earlier work this paper cites.
Retentive network: A successor to transformer for large language models (2023)
Yutao Sun, Li Dong, Shaohan Huang, Shuming Ma, Yuqing Xia, Jilong Xue, Jianyong Wang, and Furu Wei · 2023
Earlier work this paper cites.
Rwkv: Reinventing rnns for the transformer era
Bo Peng, Eric Alcaide, Quentin Anthony, Alon Albalak, Samuel Arcadinho, Huanqi Cao, Xin Cheng, Michael Chung, Matteo Grella, Kranthi Kiran GV, et al · 2023
Earlier work this paper cites.
Jamba: A hybrid transformer-mamba language model
Opher Lieber, Barak Lenz, Hofit Bata, Gal Cohen, Jhonathan Osin, Itay Dalmedigos, Erez Safahi, Shaked Meirom, Yonatan Belinkov, Shai Shalev-Shwartz, et al · 2024
Earlier work this paper cites.
Moe-mamba: Efficient selective state space models with mixture of experts
Maciej Pióro, Kamil Ciebiera, Krystian Król, Jan Ludziejewski, and Sebastian Jaszczur · 2024
Earlier work this paper cites.
Blackmamba: Mixture of experts for state-space models
Quentin Anthony, Yury Tokpanov, Paolo Glorioso, and Beren Millidge · 2024
Earlier work this paper cites.
Vision mamba: Efficient visual representation learning with bidirectional state space model
Lianghui Zhu, Bencheng Liao, Qian Zhang, Xinlong Wang, Wenyu Liu, and Xinggang Wang · 2024
Earlier work this paper cites.
Vmamba: Visual state space model
Yue Liu, Yunjie Tian, Yuzhong Zhao, Hongtian Yu, Lingxi Xie, Yaowei Wang, Qixiang Ye, and Yunfan Liu · 2024
Earlier work this paper cites.
Localmamba: Visual state space model with windowed selective scan
Tao Huang, Xiaohuan Pei, Shan You, Fei Wang, Chen Qian, and Chang Xu · 2024
Earlier work this paper cites.
Plainmamba: Improving non-hierarchical mamba in visual recognition
Chenhongyi Yang, Zehui Chen, Miguel Espinosa, Linus Ericsson, Zhenyu Wang, Jiaming Liu, and Elliot J Crowley · 2024
Cited alongside, same era.
Efficientvmamba: Atrous selective scan for light weight visual mamba
Xiaohuan Pei, Tao Huang, and Chang Xu · 2024
Cited alongside, same era.
Mambamixer: Efficient selective state space models with dual token and channel selection
Ali Behrouz, Michele Santacatterina, and Ramin Zabih · 2024
Cited alongside, same era.
Mamba-nd: Selective state space modeling for multi-dimensional data
Shufan Li, Harkanwar Singh, and Aditya Grover · 2024
Cited alongside, same era.
Simba: Simplified mamba-based architecture for vision and multivariate time series
Vm-unet-v2 rethinking vision mamba unet for medical image segmentation
Mingya Zhang, Yue Yu, Limei Gu, Tingsheng Lin, and Xianping Tao · 2024
Closest in time.
Medmamba: Vision mamba for medical image classification
Yubiao Yue and Zhenzhang Li · 2024
Closest in time.
Mim-istd: Mamba-in-mamba for efficient infrared small target detection
Tianxiang Chen, Zhentao Tan, Tao Gong, Qi Chu, Yue Wu, Bin Liu, Jieping Ye, and Nenghai Yu · 2024
Closest in time.
Rs3mamba: Visual state space model for remote sensing images semantic segmentation
Xianping Ma, Xiaokang Zhang, and Man-On Pun · 2024
Closest in time.
Freqmamba: Viewing mamba from a frequency perspective for image deraining
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Badri N Patro and Vijay S Agneeswaran · 2024
Cited alongside, same era.
Zigma: Zigzag mamba diffusion model
Vincent Tao Hu, Stefan Andreas Baumann, Ming Gui, Olga Grebenkova, Pingchuan Ma, Johannes Fischer, and Bjorn Ommer · 2024
Cited alongside, same era.
Vmambair: Visual state space model for image restoration
Yuan Shi, Bin Xia, Xiaoyu Jin, Xing Wang, Tianyu Zhao, Xin Xia, Xuefeng Xiao, and Wenming Yang · 2024
Cited alongside, same era.
Videomamba: State space model for efficient video understanding
Kunchang Li, Xinhao Li, Yi Wang, Yinan He, Yali Wang, Limin Wang, and Yu Qiao · 2024
Cited alongside, same era.
Zeyu Zhang, Akide Liu, Ian Reid, Richard Hartley, Bohan Zhuang, and Hao Tang · 2024
Cited alongside, same era.
Vivim: a video vision mamba for medical video object segmentation
Yijun Yang, Zhaohu Xing, and Lei Zhu · 2024
Cited alongside, same era.
Rsmamba: Remote sensing image classification with state space model
Keyan Chen, Bowen Chen, Chenyang Liu, Wenyuan Li, Zhengxia Zou, and Zhenwei Shi · 2024
Cited alongside, same era.
Harmamba: Efficient wearable sensor human activity recognition based on bidirectional selective ssm
Shuangjian Li, Tao Zhu, Furong Duan, Liming Chen, Huansheng Ning, and Yaping Wan · 2024
Cited alongside, same era.
Zou Zhen, Yu Hu, and Zhao Feng · 2024
Closest in time.
Rs-mamba for large remote sensing image dense prediction
Sijie Zhao, Hao Chen, Xueliang Zhang, Pengfeng Xiao, Lei Bai, and Wanli Ouyang · 2024
Closest in time.
State space models for event cameras
Nikola Zubić, Mathias Gehrig, and Davide Scaramuzza · 2024
Closest in time.
Spikemba: Multi-modal spiking saliency mamba for temporal video grounding
Wenrui Li, Xiaopeng Hong, and Xiaopeng Fan · 2024
Closest in time.
Cobra: Extending mamba to multi-modal large language model for efficient inference
Han Zhao, Min Zhang, Wei Zhao, Pengxiang Ding, Siteng Huang, and Donglin Wang · 2024
Closest in time.
U-shaped vision mamba for single image dehazing
Zhuoran Zheng and Chen Wu · 2024
Closest in time.
Hu Gao and Depeng Dang · 2024
Closest in time.
Mambatalk: Efficient holistic gesture synthesis with selective state space models
Zunnan Xu, Yukang Lin, Haonan Han, Sicheng Yang, Ronghui Li, Yachao Zhang, and Xiu Li · 2024
Closest in time.
Scalable diffusion models with state space backbone
Zhengcong Fei, Mingyuan Fan, Changqian Yu, and Junshi Huang · 2024
Closest in time.
3dmambacomplete: Exploring structured state space model for point cloud completion
Yixuan Li, Weidong Yang, and Ben Fei · 2024
Closest in time.
3dmambaipf: A state space model for iterative point cloud filtering via differentiable rendering
Qingyuan Zhou, Weidong Yang, Ben Fei, Jingyi Xu, Rui Zhang, Keyi Liu, Yeqi Luo, and Ying He · 2024
Closest in time.
Point could mamba: Point cloud learning via state space model
Tao Zhang, Xiangtai Li, Haobo Yuan, Shunping Ji, and Shuicheng Yan · 2024
Closest in time.
Pointmamba: A simple state space model for point cloud analysis
Dingkang Liang, Xin Zhou, Xinyu Wang, Xingkui Zhu, Wei Xu, Zhikang Zou, Xiaoqing Ye, and Xiang Bai · 2024
Closest in time.
Gamba: Marry gaussian splatting with mamba for single view 3d reconstruction
Qiuhong Shen, Xuanyu Yi, Zike Wu, Pan Zhou, Hanwang Zhang, Shuicheng Yan, and Xinchao Wang · 2024
Closest in time.
Ssm meets video diffusion models: Efficient video generation with structured state spaces
Yuta Oshima, Shohei Taniguchi, Masahiro Suzuki, and Yutaka Matsuo · 2024
Closest in time.
U-mamba: Enhancing long-range dependency for biomedical image segmentation
Jun Ma, Feifei Li, and Bo Wang · 2024
Closest in time.
Zi Ye and Tianxiang Chen · 2024
Closest in time.
Promamba: Prompt-mamba for polyp segmentation
Jianhao Xie, Ruofan Liao, Ziang Zhang, Sida Yi, Yuesheng Zhu, and Guibo Luo · 2024
Closest in time.
Ziyang Wang and Chao Ma · 2024
Closest in time.
Md-dose: A diffusion model based on the mamba for radiotherapy dose prediction
Linjie Fu, Xia Li, Xiuding Cai, Yingkai Wang, Xueyao Wang, Yali Shen, and Yu Yao · 2024
Closest in time.
Mambamil: Enhancing long sequence modeling with sequence reordering in computational pathology
Shu Yang, Yihui Wang, and Hao Chen · 2024
Closest in time.
Fd-vision mamba for endoscopic exposure correction
Zhuoran Zheng and Jun Zhang · 2024
Closest in time.
Lightm-unet: Mamba assists in lightweight unet for medical image segmentation
Weibin Liao, Yinghao Zhu, Xinyuan Wang, Cehngwei Pan, Yasha Wang, and Liantao Ma · 2024
Closest in time.
Segmamba: Long-range sequential modeling mamba for 3d medical image segmentation
Zhaohu Xing, Tian Ye, Yijun Yang, Guang Liu, and Lei Zhu · 2024
Closest in time.
T-mamba: Frequency-enhanced gated long-range dependency for tooth 3d cbct segmentation
Jing Hao, Lei He, and Kuo Feng Hung · 2024
Closest in time.
Guangqian Yang, Kangrui Du, Zhihan Yang, Ye Du, Yongping Zheng, and Shujun Wang · 2024
Closest in time.
Haifan Gong, Luoyao Kang, Yitao Wang, Xiang Wan, and Haofeng Li · 2024
Closest in time.
Tao Guo, Yinuo Wang, and Cai Meng · 2024
Closest in time.
Pan-mamba: Effective pan-sharpening with state space model
Xuanhua He, Ke Cao, Keyu Yan, Rui Li, Chengjun Xie, Jie Zhang, and Man Zhou · 2024
Closest in time.
Judy X Yang, Jun Zhou, Jing Wang, Hui Tian, and Alan Wee Chung Liew · 2024
Closest in time.
Samba: Semantic segmentation of remotely sensed images with state space model
Qinfeng Zhu, Yuanzhi Cai, Yuan Fang, Yihan Yang, Cheng Chen, Lei Fan, and Anh Nguyen · 2024
Closest in time.