Fetching the paper…
Reading the bibliography…
We propose a Mamba accelerator with reconfigurable architecture, MARCA.We propose three novel approaches in this paper.
Position-location solutions by Taylor-series estimation
Wade H Foy. 1976 · 1976
Earlier work this paper cites.
From time series to linear system—Part I. Finite dimensional linear time invariant systems
Jan C Willems. 1986 · 1986
Earlier work this paper cites.
The Einstein summation notation
Alan H Barr. 1991 · 1991
Earlier work this paper cites.
A fast, compact approximation of the exponential function
Nicol N Schraudolph. 1999 · 1999
Earlier work this paper cites.
CACTI 6.0: A tool to model large caches
Naveen Muralimanohar, Rajeev Balasubramonian, and Norman P Jouppi. 2009 · 2009
Earlier work this paper cites.
Hardware implementation of the exponential function using Taylor series. In 2014 NORCHIP . IEEE, 1–4
Peter Nilsson, Ateeq Ur Rahman Shaik, Rakesh Gangarajaiah, and Erik Hertz. 2014 · 2014
Earlier work this paper cites.
Highlights of the high-bandwidth memory (hbm) standard. In Memory forum workshop , Vol. 3
Mike O’Connor. 2014 · 2014
Earlier work this paper cites.
Deep residual learning for image recognition. In Proceedings of the IEEE conference on computer vision and pattern recognition . 770–778
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Earlier work this paper cites.
Pointer sentinel mixture models
Stephen Merity, Caiming Xiong, et al · 2016
Earlier work this paper cites.
The LAMBADA dataset: Word prediction requiring a broad discourse context
Denis Paperno, Germán Kruszewski, Angeliki Lazaridou, Quan Ngoc Pham, Raffaella Bernardi, Sandro Pezzelle, Marco Baroni, Gemma Boleda, and Raquel Fernández. 2016 · 2016
Earlier work this paper cites.
A fixed point exponential function accelerator for a neuromorphic many-core system. In 2017 IEEE International Symposium on Circuits and Systems (ISCAS) . IEEE, 1–4
Johannes Partzsch, Sebastian Höppner, Matthias Eberlein, Rene Schüffny, Christian Mayr, David R Lester, and Steve Furber. 2017 · 2017
Earlier work this paper cites.
Scaling equations for the accurate prediction of CMOS device performance from 180 nm to 7 nm
Aaron Stillmaker and Bevan Baas. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Think you have solved question answering? try arc, the ai2 reasoning challenge
Peter Clark, Isaac Cowhey, et al · 2018
Earlier work this paper cites.
Sigmoid-weighted linear units for neural network function approximation in reinforcement learning
Stefan Elfwing, Eiji Uchibe, and Kenji Doya. 2018 · 2018
Earlier work this paper cites.
Nvidia tensor core programmability, performance & precision. In 2018 IEEE international parallel and distributed processing symposium workshops (IPDPSW) . IEEE, 522–531
Stefano Markidis, Steven Wei Der Chien, Erwin Laure, Ivy Bo Peng, and Jeffrey S Vetter. 2018 · 2018
Earlier work this paper cites.
A coordinated tiling and batching framework for efficient GEMM on GPUs. In Proceedings of the 24th symposium on principles and practice of parallel programming . 229–241
Xiuhong Li, Yun Liang, Shengen Yan, Liancheng Jia, and Yinghan Li. 2019 · 2019
Cited alongside, same era.
Hellaswag: Can a machine really finish your sentence?
Rowan Zellers, Ari Holtzman, et al · 2019
Cited alongside, same era.
Piqa: Reasoning about physical commonsense in natural language. In AAAI
Yonatan Bisk, Rowan Zellers, et al · 2020
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Cited alongside, same era.
TSTC: Two-level Sparsity Tensor Core Enabling both Algorithm Flexibility and Hardware Efficiency. In 2023 IEEE/ACM International Conference on Computer Aided Design (ICCAD) . IEEE, 1–9
Jun Liu, Guohao Dai, Hao Xia, Lidong Guo, Xiangsheng Shi, Jiaming Xu, Huazhong Yang, and Yu Wang. 2023 · 2023
Later among the works it cites.
Ramulator 2.0: A Modern, Modular, and Extensible DRAM Simulator
Haocong Luo, Yahya Can Tu, F Nisa Bostancı, Ataberk Olgun, A Giray Ya, Onur Mutlu, et al · 2023
Later among the works it cites.
Introducing LLaMA: A foundational, 65-billion-parameter large language model
AI Meta. 2023 · 2023
Later among the works it cites.
Scalable diffusion models with transformers. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 4195–4205
William Peebles and Saining Xie. 2023 · 2023
Later among the works it cites.
Flex-sfu: Accelerating dnn activation functions by non-uniform piecewise approximation. In 2023 60th ACM/IEEE Design Automation Conference (DAC) . IEEE, 1–6
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Siyuan Lu, Meiqi Wang, Shuang Liang, Jun Lin, and Zhongfeng Wang. 2020 · 2020
Cited alongside, same era.
On the opportunities and risks of foundation models
Rishi Bommasani, Drew A Hudson, Ehsan Adeli, Russ Altman, Simran Arora, Sydney von Arx, Michael S Bernstein, Jeannette Bohg, Antoine Bosselut, Emma Brunskill, et al · 2021
Cited alongside, same era.
Nvidia a100 tensor core gpu: Performance and innovation
Jack Choquette, Wishwesh Gandhi, Olivier Giroux, Nick Stam, and Ronny Krashinsky. 2021 · 2021
Cited alongside, same era.
Efficiently modeling long sequences with structured state spaces
Albert Gu, Karan Goel, and Christopher Ré. 2021a · 2021
Cited alongside, same era.
Combining recurrent, convolutional, and continuous-time models with linear state space layers
Albert Gu, Isys Johnson, Karan Goel, Khaled Saab, Tri Dao, Atri Rudra, and Christopher Ré. 2021b · 2021
Cited alongside, same era.
Automatic speech recognition: a survey
Mishaim Malik, Muhammad Kamran Malik, Khawar Mehmood, and Imran Makhdoom. 2021 · 2021
Cited alongside, same era.
Winogrande: An adversarial winograd schema challenge at scale
Keisuke Sakaguchi, Ronan Le Bras, et al · 2021
Cited alongside, same era.
On the parameterization and initialization of diagonal state space models
Albert Gu, Karan Goel, Ankit Gupta, and Christopher Ré. 2022 · 2022
Cited alongside, same era.
Enrico Reggiani, Renzo Andri, and Lukas Cavigelli. 2023 · 2023
Later among the works it cites.
Graph Mamba: Towards Learning on Graphs with State Space Models
Ali Behrouz and Farnoosh Hashemi. 2024 · 2024
Closest in time.
Intel Xeon Platinum 8358P Processor
INTEL. 2024 · 2024
Closest in time.
STG-Mamba: Spatial-Temporal Graph Learning via Selective State Space Model
Lincan Li, Hanchen Wang, Wenjie Zhang, and Adelle Coster. 2024 · 2024
Closest in time.
PointMamba: A Simple State Space Model for Point Cloud Analysis
Dingkang Liang, Xin Zhou, Xinyu Wang, Xingkui Zhu, Wei Xu, Zhikang Zou, Xiaoqing Ye, and Xiang Bai. 2024 · 2024
Closest in time.
Jiuming Liu, Ruiji Yu, Yian Wang, Yu Zheng, Tianchen Deng, Weicai Ye, and Hesheng Wang. 2024 · 2024
Closest in time.
NVIDIA A100 Tensor Core GPU Architecture
NVIDIA. 2024 · 2024
Closest in time.
SiMBA: Simplified Mamba-Based Architecture for Vision and Multivariate Time series
Badri N Patro and Vijay S Agneeswaran. 2024 · 2024
Closest in time.
Graph-mamba: Towards long-range graph sequence modeling with selective state spaces
Chloe Wang, Oleksii Tsepa, Jun Ma, and Bo Wang. 2024 · 2024
Closest in time.
Medmamba: Vision mamba for medical image classification
Yubiao Yue and Zhenzhang Li. 2024 · 2024
Closest in time.
Point Could Mamba: Point Cloud Learning via State Space Model
Tao Zhang, Xiangtai Li, Haobo Yuan, Shunping Ji, and Shuicheng Yan. 2024 · 2024
Closest in time.
Vision mamba: Efficient visual representation learning with bidirectional state space model
Lianghui Zhu, Bencheng Liao, Qian Zhang, Xinlong Wang, Wenyu Liu, and Xinggang Wang. 2024 · 2024
Closest in time.