Fetching the paper…
Reading the bibliography…
Resistive Random Access Memory (ReRAM) has emerged as a promising platform for deep neural networks (DNNs) due to its support for parallel in-situ matrix-vector multiplication.
Topics in matrix analysis, 1994
Horn Roger and R Johnson Charles · 1994
Earlier work this paper cites.
Rram defect modeling and failure analysis based on march test and a novel squeeze-search scheme
Ching-Yi Chen, Hsiu-Chuan Shih, Cheng-Wen Wu, Chih-He Lin, Pi-Feng Chiu, Shyh-Shyuan Sheu, and Frederick T Chen · 2014
Earlier work this paper cites.
Vortex: Variation-aware training for memristor x-bar
Beiye Liu, Hai Li, Yiran Chen, Xin Li, Qing Wu, and Tingwen Huang · 2015
Earlier work this paper cites.
Rram defect modeling and failure analysis based on march test and a novel squeeze-search scheme
Ching-Yi Chen, Hsiu-Chuan Shih, Cheng-Wen Wu, Chih-He Lin, Pi-Feng Chiu, Shyh-Shyuan Sheu, and Frederick T. Chen · 2015
Earlier work this paper cites.
Prime: A novel processing-in-memory architecture for neural network computation in reram-based main memory
Ping Chi, Shuangchen Li, Cong Xu, Tao Zhang, Jishen Zhao, Yongpan Liu, Yu Wang, and Yuan Xie · 2016
Earlier work this paper cites.
Isaac: A convolutional neural network accelerator with in-situ analog arithmetic in crossbars
Ali Shafiee, Anirban Nag, Naveen Muralimanohar, Rajeev Balasubramonian, John Paul Strachan, Miao Hu, R Stanley Williams, and Vivek Srikumar · 2016
Earlier work this paper cites.
Deep compression: Compressing deep neural networks with pruning, trained quantization and huffman coding
Song Han, Huizi Mao, and William J Dally · 2016
Earlier work this paper cites.
Fixed point quantization of deep convolutional networks
Darryl Lin, Sachin Talathi, and Sreekanth Annapureddy · 2016
Earlier work this paper cites.
Binarized neural networks
Itay Hubara, Matthieu Courbariaux, Daniel Soudry, Ran El-Yaniv, and Yoshua Bengio · 2016
Earlier work this paper cites.
Nanoscale hafnium oxide rram devices exhibit pulse dependent behavior and multi-level resistance capability
Beckmann Karsten et al · 2016
Earlier work this paper cites.
Weighted-entropy-based quantization for deep neural networks
Eunhyeok Park, Junwhan Ahn, and Sungjoo Yoo · 2017
Cited alongside, same era.
Training sparse neural networks
Suraj Srinivas, Akshayvarun Subramanya, and R Venkatesh Babu · 2017
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Cited alongside, same era.
A systematic dnn weight pruning framework using alternating direction method of multipliers
Tianyun Zhang, Shaokai Ye, Kaiqi Zhang, Jian Tang, Wujie Wen, Makan Fardad, and Yanzhi Wang · 2018
Cited alongside, same era.
Darts: Differentiable architecture search
Hanxiao Liu, Karen Simonyan, and Yiming Yang · 2018
Cited alongside, same era.
Nest: A neural network synthesis tool based on a grow-and-prune paradigm
Ftrans: energy-efficient acceleration of transformers using fpga
Bingbing Li, Santosh Pandey, Haowen Fang, Yanjun Lyv, Ji Li, Jieyang Chen, Mimi Xie, Lipeng Wan, Hang Liu, and Caiwen Ding · 2020
Later among the works it cites.
Att: A fault-tolerant reram accelerator for attention-based neural networks
Haoqiang Guo, Lu Peng, Jian Zhang, Qing Chen, and Travis D LeCompte · 2020
Later among the works it cites.
Transformers: State-of-the-art natural language processing
Wolf Thomas et al · 2020
Later among the works it cites.
Forms: Fine-grained polarized reram-based in-situ computation for mixed-signal dnn accelerator
Geng Yuan, Payman Behnam, Zhengang Li, Ali Shafiee, Sheng Lin, Xiaolong Ma, Hang Liu, Xuehai Qian, Mahdi Nazm Bojnordi, Yanzhi Wang, and Caiwen Ding · 2021
Later among the works it cites.
A unified dnn weight pruning framework using reweighted optimization methods
Tianyun Zhang, Xiaolong Ma, Zheng Zhan, Shanglin Zhou, Caiwen Ding, Makan Fardad, and Yanzhi Wang · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Xiaoliang Dai, Hongxu Yin, and Niraj Jha · 2019
Cited alongside, same era.
Autoprune: Automatic network pruning by regularizing auxiliary parameters
Xia Xiao, Zigeng Wang, and Sanguthevar Rajasekaran · 2019
Cited alongside, same era.
Handling stuck-at-faults in memristor crossbar arrays using matrix transformations
Baogang Zhang, Necati Uysal, Deliang Fan, and Rickard Ewetz · 2019
Cited alongside, same era.
Analyzing multi-head self-attention: Specialized heads do the heavy lifting, the rest can be pruned
Elena Voita, David Talbot, Fedor Moiseev, Rico Sennrich, and Ivan Titov · 2019
Cited alongside, same era.
Efficient transformer-based large scale language representations using hardware-friendly block structured pruning
Bingbing Li, Zhenglun Kong, Tianyun Zhang, Ji Li, Zhengang Li, Hang Liu, and Caiwen Ding · 2020
Cited alongside, same era.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever
Cited in the paper.
Computing-in-memory neural network accelerators for safety-critical systems: Can small device variations be disastrous?
Zheyu Yan, Xiaobo Sharon Hu, and Yiyu Shi · 2022
Later among the works it cites.
Swim: Selectivewrite-verify for computing-in-memory neural accelerators
Zheyu Yan, Sharon X Hu, and Yiyu Shi · 2022
Later among the works it cites.
Fault-tolerant deep neural networks for processing-in-memory based autonomous edge systems
Siyue Wang, Geng Yuan, Xiaolong Ma, Yanyu Li, Xue Lin, and Bhavya Kailkhura · 2022
Later among the works it cites.
Improving realistic worst-case performance of nvcim dnn accelerators through training with right-censored gaussian noise
Zheyu Yan, Yifan Qin, Wujie Wen, Xiaobo Sharon Hu, and Yiyu Shi · 2023
Later among the works it cites.