Fetching the paper…
Reading the bibliography…
End-to-end neural network models achieve improved performance on various automatic speech recognition (ASR) tasks.
“Adam: A method for stochastic optimization,”
Diederik P Kingma and Jimmy Ba, · 2014
Earlier work this paper cites.
“Librispeech: an asr corpus based on public domain audio books,”
Vassil Panayotov, Guoguo Chen, Daniel Povey, and Sanjeev Khudanpur, · 2015
Earlier work this paper cites.
“Quantization and training of neural networks for efficient integer-arithmetic-only inference,”
Benoit Jacob, Skirmantas Kligys, Bo Chen, Menglong Zhu, Matthew Tang, Andrew Howard, Hartwig Adam, and Dmitry Kalenichenko, · 2018
Earlier work this paper cites.
“GAP-8: A RISC-V SoC for AI at the edge of the IoT,”
Eric Flamand, Davide Rossi, Francesco Conti, Igor Loi, Antonio Pullini, Florent Rotenberg, and Luca Benini, · 2018
Earlier work this paper cites.
“Jasper: An end-to-end convolutional neural acoustic model,”
Jason Li, Vitaly Lavrukhin, Boris Ginsburg, Ryan Leary, Oleksii Kuchaiev, Jonathan M Cohen, Huyen Nguyen, and Ravi Teja Gadde, · 2019
Earlier work this paper cites.
“Q8BERT: Quantized 8bit BERT,”
Ofir Zafrir, Guy Boudoukh, Peter Izsak, and Moshe Wasserblat, · 2019
Earlier work this paper cites.
“A simplified fully quantized transformer for end-to-end speech recognition,”
Alex Bie, Bharat Venkitesh, Joao Monteiro, Md Haidar, and Mehdi Rezagholizadeh, · 2019
Earlier work this paper cites.
“Data-free learning of student networks,”
Hanting Chen, Yunhe Wang, Chang Xu, Zhaohui Yang, Chuanjian Liu, Boxin Shi, Chunjing Xu, Chao Xu, and Qi Tian, · 2019
Earlier work this paper cites.
“SpecAugment: A simple data augmentation method for automatic speech recognition,”
Daniel S Park, William Chan, Yu Zhang, Chung-Cheng Chiu, Barret Zoph, Ekin D Cubuk, and Quoc V Le, · 2019
Earlier work this paper cites.
“Conformer: Convolution-augmented transformer for speech recognition,”
Anmol Gulati, James Qin, Chung-Cheng Chiu, Niki Parmar, Yu Zhang, Jiahui Yu, Wei Han, Shibo Wang, Zhengdong Zhang, Yonghui Wu, and Ruoming Pang, · 2020
Cited alongside, same era.
Wei Han, Zhengdong Zhang, Yu Zhang, Jiahui Yu, Chung-Cheng Chiu, James Qin, Anmol Gulati, Ruoming Pang, and Yonghui Wu, · 2020
Cited alongside, same era.
“Transformer Transducer: A streamable speech recognition model with transformer encoders and RNN-T loss,”
Qian Zhang, Han Lu, Hasim Sak, Anshuman Tripathi, Erik McDermott, Stephen Koo, and Shankar Kumar, · 2020
Cited alongside, same era.
“QuartzNet: Deep automatic speech recognition with 1d time-channel separable convolutions,”
Samuel Kriman, Stanislav Beliaev, Boris Ginsburg, Jocelyn Huang, Oleksii Kuchaiev, Vitaly Lavrukhin, Ryan Leary, Jason Li, and Yang Zhang, · 2020
Cited alongside, same era.
“The knowledge within: Methods for data-free model compression,”
Matan Haroush, Itay Hubara, Elad Hoffer, and Daniel Soudry, · 2020
Later among the works it cites.
“ZeroQ: A novel zero shot quantization framework,”
Yaohui Cai, Zhewei Yao, Zhen Dong, Amir Gholami, Michael W Mahoney, and Kurt Keutzer, · 2020
Later among the works it cites.
“Data-free network quantization with adversarial knowledge distillation,”
Yoojin Choi, Jihwan Choi, Mostafa El-Khamy, and Jungwon Lee, · 2020
Later among the works it cites.
“Generative low-bitwidth data free quantization,”
Shoukai Xu, Haokun Li, Bohan Zhuang, Jing Liu, Jiezhang Cao, Chuangrun Liang, and Mingkui Tan, · 2020
Later among the works it cites.
“Bayesian bits: Unifying quantization and pruning,”
Mart van Baalen, Christos Louizos, Markus Nagel, Rana Ali Amjad, Ying Wang, Tijmen Blankevoort, and Max Welling, · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhewei Yao, Zhen Dong, Zhangcheng Zheng, Amir Gholami, Jiali Yu, Eric Tan, Leyuan Wang, Qijing Huang, Yida Wang, Michael W Mahoney, and Kurt Keutzer, · 2020
Cited alongside, same era.
“Q-BERT: Hessian based ultra low precision quantization of BERT,”
Sheng Shen, Zhen Dong, Jiayu Ye, Linjian Ma, Zhewei Yao, Amir Gholami, Michael W Mahoney, and Kurt Keutzer, · 2020
Cited alongside, same era.
“Int8 winograd acceleration for conv1d equipped asr models deployed on mobile devices,”
Yiwu Yao, Yuchao Li, Chengyu Wang, Tianhang Yu, Houjiang Chen, Xiaotang Jiang, Jun Yang, Jun Huang, Wei Lin, Hui Shu, and Chengfei Lv, · 2020
Cited alongside, same era.
“Quantization aware training with absolute-cosine regularization for automatic speech recognition,”
Hieu Duy Nguyen, Anastasios Alexandridis, and Athanasios Mouchtaris, · 2020
Cited alongside, same era.
“Quantization of acoustic model parameters in automatic speech recognition framework,”
Amrutha Prasad, Petr Motlicek, and Srikanth Madikeri, · 2020
Cited alongside, same era.
“Cortex-M, https://developer.arm.com/ip-products/processors/cortex-m,”
ARM,
Cited in the paper.
“Google Edge TPU, https://cloud.google.com/edge-tpu,”
Cited in the paper.
“NeMo, https://github.com/nvidia/nemo,”
Cited in the paper.
Sehoon Kim, Amir Gholami, Zhewei Yao, Michael W Mahoney, and Kurt Keutzer, · 2021
Closest in time.
“Generative zero-shot network quantization,”
Xiangyu He, Qinghao Hu, Peisong Wang, and Jian Cheng, · 2021
Closest in time.
“A survey of quantization methods for efficient neural network inference,”
Amir Gholami, Sehoon Kim, Zhen Dong, Zhewei Yao, Michael W Mahoney, and Kurt Keutzer, · 2021
Closest in time.