Fetching the paper…
Reading the bibliography…
Large-scale transformer-based models like the Bidirectional Encoder Representations from Transformers (BERT) are widely used for Natural Language Processing (NLP) applications, wherein these models are initially pre-trained with a large corpus with millions of parameters and then fine-tuned for a downstream NLP task.
Overview of replab 2013: Evaluating online reputation monitoring systems
E. Amigó, J. Carrillo de Albornoz, I. Chugur, A. Corujo, J. Gonzalo, T. Martín, E. Meij, M. De Rijke, and D. Spina · 2013
Earlier work this paper cites.
S. Han, H. Mao, and W. J. Dally · 2015
Earlier work this paper cites.
Tensorflow: a system for large-scale machine learning
M. Abadi, P. Barham, J. Chen, Z. Chen, A. Davis, J. Dean, M. Devin, S. Ghemawat, G. Irving, M. Isard, et al · 2016
Earlier work this paper cites.
Squeezenet: Alexnet-level accuracy with 50x fewer parameters and< 0.5 mb model size
F. N. Iandola, S. Han, M. W. Moskewicz, K. Ashraf, W. J. Dally, and K. Keutzer · 2016
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova · 2018
Earlier work this paper cites.
Amc: Automl for model compression and acceleration on mobile devices
Y. He, J. Lin, Z. Liu, H. Wang, L.-J. Li, and S. Han · 2018
Earlier work this paper cites.
Quantization and training of neural networks for efficient integer-arithmetic-only inference
B. Jacob, S. Kligys, B. Chen, M. Zhu, M. Tang, A. Howard, H. Adam, and D. Kalenichenko · 2018
Earlier work this paper cites.
Parameter-efficient transfer learning for nlp
N. Houlsby, A. Giurgiu, S. Jastrzebski, B. Morrone, Q. De Laroussilhe, A. Gesmundo, M. Attariyan, and S. Gelly · 2019
Earlier work this paper cites.
Tinybert: Distilling bert for natural language understanding
X. Jiao, Y. Yin, L. Shang, X. Jiang, X. Chen, L. Li, F. Wang, and Q. Liu · 2019
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Y. Liu, M. Ott, N. Goyal, J. Du, M. Joshi, D. Chen, O. Levy, M. Lewis, L. Zettlemoyer, and V. Stoyanov · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, I. Sutskever, et al · 2019
Cited alongside, same era.
Haq: Hardware-aware automated quantization with mixed precision
K. Wang, Z. Liu, Y. Lin, J. Lin, and S. Han · 2019
Cited alongside, same era.
Tinyml: Machine learning with tensorflow lite on arduino and ultra-low-power microcontrollers
P. Warden and D. Situnayake · 2019
Cited alongside, same era.
Xlnet: Generalized autoregressive pretraining for language understanding
Z. Yang, Z. Dai, Y. Yang, J. Carbonell, R. R. Salakhutdinov, and Q. V. Le · 2019
Cited alongside, same era.
Real-time execution of large-scale language models on mobile
W. Niu, Z. Kong, G. Yuan, W. Jiang, J. Guan, C. Ding, P. Zhao, S. Liu, B. Ren, and Y. Wang · 2020
Cited alongside, same era.
Tinyml for ubiquitous edge ai
S. Soro · 2021
Later among the works it cites.
Serverless computing: a security perspective
E. Marin, D. Perino, and R. Di Pietro · 2022
Later among the works it cites.
A bert-based deep learning approach for reputation analysis in social media
M. W. U. Rahman, S. Shao, P. Satam, S. Hariri, C. Padilla, Z. Taylor, and C. Nevarez · 2022
Later among the works it cites.
Interpreter interface for running tensorflow lite models., retrieved: January 2023, available at: https://www.tensorflow.org/api_docs/python/tf/lite/interpreter
2023
Closest in time.
Post-training quantization, retrieved: January 2023, available at: https://www.tensorflow.org/lite/performance/
2023
Closest in time.
psutil documentation, retrieved: January 2023, available at: https://psutil.readthedocs.io/en/latest/
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Z. Sun, H. Yu, X. Song, R. Liu, Y. Yang, and D. Zhou · 2020
Cited alongside, same era.
Hat: Hardware-aware transformers for efficient natural language processing
H. Wang, Z. Wu, Z. Liu, H. Cai, L. Zhu, C. Gan, and S. Han · 2020
Cited alongside, same era.
Lite transformer with long-short range attention
Z. Wu, Z. Liu, J. Lin, Y. Lin, and S. Han · 2020
Cited alongside, same era.
Micronet for efficient language modeling
Z. Yan, H. Wang, D. Guo, and S. Han · 2020
Cited alongside, same era.
A compression-compilation framework for on-mobile real-time bert applications
W. Niu, Z. Kong, G. Yuan, W. Jiang, J. Guan, C. Ding, P. Zhao, S. Liu, B. Ren, and Y. Wang · 2021
Cited alongside, same era.
Acs712: Fully integrated, hall-effect-based linear current sensor ic with 2.1 kvrms voltage isolation and a low-resistance current conductor, available at: https://www.allegromicro.com/en/products/sense/current-sensor-ics/zero-to-fifty-amp-integrated-conductor-sensor-ics/acs712
Cited in the paper.
Arduino uno rev3, available at: https://store.arduino.cc/products/arduino-uno-rev3
Cited in the paper.
2023
Closest in time.
Raspberry pi products, retrieved: January 2023, available at: https://www.raspberrypi.com/products/
2023
Closest in time.
Tensorflow-lite, retrieved: January 2023, https://www.tensorflow.org/lite/guide
2023
Closest in time.
Twitter api, retrieved: January 2023, available at: https://developer.twitter.com/en/docs/twitter-api
2023
Closest in time.
Gpt-4 technical report, 2023
OpenAI · 2023
Closest in time.