Fetching the paper…
Reading the bibliography…
There has been a proliferation of artificial intelligence applications, where model training is key to promising high-quality services for these applications.
Automatically constructing a corpus of sentential paraphrases
Bill Dolan and Chris Brockett · 2005
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky and Geoffrey Hinton · 2009
Earlier work this paper cites.
Socially assistive robots in elderly care: a systematic review into effects and effectiveness
Roger Bemelmans, Gert Jan Gelderblom, Pieter Jonker, and Luc De Witte · 2012
Earlier work this paper cites.
How transferable are features in deep neural networks?
Jason Yosinski, Jeff Clune, Yoshua Bengio, and Hod Lipson · 2014
Earlier work this paper cites.
Visualizing and understanding convolutional networks
Matthew D Zeiler and Rob Fergus · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Sheng-syun Shen and Hung-yi Lee · 2016
Earlier work this paper cites.
Residual networks behave like ensembles of relatively shallow networks
Andreas Veit, Michael J Wilber, and Serge Belongie · 2016
Earlier work this paper cites.
Freezeout: Accelerate training by progressively freezing layers
Andrew Brock, Theodore Lim, James M Ritchie, and Nick Weston · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Tvm: An automated end-to-end optimizing compiler for deep learning
Tianqi Chen, Thierry Moreau, Ziheng Jiang, Lianmin Zheng, Eddie Yan, Haichen Shen, Meghan Cowan, Leyuan Wang, Yuwei Hu, Luis Ceze, Carlos Guestrin, and Arvind Krishnamurthy · 2018
Earlier work this paper cites.
Rish: A robot-integrated smart home for elderly care
Ha Manh Do, Minh Pham, Weihua Sheng, Dan Yang, and Meiqin Liu · 2018
Earlier work this paper cites.
Attention-based sequence classification for affect detection
Cristina Gorrostieta, Richard Brutti, Kye Taylor, Avi Shapiro, Joseph Moran, Ali Azarbayejani, and John Kane · 2018
Cited alongside, same era.
Attention clusters: Purely attention based local feature integration for video classification
Xiang Long, Chuang Gan, Gerard De Melo, Jiajun Wu, Xiao Liu, and Shilei Wen · 2018
Cited alongside, same era.
Mobilenetv2: Inverted residuals and linear bottlenecks
Mark Sandler, Andrew Howard, Menglong Zhu, Andrey Zhmoginov, and Liang-Chieh Chen · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin Ming-Wei Chang Kenton and Lee Kristina Toutanova · 2019
Cited alongside, same era.
Similarity of neural network representations revisited
Simon Kornblith, Mohammad Norouzi, Honglak Lee, and Geoffrey Hinton · 2019
Cited alongside, same era.
What would elsa do? freezing layers during transformer fine-tuning
Training high-performance and large-scale deep neural networks with full 8-bit integers
Yukuan Yang, Lei Deng, Shuang Wu, Tianyi Yan, Yuan Xie, and Guoqi Li · 2020
Later among the works it cites.
Accelerating training of transformer-based language models with progressive layer dropping
Minjia Zhang and Yuxiong He · 2020
Later among the works it cites.
Applications, databases and open computer vision research from drone videos and images: a survey
Younes Akbari, Noor Almaadeed, Somaya Al-Maadeed, and Omar Elharrouss · 2021
Later among the works it cites.
Cross-domain similarity learning for face recognition in unseen domains
Masoud Faraki, Xiang Yu, Yi-Hsuan Tsai, Yumin Suh, and Manmohan Chandraker · 2021
Later among the works it cites.
Pipetransformer: Automated elastic pipelining for distributed training of transformers
Chaoyang He, Shen Li, Mahdi Soltanolkotabi, and Salman Avestimehr · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jaejun Lee, Raphael Tang, and Jimmy Lin · 2019
Cited alongside, same era.
GLUE: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel R. Bowman · 2019
Cited alongside, same era.
Neural network acceptability judgments
Alex Warstadt, Amanpreet Singh, and Samuel R Bowman · 2019
Cited alongside, same era.
The fusion of internet of intelligent things (ioit) in remote diagnosis of obstructive sleep apnea: A survey and a new model
Mohamed Abdel-Basset, Weiping Ding, and Laila Abdel-Fatah · 2020
Cited alongside, same era.
Rigging the lottery: Making all tickets winners
Utku Evci, Trevor Gale, Jacob Menick, Pablo Samuel Castro, and Erich Elsen · 2020
Cited alongside, same era.
A smart, efficient, and reliable parking surveillance system with edge artificial intelligence on iot devices
Ruimin Ke, Yifan Zhuang, Ziyuan Pu, and Yinhai Wang · 2020
Cited alongside, same era.
Patdnn: Achieving real-time dnn execution on mobile devices with pattern-based weight pruning
Wei Niu, Xiaolong Ma, Sheng Lin, Shihao Wang, Xuehai Qian, Xue Lin, Yanzhi Wang, and Bin Ren · 2020
Cited alongside, same era.
Autofreeze: Automatically freezing model blocks to accelerate fine-tuning
Yuhan Liu, Saurabh Agarwal, and Shivaram Venkataraman · 2021
Later among the works it cites.
Artificial intelligence-based remote diagnosis of sleep apnea using instantaneous heart rates
Prabodh Panindre, Vijay Gandhi, and Sunil Kumar · 2021
Later among the works it cites.
Training data-efficient image transformers & distillation through attention
Hugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa, Alexandre Sablayrolles, and Herve Jegou · 2021
Later among the works it cites.
Enabling on-device self-supervised contrastive learning with selective data contrast
Yawen Wu, Zhepeng Wang, Dewen Zeng, Yiyu Shi, and Jingtong Hu · 2021
Later among the works it cites.
Mest: Accurate and fast memory-economic sparse training framework on the edge
Geng Yuan, Xiaolong Ma, Wei Niu, Zhengang Li, Zhenglun Kong, Ning Liu, Yifan Gong, Zheng Zhan, Chaoyang He, Qing Jin, et al · 2021
Later among the works it cites.
Distribution adaptive int8 quantization for training cnns
Kang Zhao, Sida Huang, Pan Pan, Yinghan Li, Yingya Zhang, Zhenyu Gu, and Yinghui Xu · 2021
Later among the works it cites.
Towards making the most of context in neural machine translation
Zaixiang Zheng, Xiang Yue, Shujian Huang, Jiajun Chen, and Alexandra Birch · 2021
Later among the works it cites.
Contextual transformer networks for visual recognition
Yehao Li, Ting Yao, Yingwei Pan, and Tao Mei · 2022
Later among the works it cites.
Layer freezing & data sieving: Missing pieces of a generic framework for sparse training
Geng Yuan, Yanyu Li, Sheng Li, Zhenglun Kong, Sergey Tulyakov, Xulong Tang, Yanzhi Wang, and Jian Ren · 2022
Later among the works it cites.