Fetching the paper…
Reading the bibliography…
Transformer has been adopted to image recognition tasks and shown to outperform CNNs and RNNs while it suffers from high training cost and computational complexity.
Backpropagation Applied to Handwritten Zip Code Recognition
LeCun Y et al · 1989
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky · 2009
Earlier work this paper cites.
An Analysis of Single-Layer Networks in Unsupervised Feature Learning
Adam Coates et al · 2011
Earlier work this paper cites.
ImageNet Large Scale Visual Recognition Challenge
Olga Russakovsky et al · 2015
Earlier work this paper cites.
Very Deep Convolutional Networks for Large-Scale Image Recognition
Karen Simonyan et al · 2015
Earlier work this paper cites.
U-Net: Convolutional Networks for Biomedical Image Segmentation
Olaf Ronneberger et al · 2015
Earlier work this paper cites.
Deep Residual Learning for Image Recognition
Kaiming He et al · 2016
Earlier work this paper cites.
Attention is All you Need
Ashish Vaswani et al · 2017
Earlier work this paper cites.
Revisiting Unreasonable Effectiveness of Data in Deep Learning Era
Chen Sun et al · 2017
Earlier work this paper cites.
MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications
Andrew G. Howard et al · 2017
Earlier work this paper cites.
Xception: Deep Learning with Depthwise Separable Convolutions
François Chollet · 2017
Earlier work this paper cites.
Neural Ordinary Differential Equations
Ricky T. Q. Chen et al · 2018
Earlier work this paper cites.
MobileNetV2: Inverted Residuals and Linear Bottlenecks
Mark Sandler et al · 2018
Earlier work this paper cites.
Neural SDE: Stabilizing Neural ODE Networks with Stochastic Noise
Xuanqing Liu et al · 2019
Earlier work this paper cites.
On Robustness of Neural Ordinary Differential Equations
Hanshu Yan et al · 2019
Earlier work this paper cites.
Augmented Neural ODEs
Emilien Dupont et al · 2019
Earlier work this paper cites.
Atention augmented convolutional networks
Irwan Bello et al · 2019
Earlier work this paper cites.
Root Mean Square Layer Normalization
Biao Zhang et al · 2019
Cited alongside, same era.
EfficientNet: Rethinking Model Scaling for Convolutional Neural Networks
Mingxing Tan and Quoc Le · 2019
Cited alongside, same era.
Reformer: The Efficient Transformer
Nikita Kitaev et al · 2020
Cited alongside, same era.
Transformers are RNNs: fast autoregressive transformers with linear attention
Angelos Katharopoulos et al · 2020
Cited alongside, same era.
Linformer: Self-Attention with Linear Complexity
Sinong Wang et al · 2020
Cited alongside, same era.
How to train your neural ODE: the world of Jacobian and Kinetic regularization
Chris Finlay et al · 2020
Cited alongside, same era.
EfficientNetV2: Smaller Models and Faster Training
Mingxing Tan and Quoc Le · 2021
Later among the works it cites.
MobileNetV3 for Image Classification
Siying Qian et al · 2021
Later among the works it cites.
Asher Trockman et al · 2022
Later among the works it cites.
How Do Vision Transformers Work?
Namuk Park et al · 2022
Later among the works it cites.
Automated Deep Transfer Learning-Based Approach for Detection of COVID-19 Infection in Chest X-Rays
N. Narayan Das et al · 2022
Later among the works it cites.
Automatic Diagnosis of Covid-19 Related Pneumonia from CXR and CT-Scan Images
N. Kumar et al · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
FTRANS: Energy-Efficient Acceleration of Transformers Using FPGA
Bingbing Li et al · 2020
Cited alongside, same era.
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Alexey Dosovitskiy et al · 2021
Cited alongside, same era.
Incorporating Convolution Designs Into Visual Transformers
Kun Yuan et al · 2021
Cited alongside, same era.
CvT: Introducing Convolutions to Vision Transformers
Haiping Wu et al · 2021
Cited alongside, same era.
Bottleneck Transformers for Visual Recognition
Aravind Srinivas et al · 2021
Cited alongside, same era.
CoAtNet: Marrying Convolution and Attention for All Data Sizes
Zihang Dai et al · 2021
Cited alongside, same era.
A Hybrid Network of CNN and Transformer for Lightweight Image Super-Resolution
Jinsheng Fang et al · 2022
Later among the works it cites.
FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness
Tri Dao et al · 2022
Later among the works it cites.
VAQF: Fully Automatic Software-Hardware Co-Design Framework for Low-Bit Vision Transformer
Mengshu Sun et al · 2022
Later among the works it cites.
Learnable Lookup Table for Neural Network Quantization
Longguang Wang et al · 2022
Later among the works it cites.
dsODENet: Neural ODE and Depthwise Separable Convolution for Domain Adaptation on FPGAs
Hiroki Kawakami et al · 2022
Later among the works it cites.
A Lightweight Transformer Model using Neural ODE for FPGAs
Ikumi Okubo et al · 2023
Later among the works it cites.
Hybrid CNN-Transformer Feature Fusion for Single Image Deraining
Xiang Chen et al · 2023
Later among the works it cites.
HCformer: Hybrid CNN-Transformer for LDCT Image Denoising
Jinli Yuan et al · 2023
Later among the works it cites.
A Tiny Transformer-Based Anomaly Detection Framework for IoT Solutions
Luca Barbieri et al · 2023
Later among the works it cites.
ME-ViT: A Single-Load Memory-Efficient FPGA Accelerator for Vision Transformers
Kyle Marino et al · 2024
Closest in time.