Fetching the paper…
Reading the bibliography…
Spatial redundancy widely exists in visual recognition tasks, i.e., discriminative features in an image or video frame usually correspond to only a subset of pixels, while the remaining regions are irrelevant to the task at hand.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in ICML , 2016, pp. 1928–1937
1937
Earlier work this paper cites.
Y. LeCun, J. S. Denker, and S. A. Solla, “Optimal brain damage,” in NeurIPS , 1990, pp. 598–605
1990
Earlier work this paper cites.
R. J. Williams, “Simple statistical gradient-following algorithms for connectionist reinforcement learning,” Machine learning , vol. 8, no. 3-4, pp. 229–256, 1992
1992
Earlier work this paper cites.
D. Whitley, “A genetic algorithm tutorial,” Statistics and computing , vol. 4, no. 2, pp. 65–85, 1994
1994
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in ICML , 2009, pp. 248–255
2009
Earlier work this paper cites.
H. Larochelle and G. E. Hinton, “Learning to combine foveal glimpses with a third-order boltzmann machine,” in NeurIPS , 2010, pp. 1243–1251
2010
Earlier work this paper cites.
F. Larsson and M. Felsberg, “Using fourier descriptors and spatial models for traffic sign recognition,” in Scandinavian conference on image analysis . Springer, 2011, pp. 238–249
2011
Earlier work this paper cites.
A. Grubb and D. Bagnell, “Speedboost: Anytime prediction with uniform near-optimality,” in AISTATS , 2012, pp. 458–466
2012
Earlier work this paper cites.
J. Wan, D. Wang, S. C. H. Hoi, P. Wu, J. Zhu, Y. Zhang, and J. Li, “Deep learning for content-based image retrieval: A comprehensive study,” in ACM MM , 2014, pp. 157–166
2014
Earlier work this paper cites.
V. Mnih, N. Heess, A. Graves et al. , “Recurrent models of visual attention,” in NeurIPS , 2014, pp. 2204–2212
2014
Earlier work this paper cites.
G. Hinton, O. Vinyals, and J. Dean, “Distilling the knowledge in a neural network,” in NeurIPS Deep Learning Workshop , 2014
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Very deep convolutional networks for large-scale image recognition,” in ICLR , 2015
2015
Earlier work this paper cites.
O. Vinyals, A. Toshev, S. Bengio, and D. Erhan, “Show and tell: A neural image caption generator,” in CVPR , 2015, pp. 3156–3164
2015
Earlier work this paper cites.
R. Vedantam, C. Lawrence Zitnick, and D. Parikh, “Cider: Consensus-based image description evaluation,” in CVPR , 2015, pp. 4566–4575
2015
Earlier work this paper cites.
A. Karpathy and L. Fei-Fei, “Deep visual-semantic alignments for generating image descriptions,” in CVPR , 2015, pp. 3128–3137
2015
Earlier work this paper cites.
J. Ba, V. Mnih, and K. Kavukcuoglu, “Multiple object recognition with visual attention,” in ICLR , 2015
2015
Earlier work this paper cites.
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich, “Going deeper with convolutions,” in CVPR , 2015, pp. 1–9
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in CVPR , 2016, pp. 770–778
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
M. Rastegari, V. Ordonez, J. Redmon, and A. Farhadi, “Xnor-net: Imagenet classification using binary convolutional neural networks,” in ECCV . Springer, 2016, pp. 525–542
2016
Earlier work this paper cites.
I. Hubara, M. Courbariaux, D. Soudry, R. El-Yaniv, and Y. Bengio, “Binarized neural networks,” in NeurIPS , 2016, pp. 4107–4115
2016
Earlier work this paper cites.
Z. Yang, X. He, J. Gao, L. Deng, and A. Smola, “Stacked attention networks for image question answering,” in CVPR , 2016, pp. 21–29
2016
Earlier work this paper cites.
B. Zhou, A. Khosla, A. Lapedriza, A. Oliva, and A. Torralba, “Learning deep features for discriminative localization,” in CVPR , 2016, pp. 2921–2929
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
S. Xie, R. Girshick, P. Dollár, Z. Tu, and K. He, “Aggregated residual transformations for deep neural networks,” in CVPR , 2017, pp. 1492–1500
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
J. Fu, H. Zheng, and T. Mei, “Look closer to see better: Recurrent attention convolutional neural network for fine-grained image recognition,” in CVPR , 2017, pp. 4438–4446
2017
Earlier work this paper cites.
R. Goyal, S. Ebrahimi Kahou, V. Michalski, J. Materzynska, S. Westphal, H. Kim, V. Haenel, I. Fruend, P. Yianilos, M. Mueller-Freitag et al. , “The "something something" video database for learning and evaluating visual common sense,” in ICCV , 2017, pp. 5842–5850
2017
Earlier work this paper cites.
H. Li, A. Kadav, I. Durdanovic, H. Samet, and H. P. Graf, “Pruning filters for efficient convnets,” in ICLR , 2017
2017
Earlier work this paper cites.
Z. Liu, J. Li, Z. Shen, G. Huang, S. Yan, and C. Zhang, “Learning efficient convolutional networks through network slimming,” in CVPR , 2017, pp. 2736–2744
2017
Earlier work this paper cites.
T. Bolukbasi, J. Wang, O. Dekel, and V. Saligrama, “Adaptive neural networks for efficient inference,” in ICML . JMLR. org, 2017, pp. 527–536
2017
Earlier work this paper cites.
N. Shazeer, A. Mirhoseini, K. Maziarz, A. Davis, Q. Le, G. Hinton, and J. Dean, “Outrageously large neural networks: The sparsely-gated mixture-of-experts layer,” in ICLR , 2017
2017
Cited alongside, same era.
J. Lin, Y. Rao, J. Lu, and J. Zhou, “Runtime neural pruning,” in NeurIPS , 2017, pp. 2181–2191
2017
Cited alongside, same era.
M. Figurnov, M. D. Collins, Y. Zhu, L. Zhang, J. Huang, D. Vetrov, and R. Salakhutdinov, “Spatially adaptive computation time for residual networks,” in CVPR , 2017, pp. 1039–1048
2017
Cited alongside, same era.
H. Zheng, J. Fu, T. Mei, and J. Luo, “Learning multi-attention convolutional neural network for fine-grained image recognition,” in ICCV , 2017, pp. 5209–5217
2017
Cited alongside, same era.
R. R. Selvaraju, M. Cogswell, A. Das, R. Vedantam, D. Parikh, and D. Batra, “Grad-cam: Visual explanations from deep networks via gradient-based localization,” in CVPR , 2017, pp. 618–626
A. Ruiz and J. Verbeek, “Adaptative inference cost with convolutional neural mixture models,” in ICCV , 2019, pp. 1872–1881
2019
Later among the works it cites.
Y. Chen, H. Fan, B. Xu, Z. Yan, Y. Kalantidis, M. Rohrbach, S. Yan, and J. Feng, “Drop an octave: Reducing spatial redundancy in convolutional neural networks with octave convolution,” in CVPR , 2019, pp. 3435–3444
2019
Later among the works it cites.
W. Hua, Y. Zhou, C. M. De Sa, Z. Zhang, and G. E. Suh, “Channel gating neural networks,” in NeurIPS , 2019, pp. 1884–1894
2019
Later among the works it cites.
M. Najibi, B. Singh, and L. S. Davis, “Autofocus: Efficient multi-scale inference,” in ICCV , 2019, pp. 9745–9755
2019
Later among the works it cites.
Y. Zhu, W. Sun, X. Cao, C. Wang, D. Wu, Y. Yang, and N. Ye, “Ta-cnn: Two-way attention models in deep convolutional neural network for plant recognition,” Neurocomputing , vol. 365, pp. 191–200, 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
2017
Cited alongside, same era.
J. Hu, L. Shen, and G. Sun, “Squeeze-and-excitation networks,” in CVPR , 2018, pp. 7132–7141
2018
Cited alongside, same era.
G. Huang, S. Liu, L. Van der Maaten, and K. Q. Weinberger, “Condensenet: An efficient densenet using learned group convolutions,” in CVPR , 2018, pp. 2752–2761
2018
Cited alongside, same era.
M. Sandler, A. Howard, M. Zhu, A. Zhmoginov, and L.-C. Chen, “Mobilenetv2: Inverted residuals and linear bottlenecks,” in CVPR , 2018, pp. 4510–4520
2018
Cited alongside, same era.
X. Zhang, X. Zhou, M. Lin, and J. Sun, “Shufflenet: An extremely efficient convolutional neural network for mobile devices,” in CVPR , 2018, pp. 6848–6856
2018
Cited alongside, same era.
N. Ma, X. Zhang, H.-T. Zheng, and J. Sun, “Shufflenet v2: Practical guidelines for efficient cnn architecture design,” in ECCV , 2018, pp. 116–131
2018
Cited alongside, same era.
G. Huang, D. Chen, T. Li, F. Wu, L. van der Maaten, and K. Q. Weinberger, “Multi-scale dense networks for resource efficient image classification,” in ICLR , 2018
2018
Cited alongside, same era.
2019
Later among the works it cites.
C. Chen, O. Li, D. Tao, A. Barnett, C. Rudin, and J. K. Su, “This looks like that: deep learning for interpretable image recognition,” in NeurIPS , 2019
2019
Later among the works it cites.
H. Zheng, J. Fu, Z.-J. Zha, and J. Luo, “Looking for the devil in the details: Learning trilinear attention sampling network for fine-grained image recognition,” in CVPR , 2019, pp. 5012–5021
2019
Later among the works it cites.
Y. Ding, Y. Zhou, Y. Zhu, Q. Ye, and J. Jiao, “Selective sparse sampling for fine-grained image recognition,” in ICCV , 2019, pp. 6599–6608
2019
Later among the works it cites.
L. Zhang, S. Huang, W. Liu, and D. Tao, “Learning a mixture of granularity-specific experts for fine-grained categorization,” in ICCV , 2019, pp. 8331–8340
2019
Later among the works it cites.
B. He, J. Li, Y. Zhao, and Y. Tian, “Part-regularized near-duplicate vehicle re-identification,” in CVPR , 2019, pp. 3997–4005
2019
Later among the works it cites.
Y. Zhu, J. Xie, Z. Tang, X. Peng, and A. Elgammal, “Semantic-guided multi-attention localization for zero-shot learning,” NeurIPS , vol. 32, 2019
2019
Later among the works it cites.
B. Chen and W. Deng, “Hybrid-attention based decoupled metric learning for zero-shot image retrieval,” in CVPR , 2019, pp. 2750–2759
2019
Later among the works it cites.
Y. Wang, X. Pan, S. Song, H. Zhang, G. Huang, and C. Wu, “Implicit semantic data augmentation for deep networks,” in NeurIPS , 2019, pp. 12 635–12 644
2019
Later among the works it cites.
A. Katharopoulos and F. Fleuret, “Processing megapixel images with deep attention-sampling models,” in ICML . PMLR, 2019, pp. 3282–3291
2019
Later among the works it cites.
M. Tan, B. Chen, R. Pang, V. Vasudevan, M. Sandler, A. Howard, and Q. V. Le, “Mnasnet: Platform-aware neural architecture search for mobile,” in CVPR , 2019, pp. 2820–2828
2019
Later among the works it cites.
B. Wu, X. Dai, P. Zhang, Y. Wang, F. Sun, Y. Wu, Y. Tian, P. Vajda, Y. Jia, and K. Keutzer, “Fbnet: Hardware-aware efficient convnet design via differentiable neural architecture search,” in CVPR , 2019, pp. 10 734–10 742
2019
Later among the works it cites.
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga et al. , “Pytorch: An imperative style, high-performance deep learning library,” in NeurIPS , 2019, pp. 8024–8035
2019
Later among the works it cites.
L. Yang, Y. Han, X. Chen, S. Song, J. Dai, and G. Huang, “Resolution adaptive networks for efficient inference,” in CVPR , 2020
2020
Later among the works it cites.
I. Radosavovic, R. P. Kosaraju, R. Girshick, K. He, and P. Dollár, “Designing network design spaces,” in CVPR , 2020
2020
Later among the works it cites.
Y. Wang, K. Lv, R. Huang, S. Song, L. Yang, and G. Huang, “Glance and focus: a dynamic approach to reducing spatial redundancy in image classification,” in NeurIPS , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
Q. Wang, W. Huang, Z. Xiong, and X. Li, “Looking closer at the scene: Multiscale representation learning for remote sensing image scene classification,” IEEE Transactions on Neural Networks and Learning Systems , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
L. Yang, H. Jiang, R. Cai, Y. Wang, S. Song, G. Huang, and Q. Tian, “Condensenet v2: Sparse feature reactivation for deep networks,” in CVPR , 2021
2021
Later among the works it cites.
Y. Han, G. Huang, S. Song, L. Yang, H. Wang, and Y. Wang, “Dynamic neural networks: A survey,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2021
2021
Later among the works it cites.
Y. Wang, R. Huang, S. Song, Z. Huang, and G. Huang, “Not all images are worth 16x16 words: Dynamic transformers for efficient image recognition,” in NeurIPS , 2021
2021
Later among the works it cites.
Y. Wang, Z. Chen, H. Jiang, S. Song, Y. Han, and G. Huang, “Adaptive focus for efficient video recognition,” in ICCV , October 2021
2021
Later among the works it cites.
Y. Han, G. Huang, S. Song, L. Yang, Y. Zhang, and H. Jiang, “Spatially adaptive feature refinement for efficient inference,” IEEE Transactions on Image Processing , vol. 30, pp. 9345–9358, 2021
2021
Later among the works it cites.
Y. Meng, R. Panda, C.-C. Lin, P. Sattigeri, L. Karlinsky, K. Saenko, A. Oliva, and R. Feris, “Adafuse: Adaptive temporal fusion network for efficient action recognition,” in ICLR , 2021. [Online]. Available: https://openreview.net/forum?id=bM3L3I_853
2021
Later among the works it cites.
2022
Closest in time.
Y. Wang, Y. Yue, Y. Lin, H. Jiang, Z. Lai, V. Kulikov, N. Orlov, H. Shi, and G. Huang, “Adafocus v2: End-to-end training of spatial dynamic networks for video recognition,” in CVPR , 2022
2022
Closest in time.
M. Jaderberg, K. Simonyan, A. Zisserman et al. , “Spatial transformer networks,” in NeurIPS , 2015, pp. 2017–2025
2025
Closest in time.
K. Xu, J. Ba, R. Kiros, K. Cho, A. Courville, R. Salakhudinov, R. Zemel, and Y. Bengio, “Show, attend and tell: Neural image caption generation with visual attention,” in ICML , 2015, pp. 2048–2057
2057
Closest in time.