Fetching the paper…
Reading the bibliography…
This paper presents LiteEval, a simple yet effective coarse-to-fine framework for resource efficient video recognition, suitable for both online and offline scenarios.
Reinforcement learning: An introduction
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
Robust real-time face detection
P. Viola and M. J. Jones · 2004
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
On the partition function and random maximum a-posteriori perturbations
T. Hazan and T. S. Jaakkola · 2012
Earlier work this paper cites.
Predicting domain adaptivity: Redo or recycle?
T. Yao, C.-W. Ngo, and S. Zhu · 2012
Earlier work this paper cites.
Microsoft coco: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, L. Bourdev, R. Girshick, J. Hays, P. Perona, D. Ramanan, C. L. Zitnick, and P. Dollár · 2014
Earlier work this paper cites.
Compressing neural networks with the hashing trick
W. Chen, J. Wilson, S. Tyree, K. Weinberger, and Y. Chen · 2015
Earlier work this paper cites.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
K. He, X. Zhang, S. Ren, and J. Sun · 2015
Earlier work this paper cites.
Activitynet: A large-scale video benchmark for human activity understanding
F. C. Heilbron, V. Escorcia, B. Ghanem, and J. C. Niebles · 2015
Earlier work this paper cites.
Conditional computation in neural networks for faster models
E. Bengio, P.-L. Bacon, J. Pineau, and D. Precup · 2016
Earlier work this paper cites.
Adaptive computation time for recurrent neural networks
A. Graves · 2016
Earlier work this paper cites.
Squeezenet: Alexnet-level accuracy with 50x fewer parameters and
F. N. Iandola, S. Han, M. W. Moskewicz, K. Ashraf, W. J. Dally, and K. Keutzer · 2016
Earlier work this paper cites.
Xnor-net: Imagenet classification using binary convolutional neural networks
M. Rastegari, V. Ordonez, J. Redmon, and A. Farhadi · 2016
Earlier work this paper cites.
Branchynet: Fast inference via early exiting from deep neural networks
S. Teerapittayanon, B. McDanel, and H. Kung · 2016
Earlier work this paper cites.
Temporal segment networks: Towards good practices for deep action recognition
L. Wang, Y. Xiong, Z. Wang, Y. Qiao, D. Lin, X. Tang, and L. Van Gool · 2016
Earlier work this paper cites.
End-to-end learning of action detection from frame glimpses in videos
S. Yeung, O. Russakovsky, G. Mori, and L. Fei-Fei · 2016
Cited alongside, same era.
Real-time action recognition with enhanced motion vector cnns
B. Zhang, L. Wang, Z. Wang, Y. Qiao, and H. Wang · 2016
Cited alongside, same era.
Spatially adaptive computation time for residual networks
M. Figurnov, M. D. Collins, Y. Zhu, L. Zhang, J. Huang, D. Vetrov, and R. Salakhutdinov · 2017
Cited alongside, same era.
Mask r-cnn
K. He, G. Gkioxari, P. Dollár, and R. Girshick · 2017
Cited alongside, same era.
Mobilenets: Efficient convolutional neural networks for mobile vision applications
A. G. Howard, M. Zhu, B. Chen, D. Kalenichenko, W. Wang, T. Weyand, M. Andreetto, and H. Adam · 2017
Cited alongside, same era.
Categorical reparameterization with gumbel-softmax
E. Jang, S. Gu, and B. Poole · 2017
Cited alongside, same era.
Multi-scale dense convolutional networks for efficient prediction
G. Huang, D. Chen, T. Li, F. Wu, L. van der Maaten, and K. Q. Weinberger · 2018
Later among the works it cites.
Exploiting feature and class relationships in video categorization with regularized deep neural networks
Y.-G. Jiang, Z. Wu, J. Wang, X. Xue, and S.-F. Chang · 2018
Later among the works it cites.
Mobilenetv2: Inverted residuals and linear bottlenecks
M. Sandler, A. Howard, M. Zhu, A. Zhmoginov, and L.-C. Chen · 2018
Later among the works it cites.
Sniper: Efficient multi-scale training
B. Singh, M. Najibi, and L. S. Davis · 2018
Later among the works it cites.
Convolutional networks with adaptive inference graphs
A. Veit and S. Belongie · 2018
Later among the works it cites.
Non-local neural networks
X. Wang, R. Girshick, A. Gupta, and K. He · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Pruning filters for efficient convnets
H. Li, A. Kadav, I. Durdanovic, H. Samet, and H. P. Graf · 2017
Cited alongside, same era.
Dynamic deep neural networks: Optimizing accuracy-efficiency trade-offs by selective execution
L. Liu and J. Deng · 2017
Cited alongside, same era.
The Concrete Distribution: A Continuous Relaxation of Discrete Random Variables
C. J. Maddison, A. Mnih, and Y. W. Teh · 2017
Cited alongside, same era.
Deciding how to decide: Dynamic routing in artificial neural networks
M. McGill and P. Perona · 2017
Cited alongside, same era.
Aggregated residual transformations for deep neural networks
S. Xie, R. Girshick, P. Dollár, Z. Tu, and K. He · 2017
Cited alongside, same era.
Watching a small portion could be as good as watching all: Towards efficient video classification
H. Fan, Z. Xu, L. Zhu, C. Yan, J. Ge, and Y. Yang · 2018
Cited alongside, same era.
Skipnet: Learning dynamic routing in convolutional networks
X. Wang, F. Yu, Z.-Y. Dou, and J. E. Gonzalez · 2018
Later among the works it cites.
Compressed video action recognition
C.-Y. Wu, M. Zaheer, H. Hu, R. Manmatha, A. J. Smola, and P. Krähenbühl · 2018
Later among the works it cites.
Blockdrop: Dynamic inference paths in residual networks
Z. Wu, T. Nagarajan, A. Kumar, S. Rennie, L. S. Davis, K. Grauman, and R. Feris · 2018
Later among the works it cites.
Eco: Efficient convolutional network for online video understanding
M. Zolfaghari, K. Singh, and T. Brox · 2018
Later among the works it cites.
Slowfast networks for video recognition
C. Feichtenhofer, H. Fan, J. Malik, and K. He · 2019
Closest in time.
Scsampler: Sampling salient clips from video for efficient action recognition
B. Korbar, D. Tran, and L. Torresani · 2019
Closest in time.
Autofocus: Efficient multi-scale inference
M. Najibi, B. Singh, and L. S. Davis · 2019
Closest in time.
Adaframe: Adaptive frame selection for fast video recognition
Z. Wu, C. Xiong, C.-Y. Ma, R. Socher, and L. S. Davis · 2019
Closest in time.