Feichtenhofer, C., Pinz, A., Zisserman, A.: Convolutional Two-Stream Network Fusion for Video Action Recognition. In: Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition. vol. 2016-Decem, pp. 1933–1941. IEEE Computer Society (12 2016). https://doi.org/10.1109/CVPR.2016.213
2016
Cited alongside, same era.
Hendrycks, D., Gimpel, K.: Bridging nonlinearities and stochastic regularizers with gaussian error linear units. CoRR abs/1606.08415
Original
2016
Cited alongside, same era.
Carreira, J., Zisserman, A.: Quo Vadis, action recognition? A new model and the kinetics dataset. In: Proceedings - 30th IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2017. vol. 2017-Janua, pp. 4724–4733. Institute of Electrical and Electronics Engineers Inc. (11 2017). https://doi.org/10.1109/CVPR.2017.502
2017
Cited alongside, same era.
Du, W., Wang, Y., Qiao, Y.: Recurrent spatial-temporal attention network for action recognition in videos. IEEE Transactions on Image Processing 27
2017
Cited alongside, same era.
Girdhar, R., Ramanan, D., Gupta, A., Sivic, J., Russell, B.: ActionVLAD: Learning spatio-temporal aggregation for action classification. In: Proceedings - 30th IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2017 (2017). https://doi.org/10.1109/CVPR.2017.337
2017
Cited alongside, same era.
Howard, A.G., Zhu, M., Chen, B., Kalenichenko, D., Wang, W., Weyand, T., Andreetto, M., Adam, H.: MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications (4 2017), http://arxiv.org/abs/1704.04861
Original
2017
Cited alongside, same era.
Li, Z., Gavrilyuk, K., Gavves, E., Jain, M., Snoek, C.G.: VideoLSTM convolves, attends and flows for action recognition. Computer Vision and Image Understanding 166
2017
Cited alongside, same era.
Loshchilov, I., Hutter, F.: Decoupled Weight Decay Regularization. 7th International Conference on Learning Representations, ICLR 2019 (11 2017), http://arxiv.org/abs/1711.05101
Original
2017
Cited alongside, same era.
Luo, W., Li, Y., Urtasun, R., Zemel, R.: Understanding the Effective Receptive Field in Deep Convolutional Neural Networks. Advances in Neural Information Processing Systems pp. 4905–4913 (1 2017), http://arxiv.org/abs/1701.04128
Original
2017
Cited alongside, same era.
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A.N., Kaiser, Å., Polosukhin, I.: Attention is all you need. In: Advances in Neural Information Processing Systems. vol. 2017-Decem, pp. 5999–6009. Neural information processing systems foundation (2017)
2017
Cited alongside, same era.
Chen, Y., Kalantidis, Y., Li, J., Yan, S., Feng, J.: Multi-fiber Networks for Video Recognition. In: Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics). vol. 11205 LNCS, pp. 364–380. Springer Verlag (7 2018). https://doi.org/10.1007/978-3-030-01246-5_22
2018
Cited alongside, same era.
Devlin, J., Chang, M.W., Lee, K., Toutanova, K.: BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding (10 2018), http://arxiv.org/abs/1810.04805
Original
2018
Cited alongside, same era.