Fetching the paper…
Reading the bibliography…
Wearable cameras are becoming more and more popular in several applications, increasing the interest of the research community in developing approaches for recognizing actions from the first-person point of view.
M. Ma, H. Fan, and K. M. Kitani, “Going deeper into first-person activity recognition,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2016, pp. 1894–1903
1903
Earlier work this paper cites.
1906
Earlier work this paper cites.
G. Farnebäck, “Two-frame motion estimation based on polynomial expansion,” in Image Analysis , J. Bigun and T. Gustavsson, Eds. Berlin, Heidelberg: Springer Berlin Heidelberg, 2003, pp. 363–370
2003
Earlier work this paper cites.
A. Fathi, X. Ren, and J. M. Rehg, “Learning to recognize objects in egocentric activities,” in CVPR 2011 . IEEE, Jun. 2011
2011
Earlier work this paper cites.
H. Pirsiavash and D. Ramanan, “Detecting activities of daily living in first-person camera views,” in Proc CVPR , 2012
2012
Earlier work this paper cites.
H. Wang and C. Schmid, “Action recognition with improved trajectories,” in IEEE International Conference on Computer Vision , Sydney, Australia, 2013
2013
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Two-stream convolutional networks for action recognition in videos,” in Advances in neural information processing systems , 2014, pp. 568–576
2014
Earlier work this paper cites.
R. Ghirshick, “Fast r-cnn,” in proc ICCV , 2015
2015
Earlier work this paper cites.
X. Wang and A. Gupta, “Unsupervised learning of visual representations using videos,” in International Conference on Computer Vision (ICCV) , 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
J. Walker, A. Gupta, and M. Hebert, “Dense optical flow prediction from a static image,” in 2015 IEEE International Conference on Computer Vision (ICCV) , 2015, pp. 2443–2451
2015
Earlier work this paper cites.
R. D. Geest, E. Gavves, A. Ghodrati, Z. Li, C. Snoek, and T. Tuytelaars, “Online action detection,” in Proc ECCV , 2016
2016
Earlier work this paper cites.
S. Ma, L. Sigal, and S. Sclaroff, “Learning activity progression in lstm for activity detection and early detection,” in Proc CVPR , 2016
2016
Earlier work this paper cites.
D. Pathak, P. Krähenbühl, J. Donahue, T. Darrell, and A. Efros, “Context encoders: Feature learning by inpainting,” in Computer Vision and Pattern Recognition (CVPR) , 2016
2016
Earlier work this paper cites.
L. Wang, Y. Xiong, Z. Wang, Y. Qiao, D. Lin, X. Tang, and L. Van Gool, “Temporal segment networks: Towards good practices for deep action recognition,” in Computer Vision – ECCV 2016 , B. Leibe, J. Matas, N. Sebe, and M. Welling, Eds. Cham: Springer International Publishing, 2016, pp. 20–36
2016
Earlier work this paper cites.
X. Zhang, Y. Wang, M. Gou, M. Sznaier, and O. Camps, “Efficient temporal sequence comparison and classification using gram matrix embeddings on a riemannian manifold,” in 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 4498–4507
2016
Earlier work this paper cites.
S. Sudhakaran and O. Lanz, “Convolutional long short-term memory networks for recognizing first person interactions,” in Proceedings of the IEEE International Conference on Computer Vision Workshops , 2017, pp. 2339–2346
2017
Earlier work this paper cites.
D. Pathak, R. B. Girshick, P. Dollár, T. Darrell, and B. Hariharan, “Learning features by watching objects move,” in 2017 IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2017, Honolulu, HI, USA, July 21-26, 2017 , 2017, pp. 6024–6033
2017
Cited alongside, same era.
D. Damen, H. Doughty, G. M. Farinella, S. Fidler, A. Furnari, E. Kazakos, D. Moltisanti, J. Munro, T. Perrett, W. Price, and M. Wray, “Scaling egocentric vision: The dataset,” in Computer Vision – ECCV 2018 , V. Ferrari, M. Hebert, C. Sminchisescu, and Y. Weiss, Eds. Cham: Springer International Publishing, 2018, pp. 753–771
2018
Cited alongside, same era.
S. Sudhakaran and O. Lanz, “Attention is all we need: Nailing down object-centric attention for egocentric activity recognition,” in British Machine Vision Conference , 2018
2018
Cited alongside, same era.
Y. Li, M. Liu, and J. M. Rehg, “In the eye of beholder: Joint learning of gaze and actions in first person video,” in The European Conference on Computer Vision (ECCV) , September 2018
J. Lin, C. Gan, and S. Han, “Tsm: Temporal shift module for efficient video understanding,” in Proceedings of the IEEE International Conference on Computer Vision , 2019, pp. 7083–7093
2019
Later among the works it cites.
L. Jing and Y. Tian, “Self-supervised visual feature learning with deep neural networks: A survey,” arXiv preprint:1902.06162 , 2019
2019
Later among the works it cites.
A. Mahendran, J. Thewlis, and A. Vedaldi, “Cross pixel optical-flow similarity for self-supervised learning,” in Computer Vision – ACCV 2018 , C. Jawahar, H. Li, G. Mori, and K. Schindler, Eds. Cham: Springer International Publishing, 2019, pp. 99–116
2019
Later among the works it cites.
J. Wang, J. Jiao, L. Bao, S. He, Y. Liu, and W. Liu, “Self-supervised spatio-temporal representation learning for videos by predicting motion and appearance statistics,” 2019
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
2018
Cited alongside, same era.
S. Gidaris, P. Singh, and N. Komodakis, “Unsupervised representation learning by predicting image rotations,” in ICLR , 2018
2018
Cited alongside, same era.
M. Caron, P. Bojanowski, A. Joulin, and M. Douze, “Deep clustering for unsupervised learning of visual features,” in European Conference on Computer Vision (ECCV) , 2018
2018
Cited alongside, same era.
M. Noroozi, A. Vinjimoor, P. Favaro, and H. Pirsiavash, “Boosting self-supervised learning via knowledge transfer,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2018
2018
Cited alongside, same era.
M. Lee, S. Lee, S. Son, G. Park, and N. Kwak, “Motion feature network: Fixed motion filter for action recognition,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2018, pp. 387–403
2018
Cited alongside, same era.
S. Sun, Z. Kuang, L. Sheng, W. Ouyang, and W. Zhang, “Optical flow guided feature: A fast and robust motion representation for video action recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2018, pp. 1390–1399
2018
Cited alongside, same era.
G. Garcia-Hernando, S. Yuan, S. Baek, and T.-K. Kim, “First-person hand action benchmark with rgb-d videos and 3d hand pose annotations,” in Proceedings of Computer Vision and Pattern Recognition (CVPR) , 2018
2018
Cited alongside, same era.
S. Sudhakaran, S. Escalera, and O. Lanz, “LSTA: Long Short-Term Attention for Egocentric Action Recognition,” in The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , June 2019
2019
Cited alongside, same era.
J. Zhao and C. G. M. Snoek, “Dance with flow: Two-in-one stream action detection,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , June 2019
2019
Later among the works it cites.
N. Crasto, P. Weinzaepfel, K. Alahari, and C. Schmid, “MARS: Motion-Augmented RGB Stream for Action Recognition,” in CVPR , 2019
2019
Later among the works it cites.
L. Wang, Y. Xiong, Z. Wang, Y. Qiao, D. Lin, X. Tang, and L. V. Gool, “Temporal segment networks for action recognition in videos,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 41, no. 11, pp. 2740–2755, 2019
2019
Later among the works it cites.
B. Tekin, F. Bogo, and M. Pollefeys, “H+ o: Unified egocentric recognition of 3d hand-object poses and interactions,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2019, pp. 4511–4520
2019
Later among the works it cites.
X. S. Nguyen, L. Brun, O. Lézoray, and S. Bougleux, “A neural network based on spd manifold learning for skeleton-based hand gesture recognition,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2019, pp. 12 036–12 045
2019
Later among the works it cites.
S. Sudhakaran, S. Escalera, and O. Lanz, “Fbk-hupba submission to the epic-kitchens 2019 action recognition challenge,” 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
J.-M. Perez-Rua, B. Martinez, X. Zhu, A. Toisoul, V. Escorcia, and T. Xiang, “Knowing what, where and when to look: Efficient video action modeling with attention,” 2020
2020
Closest in time.
J. Munro and D. Damen, “Multi-modal domain adaptation for fine-grained action recognition,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , June 2020
2020
Closest in time.
S. Sudhakaran, S. Escalera, and O. Lanz, “Gate-shift networks for video action recognition,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 1102–1111
2020
Closest in time.
J. Munro and D. Damen, “Multi-modal domain adaptation for fine-grained action recognition,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 122–132
2020
Closest in time.
A. Furnari and G. Farinella, “Rolling-unrolling lstms for action anticipation from first-person video,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2020
2020
Closest in time.