Fetching the paper…
Reading the bibliography…
We bring together ideas from recent work on feature design for egocentric action recognition under one framework by exploring the use of deep convolutional neural networks (CNN).
On space-time interest points
I. Laptev · 2005
Earlier work this paper cites.
A duality based approach for realtime tv-l 1 optical flow
C. Zach, T. Pock, and H. Bischof · 2007
Earlier work this paper cites.
Action recognition by learning mid-level motion features
A. Fathi and G. Mori · 2008
Earlier work this paper cites.
Learning realistic human actions from movies
I. Laptev, M. Marszałek, C. Schmid, and B. Rozenfeld · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Observing human-object interactions: Using spatial and functional compatibility for recognition
A. Gupta, A. Kembhavi, and L. S. Davis · 2009
Earlier work this paper cites.
Temporal segmentation and activity classification from first-person sensing
E. H. Spriggs, F. De La Torre, and M. Hebert · 2009
Earlier work this paper cites.
Improving the fisher kernel for large-scale image classification
F. Perronnin, J. Sánchez, and T. Mensink · 2010
Earlier work this paper cites.
Human activity analysis: A review
J. K. Aggarwal and M. S. Ryoo · 2011
Earlier work this paper cites.
Understanding egocentric activities
A. Fathi, A. Farhadi, and J. M. Rehg · 2011
Earlier work this paper cites.
Learning to recognize objects in egocentric activities
A. Fathi, X. Ren, and J. M. Rehg · 2011
Earlier work this paper cites.
Fast unsupervised ego-action learning for first-person sports videos
K. M. Kitani, T. Okabe, Y. Sato, and A. Sugimoto · 2011
Earlier work this paper cites.
Action recognition by dense trajectories
H. Wang, A. Kläser, C. Schmid, and C.-L. Liu · 2011
Earlier work this paper cites.
Learning to recognize daily actions using gaze
A. Fathi, Y. Li, and J. M. Rehg · 2012
Cited alongside, same era.
Discovering important people and objects for egocentric video summarization
J. Ghosh, Y. J. Lee, and K. Grauman · 2012
Cited alongside, same era.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Cited alongside, same era.
Detecting activities of daily living in first-person camera views
H. Pirsiavash and D. Ramanan · 2012
Cited alongside, same era.
Ucf101: A dataset of 101 human actions classes from videos in the wild
K. Soomro, A. R. Zamir, and M. Shah · 2012
Cited alongside, same era.
Modeling actions through state changes
A. Fathi and J. M. Rehg · 2013
Cited alongside, same era.
Caffe: Convolutional architecture for fast feature embedding
Y. Jia, E. Shelhamer, J. Donahue, S. Karayev, J. Long, R. Girshick, S. Guadarrama, and T. Darrell · 2014
Later among the works it cites.
The role of context for object detection and semantic segmentation in the wild
R. Mottaghi, X. Chen, X. Liu, N.-G. Cho, S.-W. Lee, S. Fidler, R. Urtasun, et al · 2014
Later among the works it cites.
Learning and transferring mid-level image representations using convolutional neural networks
M. Oquab, L. Bottou, I. Laptev, and J. Sivic · 2014
Later among the works it cites.
Bag of visual words and fusion methods for action recognition: Comprehensive study and good practice
X. Peng, L. Wang, X. Wang, and Y. Qiao · 2014
Later among the works it cites.
Two-stream convolutional networks for action recognition in videos
K. Simonyan and A. Zisserman · 2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
3d convolutional neural networks for human action recognition
S. Ji, W. Xu, M. Yang, and K. Yu · 2013
Cited alongside, same era.
Pixel-level hand detection in ego-centric videos
C. Li and K. M. Kitani · 2013
Cited alongside, same era.
Learning to predict gaze in egocentric video
Y. Li, A. Fathi, and J. M. Rehg · 2013
Cited alongside, same era.
Dense trajectories and motion boundary descriptors for action recognition
H. Wang, A. Kläser, C. Schmid, and C.-L. Liu · 2013
Cited alongside, same era.
Action recognition with improved trajectories
H. Wang and C. Schmid · 2013
Cited alongside, same era.
Return of the devil in the details: Delving deep into convolutional nets
K. Chatfield, K. Simonyan, A. Vedaldi, and A. Zisserman · 2014
Cited alongside, same era.
S. Bambach, S. Lee, D. J. Crandall, and C. Yu · 2015
Later among the works it cites.
Delving into egocentric actions
Y. Li, Z. Ye, and J. M. Rehg · 2015
Later among the works it cites.
Fully convolutional networks for semantic segmentation
J. Long, E. Shelhamer, and T. Darrell · 2015
Later among the works it cites.
Flowing convnets for human pose estimation in videos
T. Pfister, J. Charles, and A. Zisserman · 2015
Later among the works it cites.
Pooled motion features for first-person videos
M. S. Ryoo, B. Rothrock, and L. Matthies · 2015
Later among the works it cites.
Action recognition with trajectory-pooled deep-convolutional descriptors
L. Wang, Y. Qiao, and X. Tang · 2015
Later among the works it cites.
Compact cnn for indexing egocentric videos
Y. Poleg, A. Ephrat, S. Peleg, and C. Arora · 2016
Closest in time.