Fetching the paper…
Reading the bibliography…
We introduce Egocentric Object Manipulation Graphs (Ego-OMG) - a novel representation for activity modeling and anticipation of near future actions integrating three components: 1) semantic temporal structure of activities, 2) short-term dynamics, and 3) representations for appearance.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Recognizing complex events using large margin joint low-level event model
Hamid Izadinia and Mubarak Shah · 2012
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman · 2014
Earlier work this paper cites.
A hierarchical representation for future action prediction
Tian Lan, Tsung-Chuan Chen, and Silvio Savarese · 2014
Earlier work this paper cites.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning · 2014
Earlier work this paper cites.
Generating notifications for missing actions: Don’t forget to turn the lights off!
Bilge Soran, Ali Farhadi, and Linda Shapiro · 2015
Earlier work this paper cites.
Flowing convnets for human pose estimation in videos
Tomas Pfister, James Charles, and Andrew Zisserman · 2015
Earlier work this paper cites.
Conditional random fields as recurrent neural networks
Shuai Zheng, Sadeep Jayasumana, Bernardino Romera-Paredes, Vibhav Vineet, Zhizhong Su, Dalong Du, Chang Huang, and Philip HS Torr · 2015
Earlier work this paper cites.
Anticipating visual representations from unlabeled video
Carl Vondrick, Hamed Pirsiavash, and Antonio Torralba · 2016
Earlier work this paper cites.
Structural-rnn: Deep learning on spatio-temporal graphs
Ashesh Jain, Amir R Zamir, Silvio Savarese, and Ashutosh Saxena · 2016
Earlier work this paper cites.
Semi-supervised classification with graph convolutional networks
Thomas N Kipf and Max Welling · 2016
Earlier work this paper cites.
Temporal action localization in untrimmed videos via multi-stage cnns
Zheng Shou, Dongang Wang, and Shih-Fu Chang · 2016
Earlier work this paper cites.
Temporal segment networks: Towards good practices for deep action recognition
Limin Wang, Yuanjun Xiong, Zhe Wang, Yu Qiao, Dahua Lin, Xiaoou Tang, and Luc Van Gool · 2016
Cited alongside, same era.
Quo vadis, action recognition? a new model and the kinetics dataset
Joao Carreira and Andrew Zisserman · 2017
Cited alongside, same era.
Visual forecasting by imitating dynamics in natural sequences
Kuo-Hao Zeng, William B Shen, De-An Huang, Min Sun, and Juan Carlos Niebles · 2017
Cited alongside, same era.
Next-active-object prediction from egocentric videos
Antonino Furnari, Sebastiano Battiato, Kristen Grauman, and Giovanni Maria Farinella · 2017
Cited alongside, same era.
Xception: Deep learning with depthwise separable convolutions
François Chollet · 2017
Cited alongside, same era.
Encoding crowd interaction with deep neural network for pedestrian trajectory prediction
What would you expect? anticipating egocentric actions with rolling-unrolling lstms and modality attention
Antonino Furnari and Giovanni Maria Farinella · 2019
Later among the works it cites.
Forecasting human object interaction: Joint prediction of motor attention and egocentric activity
Miao Liu, Siyu Tang, Yin Li, and James Rehg · 2019
Later among the works it cites.
Leveraging the present to anticipate the future in videos
Antoine Miech, Ivan Laptev, Josef Sivic, Heng Wang, Lorenzo Torresani, and Du Tran · 2019
Later among the works it cites.
Predicting the future: A jointly learnt model for action anticipation
Harshala Gammulle, Simon Denman, Sridha Sridharan, and Clinton Fookes · 2019
Later among the works it cites.
Human motion anticipation with symbolic label
Julian Tanke and Juergen Gall · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yanyu Xu, Zhixin Piao, and Shenghua Gao · 2018
Cited alongside, same era.
Scaling egocentric vision: The epic-kitchens dataset
Dima Damen, Hazel Doughty, Giovanni Maria Farinella, Sanja Fidler, Antonino Furnari, Evangelos Kazakos, Davide Moltisanti, Jonathan Munro, Toby Perrett, Will Price, et al · 2018
Cited alongside, same era.
In the eye of beholder: Joint learning of gaze and actions in first person video
Yin Li, Miao Liu, and James M Rehg · 2018
Cited alongside, same era.
Videos as space-time region graphs
Xiaolong Wang and Abhinav Gupta · 2018
Cited alongside, same era.
Spatial temporal graph convolutional networks for skeleton-based action recognition
Sijie Yan, Yuanjun Xiong, and Dahua Lin · 2018
Cited alongside, same era.
Leveraging uncertainty to rethink loss functions and evaluation measures for egocentric action anticipation
Antonino Furnari, Sebastiano Battiato, and Giovanni Maria Farinella · 2018
Cited alongside, same era.
Video classification with channel-separated convolutional networks
Du Tran, Heng Wang, Lorenzo Torresani, and Matt Feiszli · 2019
Cited alongside, same era.
Zero-shot anticipation for instructional activities
Fadime Sener and Angela Yao · 2019
Later among the works it cites.
Graph-based global reasoning networks
Yunpeng Chen, Marcus Rohrbach, Zhicheng Yan, Yan Shuicheng, Jiashi Feng, and Yannis Kalantidis · 2019
Later among the works it cites.
Large-scale weakly-supervised pre-training for video action recognition
Deepti Ghadiyaram, Du Tran, and Dhruv Mahajan · 2019
Later among the works it cites.
Ego-topo: Environment affordances from egocentric video
Tushar Nagarajan, Yanghao Li, Christoph Feichtenhofer, and Kristen Grauman · 2020
Closest in time.
Contact anticipation maps for forecasting hand object interaction
A. Anonymous · 2020
Closest in time.
Knowledge distillation for action anticipation via label smoothing
Guglielmo Camporese, Pasquale Coscia, Antonino Furnari, Giovanni Maria Farinella, and Lamberto Ballan · 2020
Closest in time.