Fetching the paper…
Reading the bibliography…
Anticipating actions before they are executed is crucial for a wide range of practical applications, including autonomous driving and robotics.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
M. Gutmann and A. Hyvärinen, “Noise-contrastive estimation: A new estimation principle for unnormalized statistical models,” in AISTATS , 2010, pp. 297–304
2010
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Two-stream convolutional networks for action recognition in videos,” in NIPS , 2014
2014
Earlier work this paper cites.
D. Tran, L. Bourdev, R. Fergus, L. Torresani, and M. Paluri, “Learning spatiotemporal features with 3d convolutional networks,” in ICCV , 2015
2015
Earlier work this paper cites.
H. S. Koppula and A. Saxena, “Anticipating human activities using object affordances for reactive robotic response,” T-PAMI , vol. 38, no. 1, pp. 14–29, 2015
2015
Earlier work this paper cites.
S. Ioffe and C. Szegedy, “Batch normalization: Accelerating deep network training by reducing internal covariate shift,” in ICML , 2015
2015
Earlier work this paper cites.
L. Wang, Y. Xiong, Z. Wang, Y. Qiao, D. Lin, X. Tang, and L. Val Gool, “Temporal segment networks: Towards good practices for deep action recognition,” in ECCV , 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
R. De Geest, E. Gavves, A. Ghodrati, Z. Li, C. Snoek, and T. Tuytelaars, “Online action detection,” in ECCV , 2016
2016
Earlier work this paper cites.
S. Ma, L. Sigal, and S. Sclaroff, “Learning activity progression in lstms for activity detection and early detection,” in CVPR , 2016
2016
Earlier work this paper cites.
S. Huang, X. Li, Z. Zhang, Z. He, F. Wu, W. Liu, J. Tang, and Y. Zhuang, “Deep learning driven visual path prediction from a single image,” IEEE Transactions on Image Processing , vol. 25, no. 12, pp. 5892–5904, Dec 2016
2016
Earlier work this paper cites.
C. Vondrick, H. Pirsiavash, and A. Torralba, “Anticipating visual representations from unlabeled video,” in CVPR , 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
L. Zhu, Z. Xu, and Y. Yang, “Bidirectional multirate reconstruction for temporal modeling in videos,” in CVPR , 2017
2017
Earlier work this paper cites.
M. Sadegh Aliakbarian, F. Sadat Saleh, M. Salzmann, B. Fernando, L. Petersson, and L. Andersson, “Encouraging lstms to anticipate actions very early,” in ICCV , 2017
2017
Earlier work this paper cites.
F. Becattini, T. Uricchio, L. Seidenari, A. Del Bimbo, and L. Ballan, “Am I done? predicting action progress in videos,” BMVC , 2017
2017
Earlier work this paper cites.
J. Gao, Z. Yang, and R. Nevatia, “Red: Reinforced encoder-decoder networks for action anticipation,” BMVC , 2017
2017
Earlier work this paper cites.
P. Felsen, P. Agrawal, and J. Malik, “What will happen next? forecasting player moves in sports videos,” in ICCV , 2017
2017
Earlier work this paper cites.
A. Furnari, S. Battiato, K. Grauman, and G. M. Farinella, “Next-active-object prediction from egocentric videos,” Journal of Visual Communication and Image Representation , vol. 49, pp. 401–411, 2017
2017
Earlier work this paper cites.
T. Mahmud, M. Hasan, and A. K. Roy-Chowdhury, “Joint prediction of activity labels and starting times in untrimmed videos,” in ICCV , 2017
2017
Earlier work this paper cites.
K.-H. Zeng, W. B. Shen, D.-A. Huang, M. Sun, and J. Carlos Niebles, “Visual forecasting by imitating dynamics in natural sequences,” in ICCV , 2017
2017
Cited alongside, same era.
N. Rhinehart and K. M. Kitani, “First-person activity forecasting with online inverse reinforcement learning,” in ICCV , 2017
2017
Cited alongside, same era.
A. Paszke, S. Gross, S. Chintala, G. Chanan, E. Yang, Z. DeVito, Z. Lin, A. Desmaison, L. Antiga, and A. Lerer, “Automatic differentiation in pytorch,” 2017
2017
Cited alongside, same era.
L. Chen, J. Lu, Z. Song, and J. Zhou, “Part-activated deep reinforcement learning for action prediction,” in ECCV , 2018, pp. 421–436
2018
Cited alongside, same era.
D. Damen, H. Doughty, G. M. Farinella, S. Fidler, A. Furnari, E. Kazakos, D. Moltisanti, J. Munro, T. Perrett, W. Price, and M. Wray, “Scaling egocentric vision: The epic-kitchens dataset,” in ECCV , 2018
A. Miech, I. Laptev, J. Sivic, H. Wang, L. Torresani, and D. Tran, “Leveraging the present to anticipate the future in videos,” in CVPR-W , 2019
2019
Later among the works it cites.
A. Furnari and G. M. Farinella, “What would you expect? anticipating egocentric actions with rolling-unrolling lstms and modality attention.” in ICCV , 2019
2019
Later among the works it cites.
Y. Tang, Z. Wang, J. Lu, J. Feng, and J. Zhou, “Multi-stream deep neural networks for rgb-d egocentric action recognition,” IEEE Transactions on Circuits and Systems for Video Technology , vol. 29, no. 10, pp. 3001–3015, Oct 2019
2019
Later among the works it cites.
H. Gammulle, S. Denman, S. Sridharan, and C. Fookes, “Predicting the future: A jointly learnt model for action anticipation,” in ICCV , 2019, pp. 5562–5571
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
W. Liu, W. Luo, D. Lian, and S. Gao, “Future frame prediction for anomaly detection–a new baseline,” in CVPR , 2018, pp. 6536–6545
2018
Cited alongside, same era.
P. Luc, C. Couprie, Y. Lecun, and J. Verbeek, “Predicting future instance segmentation by forecasting convolutional features,” in ECCV , 2018, pp. 584–599
2018
Cited alongside, same era.
Y. Shi, B. Fernando, and R. Hartley, “Action anticipation with rbf kernelized feature mapping rnn,” in ECCV , 2018, pp. 301–317
2018
Cited alongside, same era.
C. Rodriguez, B. Fernando, and H. Li, “Action anticipation by predicting future dynamic images,” in ECCV-W , 2018
2018
Cited alongside, same era.
Y. Li, M. Liu, and J. M. Rehg, “In the eye of beholder: Joint learning of gaze and actions in first person video,” in ECCV , 2018
2018
Cited alongside, same era.
S. Sudhakaran and O. Lanz, “Attention is all we need: Nailing down object-centric attention for egocentric activity recognition,” BMVC , 2018
2018
Cited alongside, same era.
R. De Geest and T. Tuytelaars, “Modeling temporal structure with lstm for online action detection,” in WACV , 2018
2018
Cited alongside, same era.
2019
Later among the works it cites.
S. Sudhakaran, S. Escalera, and O. Lanz, “Lsta: Long short-term attention for egocentric action recognition,” in CVPR , 2019
2019
Later among the works it cites.
J. Liang, L. Jiang, J. C. Niebles, A. G. Hauptmann, and L. Fei-Fei, “Peeking into the future: Predicting future person activities and locations in videos,” in CVPR , 2019, pp. 5725–5734
2019
Later among the works it cites.
C. C. d. Santos, P. Moreno, J. L. A. Samatelo, R. F. Vassallo, and J. Santos-Victor, “Action anticipation for collaborative environments: The impact of contextual information and uncertainty-based prediction,” Neurocomputing , 2019
2019
Later among the works it cites.
Y. Abu Farha and J. Gall, “Uncertainty-aware anticipation of activities,” in ICCV-W , 2019
2019
Later among the works it cites.
Y. Wu, Y. Lin, X. Dong, Y. Yan, W. Bian, and Y. Yang, “Progressive learning for person re-identification with one example,” IEEE Transactions on Image Processing , vol. 28, no. 6, pp. 2872–2881, 2019
2019
Later among the works it cites.
X. Wang, L. Zhu, Y. Wu, and Y. Yang, “Symbiotic attention for egocentric action recognition with object-centric alignment,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2020
2020
Later among the works it cites.
B. F. Yan Bin Ng, “Forecasting future action sequences with attention: a new approach to weakly supervised action forecasting,” IEEE Transactions on Image Processing , 2020
2020
Later among the works it cites.
Q. Ke, M. Bennamoun, H. Rahmani, S. An, F. Sohel, and F. Boussaid, “Learning latent global network for skeleton-based action prediction,” IEEE Transactions on Image Processing , vol. 29, pp. 959–970, 2020
2020
Later among the works it cites.
F. Sener, D. Singhania, and A. Yao, “Temporal aggregate representations for long-range video understanding,” in ECCV , 2020
2020
Later among the works it cites.
T. Nagarajan, Y. Li, C. Feichtenhofer, and K. Grauman, “Ego-topo: Environment affordances from egocentric video,” in CVPR , 2020, pp. 163–172
2020
Later among the works it cites.
2020
Later among the works it cites.
T. Chen, S. Kornblith, M. Norouzi, and G. Hinton, “A simple framework for contrastive learning of visual representations,” in ICLR , 2020
2020
Later among the works it cites.
T. Han, W. Xie, and A. Zisserman, “Memory-augmented dense predictive coding for video representation learning,” in ECCV , 2020
2020
Later among the works it cites.
K. He, H. Fan, Y. Wu, S. Xie, and R. Girshick, “Momentum contrast for unsupervised visual representation learning,” in CVPR , 2020, pp. 9729–9738
2020
Later among the works it cites.