Fetching the paper…
Reading the bibliography…
Derived from rapid advances in computer vision and machine learning, video analysis tasks have been moving from inferring the present state to predicting the future state.
Boston: Houghton Mifflin (1971)
Mass, J., Johansson, G., Jason, G., Runeson, S.: Motion perception I and II [film] · 1971
Earlier work this paper cites.
Bull. Psychon. Soc. 9
Cutting, J., Kozlowski, L.: Recognition of friends by their work: gait perception without familarity cues · 1977
Earlier work this paper cites.
Artificial Intelligence 17
Horn, B., Schunck, B.: Determining optical flow · 1981
Earlier work this paper cites.
In: Proceedings of Imaging Understanding Workshop (1981)
Lucas, B.D., Kanade, T.: An iterative image registration technique with an application to stereo vision · 1981
Earlier work this paper cites.
In: Alvey Vision Conference (1988)
Harris, C., Stephens., M.: A combined corner and edge detector · 1988
Earlier work this paper cites.
Trends in Neurosciences 15
Goodale, M.A., Milner, A.D.: Separate visual pathways for perception and action · 1992
Earlier work this paper cites.
Chicago: University of Chicago Press (1992)
Ricoeur, P.: Oneself as another (K. Blamey, Trans.) · 1992
Earlier work this paper cites.
Neural mechanisms of perception and action 3
Decety, J., Grezes, J.: Neural mechanisms subserving the perception of human actions · 1999
Earlier work this paper cites.
Swiss J. Psychol. 59
Sumi, S.: Perception of point-light walker produced by eight lights attached to the back of the walker · 2000
Earlier work this paper cites.
IEEE Trans Pattern Analysis and Machine Intelligence 23
Bobick, A., Davis, J.: The recognition of human movement using temporal templates · 2001
Earlier work this paper cites.
J. Vis. 2
Troje, N.: Decomposing biological motion: a framework for analysis and synthesis of human gait patterns · 2002
Earlier work this paper cites.
In: ICCV, vol. 2, pp. 726 –733 (2003)
Efros, A., Berg, A., Mori, G., Malik, J.: Recognizing action at a distance · 2003
Earlier work this paper cites.
In: ICCV, pp. 432–439 (2003)
Laptev, I., Lindeberg, T.: Space-time interest points · 2003
Earlier work this paper cites.
In: ICML (2004)
Abbeel, P., Ng, A.: Apprenticeship learning via inverse reinforcement learning · 2004
Earlier work this paper cites.
Annu. Rev. Neurosci. 27
Rizzolatti, G., Craighero, L.: The mirror-neuron system · 2004
Earlier work this paper cites.
In: IEEE ICPR (2004)
Schüldt, C., Laptev, I., Caputo, B.: Recognizing human actions: A local svm approach · 2004
Earlier work this paper cites.
In: Proc. ICCV (2005)
Blank, M., Gorelick, L., Shechtman, E., Irani, M., Basri, R.: Actions as space-time shapes · 2005
Earlier work this paper cites.
Perception 24
Clarke, T., Bradshaw, M., Field, D., Hampson, S., Rose, D.: The perception of emotion from body movement in point-light displays of interpersonal dialogue · 2005
Earlier work this paper cites.
In: CVPR (2005)
Dalal, N., Triggs, B.: Histograms of oriented gradients for human detection · 2005
Earlier work this paper cites.
In: ICCV VS-PETS (2005)
Dollar, P., Rabaud, V., Cottrell, G., Belongie, S.: Behavior recognition via sparse spatio-temporal features · 2005
Earlier work this paper cites.
In: CVPR (2005)
Duong, T.V., Bui, H.H., Phung, D.Q., Venkatesh, S.: Activity recognition and abnormality detection with the switching hidden semi-markov model · 2005
Earlier work this paper cites.
In: CVPR (2005)
Fanti, C., Zelnik-Manor, L., Perona, P.: Hybrid models for human motion recognition · 2005
Earlier work this paper cites.
In: International Conference on Computer Vision (2005)
Sminchisescu, C., Kanaujia, A., Li, Z., Metaxas, D.: Conditional models for contextual human motion recognition · 2005
Earlier work this paper cites.
Percept. Psychophys 67
Troje, N., Westhoff, C., Lavrov, M.: Person identification from biological motion: effects of structural and kinematic cues · 2005
Earlier work this paper cites.
In: CVPR (2005)
Yilmaz, A., Shah, M.: Actions sketch: A novel action representation · 2005
Earlier work this paper cites.
In: CVPR (2006)
Perronnin, F., Dance, C.: Fisher kernels on visual vocabularies for image categorization · 2006
Earlier work this paper cites.
In: CVPR, vol. 2, pp. 1709–1718 (2006)
Ryoo, M., Aggarwal, J.: Recognition of composite human activities through context-free grammar based representation · 2006
Earlier work this paper cites.
In: CVPR (2006)
Wang, S.B., Quattoni, A., Morency, L.P., Demirdjian, D., Darrell, T.: Hidden conditional random fields for gesture recognition · 2006
Earlier work this paper cites.
Computer Vision and Image Understanding 104
Weinland, D., Ronfard, R., Boyer, E.: Free viewpoint action recognition using motion history volumes · 2006
Earlier work this paper cites.
Annu. Rev. Psychol. 58
Blake, R., Shiffrar, M.: Perception of human motion · 2007
Earlier work this paper cites.
Transactions on Pattern Analysis and Machine Intelligence 29
Gorelick, L., Blank, M., Shechtman, E., Irani, M., Basri, R.: Actions as space-time shapes · 2007
Earlier work this paper cites.
Image Processing, IEEE Transactions on 16
Hu, W., Xie, D., Fu, Z., Zeng, W., Maybank, S.: Semantic-based surveillance video retrieval · 2007
Earlier work this paper cites.
In: CVPR (2007)
Ikizler, N., Forsyth, D.: Searching video for complex activities with finite state models · 2007
Earlier work this paper cites.
In: ICCV (2007)
Laptev, I., Perez, P.: Retrieving actions in movies · 2007
Earlier work this paper cites.
In: CVPR (2007)
Morency, L.P., Quattoni, A., Darrell, T.: Latent-dynamic discriminative models for continuous gesture recognition · 2007
Earlier work this paper cites.
In: CVPR (2007)
Niebles, J.C., Fei-Fei, L.: A hierarchical model of shape and appearance for human action classification · 2007
Earlier work this paper cites.
In: CVPR (2007)
Rajko, S., Qian, G., Ingalls, T., James, J.: Real-time gesture recognition with minimal training requirements and on-line learning · 2007
Earlier work this paper cites.
In: Proc. ACM Multimedia (2007)
Scovanner, P., Ali, S., Shah, M.: A 3-dimensional sift descriptor and its application to action recognition · 2007
Earlier work this paper cites.
In: CVPR (2007)
Wang, L., Suter, D.: Recognizing human activities from silhouettes: Motion subspace and factorial discriminative graphical model · 2007
Earlier work this paper cites.
In: CVPR (2007)
Wong, S.F., Kim, T.K., Cipolla, R.: Learning motion categories using both semantic and structural information · 2007
Earlier work this paper cites.
In: CVPR (2008)
Felzenszwalb, P., McAllester, D., Ramanan, D.: A discriminatively trained, multiscale, deformable part model · 2008
Earlier work this paper cites.
In: ECCV (2008)
Huang, D.A., Kitani, K.M.: Action-reaction: Forecasting the dynamics of human interaction · 2008
Earlier work this paper cites.
In: CVPR (2008)
Jia, K., Yeung, D.Y.: Human action recognition using local spatio-temporal discriminant embedding · 2008
Earlier work this paper cites.
In: BMVC (2008)
Klaser, A., Marszalek, M., Schmid, C.: A spatio-temporal descriptor based on 3d-gradients · 2008
Earlier work this paper cites.
In: CVPR (2008)
Laptev, I., Marszalek, M., Schmid, C., Rozenfeld, B.: Learning realistic human actions from movies · 2008
Earlier work this paper cites.
Laptev, I., Marszałek, M., Schmid, C., Rozenfeld, B.: Learning realistic human actions from movies (2008)
2008
Earlier work this paper cites.
International Journal of Computer Vision 79
Niebles, J.C., Wang, H., Fei-Fei, L.: Unsupervised learning of human action categories using spatial-temporal words · 2008
Earlier work this paper cites.
In: CVPR (2008)
Rodriguez, M.D., Ahmed, J., Shah, M.: Action mach: A spatio-temporal maximum average correlation height filter for action recognition · 2008
Earlier work this paper cites.
In: ECCV (2008)
Tran, D., Sorokin, A.: Human activity recognition with metric learning · 2008
Earlier work this paper cites.
In: NIPS (2008)
Wang, Y., Mori, G.: Learning a discriminative hidden part model for human action recognition · 2008
Earlier work this paper cites.
In: ECCV (2008)
Willems, G., Tuytelaars, T., Gool, L.: An efficient dense and scale-invariant spatio-temporal interest poing detector · 2008
Earlier work this paper cites.
In: AAAI (2008)
Ziebart, B., Maas, A., Bagnell, J., Dey, A.: Maximum entropy inverse reinforcement learning · 2008
Earlier work this paper cites.
In: CVPR (2009)
Bregonzio, M., Gong, S., Xiang, T.: Recognizing action as clouds of space-time interest points · 2009
Earlier work this paper cites.
In: Computer Vision Workshops (ICCV Workshops), 2009 IEEE 12th International Conference on, pp. 1282 –1289 (2009)
Choi, W., Shahid, K., Savarese, S.: What are they doing? : Collective activity classification using spatio-temporal relationship among people · 2009
Earlier work this paper cites.
In: 2009 IEEE 12th International Conference on Computer Vision, pp. 1491–1498. IEEE (2009)
Duchenne, O., Laptev, I., Sivic, J., Bach, F., Ponce, J.: Automatic annotation of human actions in video · 2009
Earlier work this paper cites.
In: CVPR (2009)
Jingen Liu, J.L., Shah, M.: Recognizing realistic actions from videos ”in the wild” · 2009
Earlier work this paper cites.
In: Proc. IEEE Conf. on Computer Vision and Pattern Recognition (2009)
Liu, J., Luo, J., Shah, M.: Recognizing realistic actions from videos “in the wild” · 2009
Earlier work this paper cites.
In: IEEE Conference on Computer Vision & Pattern Recognition (2009)
Marszałek, M., Laptev, I., Schmid, C.: Actions in context · 2009
Earlier work this paper cites.
In: ICCV (2009)
Messing, R., Pal, C., Kautz, H.: Activity recognition using the velocity histories of tracked keypoints · 2009
Earlier work this paper cites.
In: ICCV, pp. 1593–1600 (2009)
Ryoo, M., Aggarwal, J.: Spatio-temporal relationship match: Video structure comparison for recognition of complex human activities · 2009
Earlier work this paper cites.
In: CVPR (2009)
Sun, J., Wu, X., Yan, S., Cheong, L., Chua, T., Li, J.: Hierarchical spatio-temporal context modeling for action recognition · 2009
Earlier work this paper cites.
In: BMVC (2009)
Wang, H., Ullah, M.M., Kl a \>{a} ser, A., Laptev, I., Schmid, C.: Evaluation of local spatio-temporal features for action recognition · 2009
Earlier work this paper cites.
In: CVPR (2009)
Yeffet, L., Wolf, L.: Local trinary patterns for human action recognition · 2009
Earlier work this paper cites.
In: IEEE Conference on Computer Vision and Pattern Recognition (2009)
Yuan, J., Liu, Z., Wu, Y.: Discriminative subvolume search for efficient action detection · 2009
Earlier work this paper cites.
In: IROS (2009)
Ziebart, B., Ratliff, N., Gallagher, G., Mertz, C., Peterson, K., Bagnell, J., Hebert, M., Dey, A., Srinivasa, S.: Planning-based prediction for pedestrians · 2009
Earlier work this paper cites.
In: ICML (2010)
Ji, S., Xu, W., Yang, M., Yu, K.: 3d convolutional neural networks for human action recognition · 2010
Earlier work this paper cites.
In: CVPR workshop (2010)
Li, W., Zhang, Z., Liu, Z.: Action recognition based on a bag of 3d points · 2010
Earlier work this paper cites.
In: ECCV (2010)
Niebles, J.C., Chen, C.W., Fei-Fei, L.: Modeling temporal structure of decomposable motion segments for activity classification · 2010
Earlier work this paper cites.
In: Proc. British Conference on Machine Vision (2010)
Patron-Perez, A., Marszalek, M., Zisserman, A., Reid, I.: High five: Recognising human interactions in tv shows · 2010
Earlier work this paper cites.
Image and Vision Computing 28
Poppe, R.: A survey on vision-based human action recognition · 2010
Earlier work this paper cites.
In: ECCV (2010)
Raptis, M., Soatto, S.: Tracklet descriptors for action modeling and video analysis · 2010
Earlier work this paper cites.
Nat. Rev. Neurosci. 11
Rizzolatti, G., Sinigaglia, C.: The functional role of the parieto-frontal mirror circuit: interpretations and misinterpretations · 2010
Earlier work this paper cites.
http://cvrc.ece.utexas.edu/SDHA2010/Human_Interaction.html (2010)
Ryoo, M.S., Aggarwal, J.K.: UT-Interaction Dataset, ICPR contest on Semantic Description of Human Activities (SDHA) · 2010
Earlier work this paper cites.
In: 2nd Workshop on Activity monitoring by multi-camera surveillance systems (AMMCSS), pp. 48–55 (2010)
S Singh, S.V., Ragheb, H.: Muhavi: A multicamera human action video dataset for the evaluation of action recognition methods · 2010
Earlier work this paper cites.
In: ECCV (2010)
Satkin, S., Hebert, M.: Modeling the temporal extent of actions · 2010
Earlier work this paper cites.
In: Advanced Video and Signal Based Surveillance (AVSS), 2010 Seventh IEEE International Conference on, pp. 48–55. IEEE (2010)
Singh, S., Velastin, S.A., Ragheb, H.: Muhavi: A multicamera human action video dataset for the evaluation of action recognition methods · 2010
Earlier work this paper cites.
In: CVPR (2010)
Sun, D., Roth, S., Black, M.J.: Secrets of optical flow estimation and their principles · 2010
Earlier work this paper cites.
In: ECCV (2010)
Taylor, G.W., Fergus, R., LeCun, Y., Bregler, C.: Convolutional learning of spatio-temporal features · 2010
Earlier work this paper cites.
In: ECCV (2010)
Turek, M., Hoogs, A., Collins, R.: Unsupervised learning of functional categories in video scenes · 2010
Earlier work this paper cites.
PAMI (2010)
Wang, Y., Mori, G.: Hidden part models for human action recognition: Probabilistic vs. max-margin · 2010
Earlier work this paper cites.
In: BMVC (2010)
Yu, T.H., Kim, T.K., Cipolla, R.: Real-time action recognition by spatiotemporal semantic and structural forests · 2010
Earlier work this paper cites.
IEEE Transactions on Pattern Analysis and Machine Intelligence (2010)
Yuan, J., Liu, Z., Wu, Y.: Discriminative video pattern search for efficient action detection · 2010
Earlier work this paper cites.
In: CVPR (2011)
Choi, W., Shahid, K., Savarese, S.: Learning context for collective activity recognition · 2011
Earlier work this paper cites.
In: ICRA (2011)
Dragan, A., Ratliff, N., Srinivasa, S.: Manipulation planning with goal sets using constrained trajectory optimization · 2011
Earlier work this paper cites.
In: ICCV (2011)
Kim, K., Lee, D., Essa, I.: Gaussian process regression flow for analysis of motion trajectories · 2011
Earlier work this paper cites.
In: ICCV (2011)
Kuehne, H., Jhuang, H., Garrote, E., Poggio, T., Serre, T.: Hmdb: A large video database for human motion recognition · 2011
Earlier work this paper cites.
In: CVPR (2011)
Le, Q.V., Zou, W.Y., Yeung, S.Y., Ng, A.Y.: Learning hierarchical invariant spatio-temporal features for action recognition with independent subspace analysis · 2011
Earlier work this paper cites.
In: CVPR (2011)
Liu, J., Kuipers, B., Savarese, S.: Recognizing human actions by attributes · 2011
Earlier work this paper cites.
Pattern Analysis and Machine Intelligence, IEEE Transactions on 33
Morrisand, B., Trivedi, M.: Trajectory learning for activity understanding: Unsupervised, multilevel, and long-term adaptive approach · 2011
Earlier work this paper cites.
In: ICCV Workshop on CDC3CV (2011)
Ni, B., Wang, G., Moulin, P.: RGBD-HuDaAct: A color-depth video database for human daily activity recognition · 2011
Earlier work this paper cites.
In: ICCV, pp. 487–494. IEEE (2011)
Pei, M., Jia, Y., Zhu, S.C.: Parsing video events with goal inference and intent prediction · 2011
Earlier work this paper cites.
In: IJCAI (2011)
Plotz, T., Hammerla, N.Y., Olivier, P.: Feature learning for activity recognition in ubiquitous computing · 2011
Earlier work this paper cites.
In: ICCV (2011)
Ryoo, M.S.: Human activity prediction: Early recognition of ongoing activities from streaming videos · 2011
Earlier work this paper cites.
In: AAAI workshop on Pattern, Activity and Intent Recognition (2011)
Sung, J., Ponce, C., Selman, B., Saxena, A.: Human activity detection from rgbd images · 2011
Earlier work this paper cites.
In: ICCV Workshops, pp. 1729 –1736 (2011)
Vahdat, A., Gao, B., Ranjbar, M., Mori, G.: A discriminative key pose sequence model for recognizing human interactions · 2011
Earlier work this paper cites.
In: IEEE Conference on Computer Vision & Pattern Recognition, pp. 3169–3176. Colorado Springs, United States (2011)
Wang, H., Kläser, A., Schmid, C., Liu, C.L.: Action Recognition by Dense Trajectories · 2011
Earlier work this paper cites.
In: CVPR (2011)
Wu, X., Xu, D., Duan, L., Luo, J.: Action recognition using context and appearance distribution features · 2011
Earlier work this paper cites.
In: CVPR (2011)
Zhou, B., Wang, X., Tang, X.: Random field topic model for semantic region analysis in crowded scenes from tracklets · 2011
Earlier work this paper cites.
In: ECCV, pp. 215–230. Springer (2012)
Choi, W., Savarese, S.: A unified framework for multi-target tracking and collective activity recognition · 2012
Earlier work this paper cites.
In: CVPR (2012)
Hoai, M., la Torre, F.D.: Max-margin early event detectors · 2012
Earlier work this paper cites.
CRCV-TR-12-01
Khurram Soomro, A.R.Z., Shah, M.: Ucf101: A dataset of 101 human action classes from videos in the wild (2012) · 2012
Earlier work this paper cites.
In: ECCV (2012)
Kitani, K.M., Ziebart, B.D., Bagnell, J.A., Hebert, M.: Activity forecasting · 2012
Earlier work this paper cites.
IEEE Transactions on Pattern Analysis and Machine Intelligence 34
Kliper-Gross, O., Hassner, T., Wolf, L.: The action similarity labeling challenge · 2012
Earlier work this paper cites.
In: Proc. European Conf. on Computer Vision (2012)
Kong, Y., Jia, Y., Fu, Y.: Learning human interaction by interactive phrases · 2012
Earlier work this paper cites.
In: EUSIPCO (2012)
Kurakin, A., Zhang, Z., Liu, Z.: A real-time system for dynamic hand gesture recognition with a depth sensor · 2012
Earlier work this paper cites.
In: CVPR (2012)
Lan, T., Sigal, L., Mori, G.: Social roles in hierarchical models for human activity · 2012
Earlier work this paper cites.
TPAMI 34
Lan, T., Wang, Y., Yang, W., Robinovitch, S.N., Mori, G.: Discriminative latent models for recognizing contextual group activities · 2012
Earlier work this paper cites.
In: ECCV (2012)
Li, K., Hu, J., Fu, Y.: Modeling complex temporal composition of actionlets for activity prediction · 2012
Earlier work this paper cites.
Machine Vision and Applications Journal (2012)
Reddy, K.K., Shah, M.: Recognizing 50 human action categories of web videos · 2012
Earlier work this paper cites.
IEEE transactions on pattern analysis and machine intelligence 35
Scheirer, W.J., de Rezende Rocha, A., Sapkota, A., Boult, T.E.: Toward open set recognition · 2012
Cited alongside, same era.
In: ICRA (2012)
Sung, J., Ponce, C., Selman, B., Saxena, A.: Unstructured human activity detection from rgbd images · 2012
Cited alongside, same era.
In: CVPR (2012)
Tang, K., Fei-Fei, L., Koller, D.: Learning latent temporal structure for complex event detection · 2012
Cited alongside, same era.
In: Advances in Neural Information Processing Systems (2012)
Tang, K., Ramanathan, V., Fei-Fei, L., Koller, D.: Shifting weights: Adapting object detectors from image to video · 2012
Cited alongside, same era.
In: ECCV (2012)
Wang, J., Liu, Z., Chorowski, J., Chen, Z., Wu, Y.: Robust 3d action recognition with random occupancy patterns · 2012
Cited alongside, same era.
In: CVPR (2012)
Wang, J., Liu, Z., Wu, Y., Yuan, J.: Mining actionlet ensemble for action recognition with depth cameras · 2012
Cited alongside, same era.
In: CVPR (2017)
Diba, A., Sharma, V., Gool, L.V.: Deep temporal linear encoding networks · 2017
Later among the works it cites.
In: CVPR (2017)
Duta, I.C., Ionescu, B., Aizawa, K., Sebe, N.: spatio-temporal vector of locally max pooled features for action recognition in videos · 2017
Later among the works it cites.
In: 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp. 7445–7454. IEEE (2017)
Feichtenhofer, C., Pinz, A., Wildes, R.P.: Spatiotemporal multiplier networks for video action recognition · 2017
Later among the works it cites.
In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp. 3636–3645 (2017)
Fernando, B., Bilen, H., Gavves, E., Gould, S.: Self-supervised video representation learning with odd-one-out networks · 2017
Later among the works it cites.
In: ICCV (2017)
Gao, J., Yang, Z., Chen, K., Sun, C., Nevatia, R.: TURN TAP: Temporal unit regression network for temporal action proposals · 2017
Later among the works it cites.
In: CVPR (2017)
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
In: CVPR (2012)
Wang, Z., Wang, J., Xiao, J., Lin, K.H., Huang, T.S.: Substructural and boundary modeling for continuous action recognition · 2012
Cited alongside, same era.
In: Computer Vision and Pattern Recognition Workshops (CVPRW), 2012 IEEE Computer Society Conference on, pp. 20–27. IEEE (2012)
Xia, L., Chen, C., Aggarwal, J.: View invariant human action recognition using histograms of 3d joints · 2012
Cited alongside, same era.
In: CVPRW (2012)
Xia, L., Chen, C.C., Aggarwal, J.K.: View invariant human action recognition using histograms of 3d joints · 2012
Cited alongside, same era.
In: ECCV (2012)
Yang, Y., Shah, M.: Complex events detection using data-driven concepts · 2012
Cited alongside, same era.
In: ECCV (2012)
Yao, B., Fei-Fei, L.: Action recognition with exemplar based 2.5d graph matching · 2012
Cited alongside, same era.
TPAMI 34
Yao, B., Fei-Fei, L.: Recognizing human-object interactions in still images by modeling the mutual context of objects and human poses · 2012
Cited alongside, same era.
Girdhar, R., Ramanan, D., Gupta, A., Sivic, J., Russell, B.: Actionvlad: Learning spatio-temporal aggregation for action classification · 2017
Later among the works it cites.
In: Proc. ICCV (2017)
Goyal, R., Kahou, S.E., Michalski, V., Materzynska, J., Westphal, S., Kim, H., Haenel, V., Fruend, I., Yianilos, P., Mueller-Freitag, M., et al.: The” something something” video database for learning and evaluating visual common sense · 2017
Later among the works it cites.
arXiv preprint arXiv:1705.08421 (2017)
Gu, C., Sun, C., Vijayanarasimhan, S., Pantofaru, C., Ross, D.A., Toderici, G., Li, Y., Ricco, S., Sukthankar, R., Schmid, C., et al.: Ava: A video dataset of spatio-temporally localized atomic visual actions · 2017
Later among the works it cites.
Image and Vision Computing (2017)
Herath, S., Harandi, M., Porikli, F.: Going deeper into action recognition: A survey · 2017
Later among the works it cites.
IEEE Transactions on Pattern Analysis and Machine Intelligence 40
Jiang, Y.G., Wu, Z., Wang, J., Xue, X., Chang, S.F.: Exploiting feature and class relationships in video categorization with regularized deep neural networks · 2017
Later among the works it cites.
In: BMVC (2017)
Jiyang Gao Zhenheng Yang, R.N.: Red: Reinforced encoder-decoder networks for action anticipation · 2017
Later among the works it cites.
In: CVPR (2017)
Kar, A., Rai, N., Sikka, K., Sharma, G.: Adascan: Adaptive scan pooling in deep convolutional neural networks for human action recognition in videos · 2017
Later among the works it cites.
arXiv preprint arXiv:1705.06950 (2017)
Kay, W., Carreira, J., Simonyan, K., Zhang, B., Hillier, C., Vijayanarasimhan, S., Viola, F., Green, T., Back, T., Natsev, P., et al.: The kinetics human action video dataset · 2017
Later among the works it cites.
In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp. 3288–3297 (2017)
Ke, Q., Bennamoun, M., An, S., Sohel, F., Boussaid, F.: A new representation of skeleton sequences for 3d action recognition · 2017
Later among the works it cites.
International Journal of Computer Vision (IJCV) 123
Kong, Y., Fu, Y.: Max-margin heterogeneous information machine for rgb-d action recognition · 2017
Later among the works it cites.
In: CVPR (2017)
Kong, Y., Tao, Z., Fu, Y.: Deep sequential context networks for action prediction · 2017
Later among the works it cites.
In: Proceedings of the IEEE International Conference on Computer Vision, pp. 667–676 (2017)
Lee, H.Y., Huang, J.B., Singh, M., Yang, M.H.: Unsupervised representation learning by sorting sequences · 2017
Later among the works it cites.
In: CVPR (2017)
Lee, N., Choi, W., Vernaza, P., Choy, C.B., Torr, P.H., Chandraker, M.: Desire: Distant future prediction in dynamic scenes with interacting agents · 2017
Later among the works it cites.
In: ICCV (2017)
Qiu, Z., Yao, T., Mei, T.: Learning spatio-temporal representation with pseudo-3d residual network · 2017
Later among the works it cites.
In: CVPR (2017)
Shou, Z., Chan, J., Zareian, A., Miyazawa, K., Chang, S.F.: CDC: Convolutional-de-convolutional networks for precise temporal action localization in untrimmed videos · 2017
Later among the works it cites.
In: IJCAI (2017)
Su, H., Zhu, J., Dong, Y., Zhang, B.: Forecast the plausible paths in crowd scenes · 2017
Later among the works it cites.
IEEE Transactions on Pattern Analysis and Machine Intelligence (2017)
Varol, G., Laptev, I., Schmid, C.: Long-term temporal convolutions for action recognition · 2017
Later among the works it cites.
In: CVPR (2017)
Wang, L., Xiong, Y., Lin, D., Van Gool, L.: UntrimmedNets for weakly supervised action recognition and detection · 2017
Later among the works it cites.
In: Proceedings of the IEEE international conference on computer vision, pp. 1329–1338 (2017)
Wang, X., He, K., Gupta, A.: Transitive invariance for self-supervised visual representation learning · 2017
Later among the works it cites.
In: Proceedings of the IEEE international conference on computer vision, pp. 5783–5792 (2017)
Xu, H., Das, A., Saenko, K.: R-c3d: Region convolutional 3d network for temporal activity detection · 2017
Later among the works it cites.
arXiv preprint arXiv:1712.09374 (2017)
Zhao, H., Yan, Z., Wang, H., Torresani, L., Torralba, A.: Slac: A sparsely labeled dataset for action classification and localization · 2017
Later among the works it cites.
In: ICCV (2017)
Zhao, Y., Xiong, Y., Wang, L., Wu, Z., Tang, X., Lin, D.: Temporal action detection with structured segment networks · 2017
Later among the works it cites.
In: Proceedings of the European Conference on Computer Vision (ECCV), pp. 770–786 (2018)
Buchler, U., Brattoli, B., Ommer, B.: Improving spatiotemporal self-supervision by deep reinforcement learning · 2018
Closest in time.
In: CVPR (2018)
Chao, Y.W., Vijayanarasimhan, S., Seybold, B., Ross, D.A., Deng, J., Sukthankar, R.: Rethinking the Faster R-CNN architecture for temporal action localization · 2018
Closest in time.
In: European Conference on Computer Vision (2018)
Damen, D., Doughty, H., Farinella, G.M., Fidler, S., Furnari, A., Kazakos, E., Moltisanti, D., Munro, J., Perrett, T., Price, W., Wray, M.: Scaling egocentric vision: The epic-kitchens dataset · 2018
Closest in time.
IEEE Sensors Journal 18
Dawar, N., Kehtarnavaz, N.: Action detection and recognition in continuous action streams by deep learning-based sensing fusion · 2018
Closest in time.
In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 5589–5597 (2018)
Gan, C., Gong, B., Liu, K., Su, H., Guibas, L.J.: Geometry guided convolutional neural networks for self-supervised video representation learning · 2018
Closest in time.
In: CVPR (2018)
Gu, C., Sun, C., Ross, D.A., Vondrick, C., Pantofaru, C., Li, Y., Vijayanarasimhan, S., Toderici, G., Ricco, S., Sukthankar, R., et al.: AVA: A video dataset of spatio-temporally localized atomic visual actions · 2018
Closest in time.
In: ECCV (2018)
Guo, M., Chou, E., Huang, D.A., Song, S., Yeung, S., Fei-Fei, L.: Neural graph matching networks for fewshot 3d action recognition · 2018
Closest in time.
In: CVPR (2018)
Gupta, A., Johnson, J., Fei-Fei, L., Savarese, S., Alahi, A.: Social gan: Socially acceptable trajectories with generative adversarial networks · 2018
Closest in time.
In: AAAI (2018)
Kong, Y., Gao, S., Sun, B., Fu, Y.: Action prediction from videos via memorizing hard-to-predict samples · 2018
Closest in time.
IEEE TPAMI (2018)
Kong, Y., Tao, Z., Fu, Y.: Adversarial action prediction networks · 2018
Closest in time.
IEEE Transactions on Image Processing 27
Lai, S., Zhang, W.S., Hu, J.F., Zhang, J.: Global-local temporal saliency action prediction · 2018
Closest in time.
In: Proceedings of the European Conference on Computer Vision (ECCV), pp. 3–19 (2018)
Lin, T., Zhao, X., Su, H., Wang, C., Yang, M.: Bsn: Boundary sensitive network for temporal action proposal generation · 2018
Closest in time.
In: ECCV (2018)
Luo, Z., Hsieh, J.T., Jiang, L., Carlos Niebles, J., Fei-Fei, L.: Graph distillation for action detection with privileged modalities · 2018
Closest in time.
Mishra, A., Verma, V., Reddy, M.K.K., Subramaniam, A., Rai, P., Mittal, A.: A generative approach to zero-shot and few-shot action recognition (2018)
2018
Closest in time.
In: ICME (2018)
Shu, Y., Shi, Y., Wang, Y., Zou, Y., Yuan, Q., Tian, Y.: ODN: Opening the deep network for open-set action recognition · 2018
Closest in time.
IEEE Transactions on Image Processing (TIP) 27
Song, S., Lan, C., Xing, J., Zeng, W., Liu, J.: Spatio-temporal attention-based LSTM networks for 3d action recognition and detection · 2018
Closest in time.
In: Thirty-Second AAAI Conference on Artificial Intelligence (2018)
Yan, S., Xiong, Y., Lin, D.: Spatial temporal graph convolutional networks for skeleton-based action recognition · 2018
Closest in time.
In: CVPR (2018)
Yang, H., He, X., Porikli, F.: One-shot action localization by learning sequence matching network · 2018
Closest in time.
In: Proceedings of the European Conference on Computer Vision (ECCV), pp. 803–818 (2018)
Zhou, B., Andonian, A., Oliva, A., Torralba, A.: Temporal relational reasoning in videos · 2018
Closest in time.
In: ECCV (2018)
Zhu, L., Yang, Y.: Compound memory networks for few-shot video classification · 2018
Closest in time.
In: BMVC (2019)
Bishay, M., Zoumpourlis, G., Patras, I.: Tarn: Temporal attentive relation network for few-shot and zero-shot action recognition · 2019
Closest in time.
In: ICCVW (2019)
Dwivedi, S.K., Gupta, V., Mitra, R., Ahmed, S., Jain, A.: Protogan: Towards few shot learning for action recognition · 2019
Closest in time.
In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 12046–12055 (2019)
Ghadiyaram, D., Tran, D., Mahajan, D.: Large-scale weakly-supervised pre-training for video action recognition · 2019
Closest in time.
In: CVPR (2019)
Ke, Q., Fritz, M., Schiele, B.: Time-conditioned action anticipation in one shot · 2019
Closest in time.
arXiv preprint arXiv:1907.03395 (2019)
Kosaraju, V., Sadeghian, A., Martín-Martín, R., Reid, I., Rezatofighi, S.H., Savarese, S.: Social-bigat: Multimodal trajectory forecasting using bicycle-gan and graph attention networks · 2019
Closest in time.
In: 2019 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), pp. 6150–6156. IEEE (2019)
Li, J., Ma, H., Tomizuka, M.: Conditional generative neural system for probabilistic trajectory prediction · 2019
Closest in time.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 5725–5734 (2019)
Liang, J., Jiang, L., Niebles, J.C., Hauptmann, A.G., Fei-Fei, L.: Peeking into the future: Predicting future person activities and locations in videos · 2019
Closest in time.
In: Proceedings of the IEEE/CVF International Conference on Computer Vision, pp. 3889–3898 (2019)
Lin, T., Liu, X., Li, X., Ding, E., Wen, S.: Bmn: Boundary-matching network for temporal action proposal generation · 2019
Closest in time.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 3604–3613 (2019)
Liu, Y., Ma, L., Zhang, Y., Liu, W., Chang, S.F.: Multi-granularity generator for temporal action proposal · 2019
Closest in time.
In: CVPR (2019)
Mehrasa, N., Jyothi, A.A., Durand, T., He, J., Sigal, L., Mori, G.: A variational auto-encoder model for stochastic point processes · 2019
Closest in time.
In: ICCV (2019)
Narayan, S., Cholakkal, H., Khan, F.S., Shao, L.: 3C-Net: Category count and center loss for weakly-supervised action localization · 2019
Closest in time.
In: CVPR (2019)
Oza, P., Patel, V.M.: C2AE: Class conditioned auto-encoder for open-set recognition · 2019
Closest in time.
In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 12056–12065 (2019)
Qiu, Z., Yao, T., Ngo, C.W., Tian, X., Mei, T.: Learning spatio-temporal representation with local and global diffusion · 2019
Closest in time.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 1349–1358 (2019)
Sadeghian, A., Kosaraju, V., Sadeghian, A., Hirose, N., Rezatofighi, H., Savarese, S.: Sophie: An attentive gan for predicting paths compliant to social and physical constraints · 2019
Closest in time.
In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 1227–1236 (2019)
Si, C., Chen, W., Wang, W., Wang, L., Tan, T.: An attention enhanced graph convolutional lstm network for skeleton-based action recognition · 2019
Closest in time.
IEEE Transactions on Multimedia (TMM) 21
Song, H., Wu, X., Zhu, B., Wu, Y., Chen, M., Jia, Y.: Temporal action localization in untrimmed videos using action pattern trees · 2019
Closest in time.
In: CVPR (2019)
Song, L., Zhang, S., Yu, G., Sun, H.: TACNet: Transition-aware context network for spatio-temporal action detection · 2019
Closest in time.
In: CVPR (2019)
Tang, Y., Ding, D., Rao, Y., Zheng, Y., Zhang, D., Zhao, L., Lu, J., Zhou, J.: COIN: A large-scale dataset for comprehensive instructional video analysis · 2019
Closest in time.
In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 10334–10343 (2019)
Xu, D., Xiao, J., Zhao, Z., Shao, J., Xie, D., Zhuang, Y.: Self-supervised spatiotemporal learning via video clip order prediction · 2019
Closest in time.
IEEE transactions on pattern analysis and machine intelligence 41
Xu, H., Das, A., Saenko, K.: Two-stream region convolutional 3d network for temporal activity detection · 2019
Closest in time.
In: ICCV (2019)
Xu, M., Gao, M., Chen, Y.T., Davis, L.S., Crandall, D.J.: Temporal recurrent networks for online action detection · 2019
Closest in time.
In: CVPR (2019)
Yang, X., Yang, X., Liu, M.Y., Xiao, F., Davis, L.S., Kautz, J.: STEP: Spatio-temporal progressive learning for video action detection · 2019
Closest in time.
Pattern Recognition 85
Yang, Y., Hou, C., Lang, Y., Guan, D., Huang, D., Xu, J.: Open-set human activity recognition based on micro-doppler signatures · 2019
Closest in time.
In: ICCV (2019)
Yu, T., Ren, Z., Li, Y., Yan, E., Xu, N., Yuan, J.: Temporal structure mining for weakly supervised action detection · 2019
Closest in time.
In: ICCV (2019)
Zeng, R., Huang, W., Tan, M., Rong, Y., Zhao, P., Huang, J., Gan, C.: Graph convolutional networks for temporal action localization · 2019
Closest in time.
In: ICCV (2019)
Zhao, H., Torralba, A., Torresani, L., Yan, Z.: HACS: Human action clips and segments dataset for recognition and temporal localization · 2019
Closest in time.
In: CVPR (2020)
Cao, K., Ji, J., Cao, Z., Chang, C.Y., Niebles, J.C.: Few-shot video classification via temporal alignment · 2020
Closest in time.
In: ECCV (2020)
Chen, G., Qiao, L., Shi, Y., Peng, P., Li, J., Huang, T., Pu, S., Tian, Y.: Learning open set network with discriminative reciprocal points · 2020
Closest in time.
IEEE Transactions on Pattern Analysis and Machine Intelligence (PAMI) (2020)
Furnari, A., Farinella, G.M.: Rolling-unrolling lstms for action anticipation from first-person video · 2020
Closest in time.
IEEE transactions on pattern analysis and machine intelligence (2020)
Geng, C., Huang, S.j., Chen, S.: Recent advances in open set recognition: A survey · 2020
Closest in time.
IEEE Transactions on Pattern Analysis and Machine Intelligence 42
Liu, J., Shahroudy, A., Perez, M., Wang, G., Duan, L.Y., Kot, A.C.: Ntu rgb+d 120: A large-scale benchmark for 3d human activity understanding · 2020
Closest in time.
arXiv preprint arXiv:2012.11717 (2020)
Liu, Y., Yan, Q., Alahi, A.: Social nce: Contrastive learning of socially-aware motion representations · 2020
Closest in time.
arXiv preprint arXiv:2012.01526 (2020)
Mangalam, K., An, Y., Girase, H., Malik, J.: From goals, waypoints & paths to long term human trajectory forecasting · 2020
Closest in time.
In: European Conference on Computer Vision, pp. 759–776. Springer (2020)
Mangalam, K., Girase, H., Agarwal, S., Lee, K.H., Adeli, E., Malik, J., Gaidon, A.: It is not the journey but the destination: Endpoint conditioned trajectory prediction · 2020
Closest in time.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 7143–7152 (2020)
Marchetti, F., Becattini, F., Seidenari, L., Bimbo, A.D.: Mantra: Memory augmented networks for multiple trajectory prediction · 2020
Closest in time.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 14424–14432 (2020)
Mohamed, A., Qian, K., Elhoseiny, M., Claudel, C.: Social-stgcnn: A social spatio-temporal graph convolutional neural network for human trajectory prediction · 2020
Closest in time.
In: CVPR (2020)
Perera, P., Morariu, V.I., Jain, R., Manjunatha, V., Wigington, C., Ordonez, V., Patel, V.M.: Generative-discriminative feature representations for open-set recognition · 2020
Closest in time.
In: IVS (2020)
Roitberg, A., Ma, C., Haurilet, M., Stiefelhagen, R.: Open set driver activity recognition · 2020
Closest in time.
In: ECCV (2020)
Zhang, H., Zhang, L., Qi, X., Li, H., Torr, P.H.S., Koniusz, P.: Few-shot action recognition with permutation-invariant attention · 2020
Closest in time.
In: ICCV (2021)
Bao, W., Yu, Q., Kong, Y.: Evidential deep learning for open set action recognition · 2021
Closest in time.
In: CVPR (2021)
Bhattacharyya, A., Reino, D.O., Fritz, M., Schiele, B.: Euro-pvi: Pedestrian vehicle interactions in dense urban centers · 2021
Closest in time.
In: Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), pp. 8178–8187 (2021)
Chen, S., Sun, P., Xie, E., Ge, C., Wu, J., Ma, L., Shen, J., Luo, P.: Watch only once: An end-to-end video action detection framework · 2021
Closest in time.
In: ICCV (2021)
Chung, J., hsin Wuu, C., ru Yang, H., Tai, Y.W., Tang, C.K.: Haa500: Human-centric atomic action dataset with curated videos · 2021
Closest in time.
In: ICCV (2021)
Dendorfer, P., Elflein, S., Leal-Taixé, L.: Mg-gan: A multi-generator model preventing out-of-distribution samples in pedestrian trajectory prediction · 2021
Closest in time.
In: CVPR (2021)
Fernando, B., Herath, S.: Anticipating human actions by correlating past with the future with jaccard similarity measures · 2021
Closest in time.
In: ICCV (2021)
Girase, H., Gang, H., Malla, S., Li, J., Kanehara, A., Mangalam, K., Choi, C.: Loki: Long term and key intentions for trajectory prediction · 2021
Closest in time.
In: 2020 25th International Conference on Pattern Recognition (ICPR), pp. 10335–10342. IEEE (2021)
Giuliari, F., Hasan, I., Cristani, M., Galasso, F.: Transformer networks for trajectory forecasting · 2021
Closest in time.
In: Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (2021)
Ke, Q., Fritz, M., Schiele, B.: Future moment assessment for action query · 2021
Closest in time.
In: ICCV (2021)
Li, Y., Chen, L., He, R., Wang, Z., Wu, G., Wang, L.: Multisports: A multi-person video dataset of spatio-temporally localized sports actions · 2021
Closest in time.
In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp. 4751–4760 (2021)
Li, Z., Yao, L.: Three birds with one stone: Multi-task temporal action detection via recycling temporal annotations · 2021
Closest in time.
In: CVPR (2021)
Liu, X., Pintea, S.L., Nejadasl, F.K., Booij, O., van Gemert, J.C.: No frame left behind: Full video action recognition · 2021
Closest in time.
In: CVPR (2021)
Mingfei Gao Yingbo Zhou, R.X.R.S.C.X.: Woad: Weakly supervised online action detection in untrimmed videos · 2021
Closest in time.
In: CVPR (2021)
Narayanan, S., Moslemi, R., Pittaluga, F., Liu, B., Chandraker, M.: Divide-and-conquer for lane-aware diverse trajectory prediction · 2021
Closest in time.
In: CVPR (2021)
Perrett, T., Masullo, A., Burghardt, T., Mirmehdi, M., Damen, D.: Temporal-relational crosstransformers for few-shot action recognition · 2021
Closest in time.
In: CVPR (2021)
Rasouli, A., Rohani, M., Luo, J.: Bifold and semantic reasoning for pedestrian behavior prediction · 2021
Closest in time.
In: ICCV (2021)
Rohit, G., Kristen, G.: Anticipative video transformer · 2021
Closest in time.
In: CVPR (2021)
Surís, D., Liu, R., Vondrick, C.: Learning the predictability of the future · 2021
Closest in time.
arXiv preprint arXiv:2103.14107 (2021)
Wang, C., Wang, Y., Xu, M., Crandall, D.J.: Stepwise goal-driven networks for trajectory prediction · 2021
Closest in time.
In: CVPR, pp. 1895–1904 (2021)
Wang, L., Tong, Z., Ji, B., Wu, G.: Tdn: Temporal difference networks for efficient action recognition · 2021
Closest in time.
In: CVPR (2021)
Yang, W., Zhang, T., Yu, X., Qi, T., Zhang, Y., Wu, F.: Uncertainty guided collaborative training for weakly supervised temporal action detection · 2021
Closest in time.
arXiv preprint arXiv:2103.14023 (2021)
Yuan, Y., Weng, X., Ou, Y., Kitani, K.: Agentformer: Agent-aware transformers for socio-temporal multi-agent forecasting · 2021
Closest in time.
In: ICCV (2021)
Zhao, H., Wildes, R.P.: Where are you heading? dynamic trajectory prediction with expert goal examples · 2021
Closest in time.