Fetching the paper…
Reading the bibliography…
In this paper we present an approach for classifying the activity performed by a group of people in a video sequence.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
M. Schuster and K. K. Paliwal, “Bidirectional recurrent neural networks,” IEEE Transcations on Signal Processing , vol. 45, no. 11, pp. 2673–2681, 1997
1997
Earlier work this paper cites.
S. S. Intille and A. Bobick, “Recognizing planned, multiperson action,” Computer Vision and Image Understanding (CVIU) , vol. 81, pp. 414–445, 2001
2001
Earlier work this paper cites.
C. Schüldt, I. Laptev, and B. Caputo, “Recognizing human actions: a local svm approach,” in International Conference on Pattern Recognition, ICPR , vol. 3. IEEE, 2004, pp. 32–36
2004
Earlier work this paper cites.
P. Nillius, J. Sullivan, and S. Carlsson, “Multi-target tracking-linking identities using bayesian network inference,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2006
2006
Earlier work this paper cites.
S. Lazebnik, C. Schmid, and J. Ponce, “Beyond bags of features: Spatial pyramid matching for recognizing natural scene categories,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2006
2006
Earlier work this paper cites.
A. Klaser, M. Marszałek, and C. Schmid, “A spatio-temporal descriptor based on 3d-gradients,” in British Machine Vision Conference (BMVC)) . British Machine Vision Association, 2008, pp. 275–1
2008
Earlier work this paper cites.
H. Wang and C. Schmid, “Action recognition with improved trajectories,” in IEEE International Conference on Computer Vision (ICCV) , Sydney, Australia, 2013. [Online]. Available: http://hal.inria.fr/hal-00873267
2008
Earlier work this paper cites.
W. Choi, K. Shahid, and S. Savarese, “What are they doing?: Collective activity classification using spatio-temporal relationship among people,” in IEEE International Conference on Computer Vision Workshops (ICCV Workshops) . IEEE, 2009, pp. 1282–1289
2009
Earlier work this paper cites.
B. Siddiquie, Y. Yacoob, and L. Davis, “Recognizing plays in american football videos,” Technical report, University of Maryland, Tech. Rep., 2009
2009
Earlier work this paper cites.
D. E. King, “Dlib-ml: A machine learning toolkit,” Journal of Machine Learning Research , vol. 10, pp. 1755–1758, 2009
2009
Earlier work this paper cites.
R. Poppe, “A survey on vision-based human action recognition,” Image and Vision Computing , vol. 28, no. 6, pp. 976–990, 2010
2010
Earlier work this paper cites.
L.-J. Li, H. Su, L. Fei-Fei, and E. P. Xing, “Object bank: A high-level image representation for scene classification & semantic feature sparsification,” in Neural Information Processing Systems (NIPS) , 2010, pp. 1378–1386
2010
Earlier work this paper cites.
M. S. Ryoo and J. K. Aggarwal, “UT-Interaction Dataset, ICPR contest on Semantic Description of Human Activities (SDHA),” http://cvrc.ece.utexas.edu/SDHA2010/Human_Interaction.html, 2010
2010
Earlier work this paper cites.
H. Wang, A. Kläser, C. Schmid, and C.-L. Liu, “Action recognition by dense trajectories,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE, 2011, pp. 3169–3176
2011
Earlier work this paper cites.
D. Weinland, R. Ronfard, and E. Boyer, “A survey of vision-based methods for action representation, segmentation and recognition,” Computer Vision and Image Understanding , vol. 115, no. 2, pp. 224–241, 2011
2011
Earlier work this paper cites.
M.-C. Chang, N. Krahnstoever, and W. Ge, “Probabilistic group-level motion analysis and scenario recognition,” in IEEE Internation Conference on Computer Vision (ICCV) . IEEE, 2011, pp. 747–754
2011
Earlier work this paper cites.
W. Choi, K. Shahid, and S. Savarese, “Learning context for collective activity recognition,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE, 2011, pp. 3273–3280
2011
Earlier work this paper cites.
V. I. Morariu and L. S. Davis, “Multi-agent event recognition in structured scenarios.” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2011
2011
Earlier work this paper cites.
H. Kuehne, H. Jhuang, E. Garrote, T. Poggio, and T. Serre, “Hmdb: a large video database for human motion recognition,” in IEEE International Conference on Computer Vision (ICCV) . IEEE, 2011, pp. 2556–2563
2011
Earlier work this paper cites.
S. Oh, A. Hoogs, A. Perera, N. Cuntoor, C.-C. Chen, J. T. Lee, S. Mukherjee, J. Aggarwal, H. Lee, L. Davis, E. Swears, X. Wang, Q. Ji, K. Reddy, M. Shah, C. Vondrick, H. Pirsiavash, D. Ramanan, J. Yuen, A. Torralba, B. Song, A. Fong, A. Roy-Chowdhury, and M. Desai, “A large-scale benchmark dataset for event recognition in surveillance video,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2011
2011
Earlier work this paper cites.
T. Lan, L. Sigal, and G. Mori, “Social roles in hierarchical models for human activity recognition,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE, 2012, pp. 1354–1361
2012
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in Neural Information Processing Systems (NIPS) , 2012, pp. 1097–1105
2012
Cited alongside, same era.
T. Lan, Y. Wang, W. Yang, S. Robinovitch, and G. Mori, “Discriminative latent models for recognizing contextual group activities,” IEEE Transactions on Pattern Analysis and Machine Intelligence (PAMI) , vol. 34, no. 8, pp. 1549–1562, 2012
2012
Cited alongside, same era.
T. Lan, L. Sigal, and G. Mori, “Social roles in hierarchical models for human activity recognition,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2012
2012
Cited alongside, same era.
X. Zhu and D. Ramanan, “Face detection, pose estimation, and landmark localization in the wild,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE, 2012, pp. 2879–2886
2012
Cited alongside, same era.
2014
Later among the works it cites.
J. J. Tompson, A. Jain, Y. LeCun, and C. Bregler, “Joint training of a convolutional network and a graphical model for human pose estimation,” in Advances in Neural Information Processing Systems 27 , Z. Ghahramani, M. Welling, C. Cortes, N. Lawrence, and K. Weinberger, Eds. Curran Associates, Inc., 2014, pp. 1799–1807
2014
Later among the works it cites.
M. Danelljan, G. Häger, F. Shahbaz Khan, and M. Felsberg, “Accurate scale estimation for robust visual tracking,” in British Machine Vision Conference (BMVC) , 2014
2014
Later among the works it cites.
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich, “Going deeper with convolutions,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015
2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
W. Choi and S. Savarese, “A unified framework for multi-target tracking and collective activity recognition,” in European Conference on Computer Vision (ECCV) . Springer, 2012, pp. 215–230
2012
Cited alongside, same era.
C. Direkoglu and N. E. O’Connor, “Team activity recognition in sports,” in European Conference on Computer Vision (ECCV) . Springer, 2012, pp. 69–83
2012
Cited alongside, same era.
2012
Cited alongside, same era.
M. R. Amer, D. Xie, M. Zhao, S. Todorovic, and S.-C. Zhu, “Cost-sensitive top-down/bottom-up inference for multiscale activity recognition,” in European Conference on Computer Vision (ECCV) . Springer, 2012, pp. 187–200
2012
Cited alongside, same era.
V. Ramanathan, B. Yao, and L. Fei-Fei, “Social role discovery in human events,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE, 2013, pp. 2475–2482
2013
Cited alongside, same era.
Y. Bo and H. Jiang, “Scale and rotation invariant approach to tracking human body part regions in videos,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Workshops , June 2013
2013
Cited alongside, same era.
S. Kwak, B. Han, and J. H. Han, “Multi-agent event detection: Localization and role assignment,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2013
2013
Cited alongside, same era.
A. Bialkowski, P. Lucey, P. Carr, S. Denman, I. Matthews, and S. Sridharan, “Recognising team activities from noisy data,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Workshops , 2013, pp. 984–990
2013
Cited alongside, same era.
Later among the works it cites.
J. Donahue, L. Anne Hendricks, S. Guadarrama, M. Rohrbach, S. Venugopalan, K. Saenko, and T. Darrell, “Long-term recurrent convolutional networks for visual recognition and description,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015, pp. 2625–2634
2015
Later among the works it cites.
T. Shu, D. Xie, B. Rothrock, S. Todorovic, and S.-C. Zhu, “Joint inference of groups, events and human roles in aerial videos,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015
2015
Later among the works it cites.
K. Soomro, S. Khokhar, and M. Shah, “Tracking when the camera looks away,” in IEEE International Conference on Computer Vision (ICCV) Workshops , 2015, pp. 25–33
2015
Later among the works it cites.
F. Turchini, L. Seidenari, and A. Del Bimbo, “Understanding sport activities from correspondences of clustered trajectories,” in IEEE International Conference on Computer Vision (ICCV) Workshops , December 2015
2015
Later among the works it cites.
X. Wei, L. Sha, P. Lucey, P. Carr, S. Sridharan, and I. Matthews, “Predicting ball ownership in basketball from a monocular view using only player trajectories,” in IEEE International Conference on Computer Vision (ICCV) Workshops , December 2015
2015
Later among the works it cites.
J. Y.-H. Ng, M. Hausknecht, S. Vijayanarasimhan, O. Vinyals, R. Monga, and G. Toderici, “Beyond short snippets: Deep networks for video classification,” IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015
2015
Later among the works it cites.
S. Venugopalan, H. Xu, J. Donahue, M. Rohrbach, R. Mooney, and K. Saenko, “Translating videos to natural language using deep recurrent neural networks,” in North American Chapter of the Association for Computational Linguistics , 2015
2015
Later among the works it cites.
A. Karpathy and L. Fei-Fei, “Deep visual-semantic alignments for generating image descriptions,” IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015
2015
Later among the works it cites.
S. Zheng, S. Jayasumana, B. Romera-Paredes, V. Vineet, Z. Su, D. Du, C. Huang, and P. H. S. Torr, “Conditional random fields as recurrent neural networks,” in IEEE International Conference on Computer Vision (ICCV) , 2015
2015
Later among the works it cites.
2015
Later among the works it cites.
Z. Deng, M. Zhai, L. Chen, Y. Liu, S. Muralidharan, M. Roshtkhari, and G. Mori, “Deep structured models for group activity recognition,” in British Machine Vision Conference (BMVC) , 2015
2015
Later among the works it cites.
2015
Later among the works it cites.
P. Over, G. Awad, M. Michel, J. Fiscus, W. Kraaij, A. F. Smeaton, G. Quéenot, and R. Ordelman, “Trecvid 2015 – an overview of the goals, tasks, data, evaluation mechanisms and metrics,” in Proceedings of TRECVID 2015 . NIST, USA, 2015
2015
Later among the works it cites.
D. Conigliaro, P. Rota, F. Setti, C. Bassetti, N. Conci, N. Sebe, and M. Cristani, “The s-hock dataset: Analyzing crowds at the stadium,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015
2015
Later among the works it cites.
H. Hajimirsadeghi, W. Yan, A. Vahdat, and G. Mori, “Visual recognition by counting instances: A multi-instance cardinality potential kernel,” IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015
2015
Later among the works it cites.
M. S. Ibrahim, S. Muralidharan, Z. Deng, A. Vahdat, and G. Mori, “A hierarchical deep temporal model for group activity recognition.” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016
2016
Closest in time.
V. Ramanathan, J. Huang, S. Abu-El-Haija, A. Gorban, K. Murphy, and L. Fei-Fei, “Detecting events and key actors in multi-person videos,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016
2016
Closest in time.