Fetching the paper…
Reading the bibliography…
The number of categories for action recognition is growing rapidly and it has become increasingly hard to label sufficient training data for learning conventional models for all categories.
Scholkopf B, Smola AJ (2002) Learning with Kernels: Support Vector Machines, Regularization, Optimization, and Beyond. MIT Press. (2002),
2002
Earlier work this paper cites.
Schuldt C, Laptev I, Caputo B (2004) Recognizing human actions: A local svm approach. In: ICPR
2004
Earlier work this paper cites.
Zhou D, Bousquet O, Weston J (2004) Learning with local and global consistency. In: NIPS
2004
Earlier work this paper cites.
Blank M, Gorelick L, Shechtman E, Irani M, Basri R (2005) Actions as space-time shapes. In: ICCV
2005
Earlier work this paper cites.
Laptev I (2005) On space-time interest points. International Journal of Computer Vision
2005
Earlier work this paper cites.
Belkin M, Niyogi P, Sindhwani V (2006) Manifold regularization: A geometric framework for learning from labeled and unlabeled examples. The Journal of Machine Learning Research
2006
Earlier work this paper cites.
Scovanner P, Ali S, Shah M (2007) A 3-dimensional sift descriptor and its application to action recognition. In: ACM Multimedia
2007
Earlier work this paper cites.
Klaser A, Marszałek M, Schmid C (2008) A spatio-temporal descriptor based on 3d-gradients. In: BMVC
2008
Earlier work this paper cites.
Larochelle H, Erhan D, Bengio Y (2008) Zero-data learning of new tasks. In: AAAI
2008
Earlier work this paper cites.
Maaten LVD, Hinton G (2008) Visualizing data using t-SNE. Journal of Machine Learning Research
2008
Earlier work this paper cites.
Mitchell J, Lapata M (2008) Vector-based Models of Semantic Composition. Computational Linguistics
2008
Earlier work this paper cites.
Deng J, Dong W, Socher R, Li L, Li K, Li F (2009) Imagenet: A large-scale hierarchical image database. In: CVPR
2009
Earlier work this paper cites.
Lampert CH, Nickisch H, Harmeling S (2009) Learning to detect unseen object classes by between-class attribute transfer. In: CVPR
2009
Earlier work this paper cites.
Marszalek M, Laptev I, Schmid C (2009) Actions in context. In: CVPR
2009
Earlier work this paper cites.
Palatucci M, Hinton G, Pomerleau D, Mitchell TM (2009) Zero-shot learning with semantic output codes. In: NIPS
2009
Earlier work this paper cites.
Yeffet L, Wolf L (2009) Local trinary patterns for human action recognition. In: ICCV
2009
Earlier work this paper cites.
Niebles CWFFL Juan Carlos Chen (2010) Modeling temporal structure of decomposable motion segments for activity classification. In: ECCV
2010
Earlier work this paper cites.
Pan SJ, Yang Q (2010) A survey on transfer learning. IEEE Transactions on Knowledge and Data Engineering
2010
Earlier work this paper cites.
Perronnin F, Sánchez J, Mensink T (2010) Improving the fisher kernel for large-scale image classification. In: ECCV
2010
Earlier work this paper cites.
Poppe R (2010) A survey on vision-based human action recognition. Image and vision computing
2010
Earlier work this paper cites.
Rohrbach M, Stark M, Szarvas G, Gurevych I, Schiele B (2010) What helps where - and why? Semantic relatedness for knowledge transfer. In: CVPR
2010
Cited alongside, same era.
Aggarwal J, Ryoo M (2011) Human activity analysis: A review. ACM Computer Survey
2011
Cited alongside, same era.
Jiang YG, Ye G, Chang SF, Ellis DPW, Loui AC (2011) Consumer video understanding: a benchmark database and an evaluation of human and machine performance. In: ICMR
2011
Cited alongside, same era.
Kuehne H, Jhuang H, Garrote E, Poggio T, Serre T (2011) Hmdb: A large video database for human motion recognition. In: ICCV
2011
Cited alongside, same era.
Liu J, Kuipers B, Savarese S (2011) Recognizing human actions by attributes. In: CVPR
2011
Cited alongside, same era.
Lazaridou A, Bruni E, Baroni M (2014) Is this a wampimuk? cross-modal mapping between distributional semantics and the visual world. In: ACL
2014
Later among the works it cites.
Mensink T, Gavves E, Snoek CG (2014) Costa: Co-occurrence statistics for zero-shot classification. In: CVPR
2014
Later among the works it cites.
Milajevs D, Kartsaklis D, Sadrzadeh M, Purver M (2014) Evaluating neural word representations in tensor-based compositional settings. In: EMNLP
2014
Later among the works it cites.
Norouzi M, Mikolov T, Bengio S, Singer Y, Shlens J, Frome A, Corrado GS, Dean J (2014) Zero-shot learning by convex combination of semantic embeddings. In: ICLR
2014
Later among the works it cites.
Wu S, Bondugula S, Luisier F, Zhuang X, Natarajan P (2014) Zero-Shot Event Detection Using Multi-modal Fusion of Weakly Supervised Concepts. In: CVPR
2014
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Rohrbach M, Stark M, Schiele B (2011) Evaluating knowledge transfer and zero-shot learning in a large-scale setting. In: CVPR
2011
Cited alongside, same era.
Fu Y, Hospedales TM, Xiang T, Gong S (2012) Attribute learning for understanding unstructured social activity. In: ECCV
2012
Cited alongside, same era.
Soomro K, Zamir AR, Shah M (2012) Ucf101: A dataset of 101 human actions classes from videos in the wild. arXiv preprint arXiv:12120402
2012
Cited alongside, same era.
Frome A, Corrado GS, Shlens J (2013) Devise: A deep visual-semantic embedding model. In: NIPS
2013
Cited alongside, same era.
Jiang YG, Liu J, Roshan Zamir A, Laptev I, Piccardi M, Shah M, Sukthankar R (2013) THUMOS challenge: Action recognition with a large number of classes
2013
Cited alongside, same era.
Mikolov T, Sutskever I, Chen K (2013) Distributed representations of words and phrases and their compositionality. In: NIPS
2013
Cited alongside, same era.
Over P, Fiscus J, Sanders G, Joy D, Michel M, Smeaton-Alan AF, Quénot-Georges G (2014) Trecvid 2013–an overview of the goals, tasks, data, evaluation mechanisms, and metrics
2013
Cited alongside, same era.
Later among the works it cites.
Zheng J, Jiang Z (2014) Submodular Attribute Selection for Action Recognition in Video. In: NIPS
2014
Later among the works it cites.
Akata Z, Reed S, Walter D, Lee H, Schiele B (2015) Evaluation of output embeddings for fine-grained image classification. In: CVPR
2015
Closest in time.
Dinu G, Lazaridou A, Baroni M (2015) Improving zero-shot learning by mitigating the hubness problem. In: ICLR, Workshop Track
2015
Closest in time.
Gan C, Lin M, Yang Y, Zhuang Y, GHauptmann A (2015) Exploring semantic inter-class relationships (sir) for zero-shot action recognition. In: AAAI
2015
Closest in time.
Jain M, Snoek CGM (2015) What do 15 , 000 object categories tell us about classifying and localizing actions ? In: CVPR
2015
Closest in time.
Jain M, van Gemert JC, Mensink T, Snoek CGM (2015) Objects2action: Classifying and localizing actions without any video example. In: ICCV
2015
Closest in time.
Jiang Y, Wu Z, Wang J, Xue X, Chang S (2015) Exploiting feature and class relationships in video categorization with regularized deep neural networks. arXiv preprint arXiv:150207209
2015
Closest in time.
Kodirov E, Xiang T, Fu Z, Gong S (2015) Unsupervised Domain Adaptation for Zero-Shot Learning. In: ICCV
2015
Closest in time.
Romera-Paredes B, Torr PHS (2015) An embarrassingly simple approach to zero-shot learning. In: ICML
2015
Closest in time.
Shao L, Zhu F, Li X (2015) Transfer learning for visual categorization: a survey. IEEE transactions on neural networks and learning systems
2015
Closest in time.
Wang H, Oneata D, Verbeek J, Schmid C (2015) A robust and efficient video representation for action recognition. International Journal of Computer Vision
2015
Closest in time.
Xu X, Hospedales T, Gong S (2015) Semantic embedding space for zero shot action recognition. In: ICIP
2015
Closest in time.
Yang Y, Hospedales T (2015) A unified perspective on multi-domain and multi-task learning. In: ICLR
2015
Closest in time.
Rohrbach M, Rohrbach A, Regneri M, Amin S, Andriluka M, Pinkal M, Schiele B (2016) Recognizing fine-grained and composite activities using hand-centric features and script data. IJCV
2016
Closest in time.