Fetching the paper…
Reading the bibliography…
This paper investigates the problem of zero-shot action recognition, in the setting where no training videos with seen actions are available.
The use of MMR, diversity-based reranking for reordering documents and producing summaries
Jaime Carbonell and Jade Goldstein · 1998
Earlier work this paper cites.
Recognizing human actions: a local SVM approach
Christian Schuldt, Ivan Laptev, and Barbara Caputo · 2004
Earlier work this paper cites.
Scene-based event detection for baseball videos
Cheng-Chang Lien, Chiu-Lung Chiang, and Chang-Hsing Lee · 2007
Earlier work this paper cites.
Recognizing human actions using multiple features
Jingen Liu, Saad Ali, and Mubarak Shah · 2008
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Introduction to modern information retrieval
Gobinda G Chowdhury · 2010
Earlier work this paper cites.
Object, Scene and Actions: Combining Multiple Features for Human Action Recognition
Nazli Ikizler-Cinbis and Stan Sclaroff · 2010
Earlier work this paper cites.
A short note on the kinetics-700-2020 human action dataset, 2020
Lucas Smaira, João Carreira, Eric Noland, Ellen Clancy, Amy Wu, and Andrew Zisserman · 2010
Earlier work this paper cites.
Recognizing human actions by attributes
Jingen Liu, Benjamin Kuipers, and Silvio Savarese · 2011
Earlier work this paper cites.
Semantic model vectors for complex video event recognition
Michele Merler, Bert Huang, Lexing Xie, Gang Hua, and Apostol Natsev · 2011
Earlier work this paper cites.
Ucf101: A dataset of 101 human actions classes from videos in the wild, 2012
Khurram Soomro, Amir Roshan Zamir, and Mubarak Shah · 2012
Earlier work this paper cites.
3D Convolutional Neural Networks for Human Action Recognition
Shuiwang Ji, Wei Xu, Ming Yang, and Kai Yu · 2013
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg Corrado, and Jeffrey Dean · 2013
Earlier work this paper cites.
Action Recognition with Improved Trajectories
Heng Wang and Cordelia Schmid · 2013
Earlier work this paper cites.
Composite concept discovery for zero-shot video event detection
Amirhossein Habibian, Thomas Mensink, and Cees GM Snoek · 2014
Earlier work this paper cites.
Attribute-Based Classification for Zero-Shot Visual Object Categorization
Christoph H. Lampert, Hannes Nickisch, and Stefan Harmeling · 2014
Earlier work this paper cites.
Object bank: An object-level image representation for high-level visual recognition
Li-Jia Li, Hao Su, Yongwhan Lim, and Li Fei-Fei · 2014
Earlier work this paper cites.
Conceptlets: Selective semantics for classifying video events
Masoud Mazloom, Efstratios Gavves, and Cees GM Snoek · 2014
Earlier work this paper cites.
Two-Stream Convolutional Networks for Action Recognition in Videos
Karen Simonyan and Andrew Zisserman · 2014
Earlier work this paper cites.
Semantic concept discovery for large-scale zero-shot event detection
Xiaojun Chang, Yi Yang, Alexander Hauptmann, Eric P Xing, and Yao-Liang Yu · 2015
Cited alongside, same era.
Camera Motion and Surrounding Scene Appearance as Context for Action Recognition
Fabian Caba Heilbron, Ali Thabet, Juan Carlos Niebles, and Bernard Ghanem · 2015
Cited alongside, same era.
Going deeper with convolutions
Christian Szegedy, Wei Liu, Yangqing Jia, Pierre Sermanet, Scott Reed, Dragomir Anguelov, Dumitru Erhan, Vincent Vanhoucke, and Andrew Rabinovich · 2015
Cited alongside, same era.
Semantic embedding space for zero-shot action recognition
Xun Xu, Timothy Hospedales, and Shaogang Gong · 2015
Cited alongside, same era.
Robust relative attributes for human action recognition
Zhong Zhang, Chunheng Wang, Baihua Xiao, Wen Zhou, and Shuang Liu · 2015
Cited alongside, same era.
Exploring synonyms as context in zero-shot action recognition
Places: A 10 Million Image Database for Scene Recognition
Bolei Zhou, Agata Lapedriza, Aditya Khosla, Aude Oliva, and Antonio Torralba · 2018
Later among the works it cites.
Towards universal representation for unseen action recognition
Yi Zhu, Yang Long, Yu Guan, Shawn Newsam, and Ling Shao · 2018
Later among the works it cites.
Video Action Transformer Network
Rohit Girdhar, Joao Joao Carreira, Carl Doersch, and Andrew Zisserman · 2019
Later among the works it cites.
Task-driven modular networks for zero-shot compositional learning
Senthil Purushwalkam, Maximilian Nickel, Abhinav Gupta, and Marc’Aurelio Ranzato · 2019
Later among the works it cites.
Sentence-bert: Sentence embeddings using siamese bert-networks
Nils Reimers and Iryna Gurevych · 2019
Later among the works it cites.
Towards a Fair Evaluation of Zero-Shot Action Recognition Using External Data
Alina Roitberg, Manuel Martinez, Monica Haurilet, and Rainer Stiefelhagen · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ioannis Alexiou, Tao Xiang, and Shaogang Gong · 2016
Cited alongside, same era.
Dynamic concept composition for zero-example event detection
Xiaojun Chang, Yi Yang, Guodong Long, Chengqi Zhang, and Alexander Hauptmann · 2016
Cited alongside, same era.
Learning Attributes Equals Multi-Source Domain Generalization
Chuang Gan, Tianbao Yang, and Boqing Gong · 2016
Cited alongside, same era.
Deep Residual Learning for Image Recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
Recognizing unseen actions in a domain-adapted embedding space
Yikang Li, Sheng-hung Hu, and Baoxin Li · 2016
Cited alongside, same era.
Harnessing Object and Scene Semantics for Large-Scale Video Understanding
Zuxuan Wu, Yanwei Fu, Yu-Gang Jiang, and Leonid Sigal · 2016
Cited alongside, same era.
Quo vadis, action recognition? a new model and the kinetics dataset
Joao Carreira and Andrew Zisserman · 2017
Cited alongside, same era.
Later among the works it cites.
Combining multiple deep cues for action recognition
Ruiqi Wang and Xinxiao Wu · 2019
Later among the works it cites.
HACS: Human Action Clips and Segments Dataset for Recognition and Temporal Localization
Hang Zhao, Antonio Torralba, Lorenzo Torresani, and Zhicheng Yan · 2019
Later among the works it cites.
Rethinking zero-shot video classification: End-to-end training for realistic applications
Biagio Brattoli, Joseph Tighe, Fedor Zhdanov, Pietro Perona, and Krzysztof Chalupka · 2020
Later among the works it cites.
Using sentences as semantic representations in large scale zero-shot learning
Yannick Le Cacheux, Hervé Le Borgne, and Michel Crucianu · 2020
Later among the works it cites.
Searching for actions on the hyperbole
Teng Long, Pascal Mettes, Heng Tao Shen, and Cees G M Snoek · 2020
Later among the works it cites.
Shuffled ImageNet Banks for Video Event Detection and Search
Pascal Mettes, Dennis C. Koelma, and Cees G. M. Snoek · 2020
Later among the works it cites.
Zero-shot learning for action recognition using synthesized features
Ashish Mishra, Anubha Pandey, and Hema A Murthy · 2020
Later among the works it cites.
Moments in Time Dataset: One Million Videos for Event Understanding
Mathew Monfort, Carl Vondrick, Aude Oliva, Alex Andonian, Bolei Zhou, Kandan Ramakrishnan, Sarah Adel Bargal, Tom Yan, Lisa Brown, Quanfu Fan, and Dan Gutfreund · 2020
Later among the works it cites.
MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained Transformers
Wenhui Wang, Furu Wei, Li Dong, Hangbo Bao, Nan Yang, and Ming Zhou · 2020
Later among the works it cites.
Open world compositional zero-shot learning
Massimiliano Mancini, Muhammad Ferjad Naeem, Yongqin Xian, and Zeynep Akata · 2021
Closest in time.
Object Priors for Classifying and Localizing Unseen Actions
Pascal Mettes, William Thong, and Cees G. M. Snoek · 2021
Closest in time.
Learning graph embeddings for compositional zero-shot learning
Muhammad Ferjad Naeem, Yongqin Xian, Federico Tombari, and Zeynep Akata · 2021
Closest in time.