Fetching the paper…
Reading the bibliography…
Activity recognition has shown impressive progress in recent years.
The representation and matching of pictorial structures
Martin Fischler and Robert Elschlager · 1973
Earlier work this paper cites.
The Handbook of Artificial Intelligence, Volume 1
Avron Barr and Edward Feigenbaum · 1981
Earlier work this paper cites.
Term-weighting approaches in automatic text retrieval
Gerard Salton and Christopher Buckley · 1988
Earlier work this paper cites.
Recognition of human body motion using phase space constraints
Lee Campbell and Aaron Bobick · 1995
Earlier work this paper cites.
Stacked generalization: when does it work?
Kai Ming Ting and Ian H Witten · 1997
Earlier work this paper cites.
WordNet: An Electronical Lexical Database
Christiane Fellbaum · 1998
Earlier work this paper cites.
Open mind common sense: Knowledge acquisition from the general public
Push Singh, Thomas Lin, Erik Mueller, Grace Lim, Travell Perkins, and Wan Zhu · 2002
Earlier work this paper cites.
Recognizing human actions: a local SVM approach
Christian Schuldt, Ivan Laptev, and Barbara Caputo · 2004
Earlier work this paper cites.
Learning with Local and Global Consistency
Dengyong Zhou, Olivier Bousquet, Thomas Navin Lal, Jason Weston, and Bernhard Schölkopf · 2004
Earlier work this paper cites.
Pictorial structures for object recognition
Pedro Felzenszwalb and Daniel Huttenlocher · 2005
Earlier work this paper cites.
On space-time interest points
Ivan Laptev · 2005
Earlier work this paper cites.
Human detection using oriented histograms of flow and appearance
Navneet Dalal, Bill Triggs, and Cordelia Schmid · 2006
Earlier work this paper cites.
Advene: an open-source framework for integrating and visualising audiovisual metadata
Olivier Aubert and Yannick Prié · 2007
Earlier work this paper cites.
PETS , 2007
James Ferryman, editor · 2007
Earlier work this paper cites.
Retrieving actions in movies
Ivan Laptev and Patrick Pérez · 2007
Earlier work this paper cites.
What, where and who? classifying events by scene and object recognition
Li-Jia Li and Fei-Fei Li · 2007
Earlier work this paper cites.
Progressive search space reduction for human pose estimation
Vittorio Ferrari, Manuel Marin, and Andrew Zisserman · 2008
Earlier work this paper cites.
Learning realistic human actions from movies
Ivan Laptev, Marcin Marszalek, Cordelia Schmid, and Benjamin Rozenfeld · 2008
Earlier work this paper cites.
View and scale invariant action recognition using multiview shape-flow models
Pradeep Natarajan and Ramakant Nevatia · 2008
Earlier work this paper cites.
Automated flower classification over a large number of classes
Maria-Elena Nilsback and Andrew Zisserman · 2008
Earlier work this paper cites.
Action MACH a spatio-temporal maximum average correlation height filter for action recognition
Mikel Rodriguez, Javed Ahmed, and Mubarak Shah · 2008
Earlier work this paper cites.
Pictorial structures revisited: People detection and articulated pose estimation
Mykhaylo Andriluka, Stefan Roth, and Bernt Schiele · 2009
Earlier work this paper cites.
Understanding videos, constructing plots learning a visually grounded storyline model from annotated videos
Abhinav Gupta, Praveen Srinivasan, Jianbo Shi, and Larry Davis · 2009
Earlier work this paper cites.
Guide to the cmu multimodal activity database
Fernando De la Torre, Jessica Hodgins, Javier Montano, Sergio Valcarcel, Ricard Forcada, and Justin Macey · 2009
Earlier work this paper cites.
Recognizing realistic actions from videos ’in the wild’
Jingen Liu, Jiebo Luo, and Mubarak Shah · 2009
Earlier work this paper cites.
Actions in context
Marcin Marszalek, Ivan Laptev, and Cordelia Schmid · 2009
Earlier work this paper cites.
Activity recognition using the velocity histories of tracked keypoints
Ross Messing, Chris Pal, and Henry Kautz · 2009
Earlier work this paper cites.
Spatio-temporal relationship match: Video structure comparison for recognition of complex human activities
Michael Ryoo and Jake Aggarwal · 2009
Earlier work this paper cites.
Feature-weighted linear stacking
Joseph Sill, Gábor Takács, Lester Mackey, and David Lin · 2009
Earlier work this paper cites.
The TUM Kitchen Data Set of Everyday Manipulation Activities for Motion Tracking and Action Recognition
Moritz Tenorth, Jan Bandouch, and Michael Beetz · 2009
Earlier work this paper cites.
Local trinary patterns for human action recognition
Lahav Yeffet and Lior Wolf · 2009
Earlier work this paper cites.
Discriminative subvolume search for efficient action detection
Junsong Yuan, Zicheng Liu, and Ying Wu · 2009
Earlier work this paper cites.
An analysis of sensor-oriented vs. model-based activity recognition
Andreas Zinnen, Ulf Blanke, and Bernt Schiele · 2009
Earlier work this paper cites.
Attribute-centric recognition for cross-category generalization
Ali Farhadi, Ian Endres, and Derek Hoiem · 2010
Earlier work this paper cites.
Object detection with discriminatively trained part-based models
Pedro Felzenszwalb, Ross Girshick, David McAllester, and Deva Ramanan · 2010
Earlier work this paper cites.
Using body-anchored priors for identifying actions in single images
Leonid Karlinsky, Michael Dinerstein, and Shimon Ullman · 2010
Earlier work this paper cites.
Modeling temporal structure of decomposable motion segments for activity classification
Juan Niebles, Chih-Wei Chen, and Li Fei-Fei · 2010
Cited alongside, same era.
High five: Recognising human interactions in TV shows
Alonso Patron-Perez, Marcin Marszalek, Andrew Zisserman, and Ian D. Reid · 2010
Cited alongside, same era.
Learning script knowledge with web experiments
Michaela Regneri, Alexander Koller, and Manfred Pinkal · 2010
Cited alongside, same era.
Collecting complex activity data sets in highly rich networked sensor environments
Daniel Roggen, Alberto Calatroni, Mirco Rossi, Thomas Holleczek, Kilian Forster, Gerhard Troster, Paul Lukowicz, David Bannach, Gerald Pirkl, Alois Ferscha, Jakob Doppler, Clemens Holzmann, Marc Kurz, Gerald Holl, Ricardo Chavarriaga, Hesam Sagha, Hamidreza Bayati, Marco Creatura, and Jose del R. Millan · 2010
Cited alongside, same era.
What helps Where - and Why? Semantic Relatedness for Knowledge Transfer
Marcus Rohrbach, Michael Stark, György Szarvas, Iryna Gurevych, and Bernt Schiele · 2010
Cited alongside, same era.
Trecvid 2012 – an overview of the goals, tasks, data, evaluation mechanisms and metrics
Paul Over, George Awad, Martial Michel, Jonathan Fiscus, Greg Sanders, B Shaw, Alan F. Smeaton, and Georges Quéenot · 2012
Later among the works it cites.
A combined pose, object, and feature model for action understanding
Benjamin Packer, Kate Saenko, and Daphne Koller · 2012
Later among the works it cites.
Detecting activities of daily living in first-person camera views
Hamed Pirsiavash and Deva Ramanan · 2012
Later among the works it cites.
Ucf101: A dataset of 101 human actions classes from videos in the wild
Khurram Soomro, Amir Roshan Zamir, and Mubarak Shah · 2012
Later among the works it cites.
Learning latent temporal structure for complex event detection
Kevin Tang, Li Fei-Fei, and Daphne Koller · 2012
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cascaded models for articulated pose estimation, 2010
Benjamin Sapp, Alexander Toshev, and Ben Taskar · 2010
Cited alongside, same era.
Connecting modalities: Semi-supervised segmentation and annotation of images using unaligned text corpora
Richard Socher and Li Fei-Fei · 2010
Cited alongside, same era.
Convolutional learning of spatio-temporal features
Graham W Taylor, Rob Fergus, Yann LeCun, and Christoph Bregler · 2010
Cited alongside, same era.
Efficient additive kernels via explicit feature maps
Andrea Vedaldi and Andrew Zisserman · 2010
Cited alongside, same era.
Caltech-ucsd birds 200
Peter Welinder, Steve Branson, Takeshi Mita, Catherine Wah, Florian Schroff, Serge Belongie, and Pietro Perona · 2010
Cited alongside, same era.
Discriminative appearance models for pictorial structures
Mykhaylo Andriluka, Stefan Roth, and Bernt Schiele · 2011
Cited alongside, same era.
Sequential deep learning for human action recognition
Moez Baccouche, Franck Mamalet, Christian Wolf, Christophe Garcia, and Atilla Baskurt · 2011
Cited alongside, same era.
Towards a watson that sees: Language-guided action recognition for robots
Ching Lik Teo, Yezhou Yang, H Daume, C Fermuller, and Yiannis Aloimonos · 2012
Later among the works it cites.
Recognizing human-object interactions in still images by modeling the mutual context of objects and human poses
Bangpeng Yao and Fei-Fei Li · 2012
Later among the works it cites.
Multi-view Pictorial Structures for 3D Human Pose Estimation
Sikandar Amin, Mykhaylo Andriluka, Marcus Rohrbach, and Bernt Schiele · 2013
Later among the works it cites.
A survey of video datasets for human action and activity recognition
Jose Chaquet, Enrique Carmona, and Antonio Fernández-Caballero · 2013
Later among the works it cites.
Thousand frames in just a few words: Lingual description of videos through latent topics and sparse object stitching
Pradipto Das, Chenliang Xu, Richard Doell, and Jason Corso · 2013
Later among the works it cites.
Write a classifier: Zero-shot learning using purely textual descriptions
Mohamed Elhoseiny, Babak Saleh, and Ahmed Elgammal · 2013
Later among the works it cites.
Devise: A deep visual-semantic embedding model
Andrea Frome, Greg Corrado, Jon Shlens, Samy Bengio, Jeffrey Dean, Marc’Aurelio Ranzato, and Tomas Mikolov · 2013
Later among the works it cites.
Learning multi-modal latent attributes
Yanwei Fu, Timothy Hospedales, Tao Xiang, and Shaogang Gong · 2013
Later among the works it cites.
Articulated pose estimation using discriminative armlet classifiers
Georgia Gkioxari, Pablo Arbelaez, Lubomir Bourdev, and Jitendra Malik · 2013
Later among the works it cites.
Youtube2text: Recognizing and describing arbitrary activities using semantic hierarchies and zero-shoot recognition
Sergio Guadarrama, Niveda Krishnamoorthy, Girish Malkarnenkar, Subhashini Venugopalan, Raymond Mooney, Trevor Darrell, and Kate Saenko · 2013
Later among the works it cites.
Towards understanding action recognition
Hueihan Jhuang, Jurgen Gall, Silvia Zuffi, Cordelia Schmid, and Michael Black · 2013
Later among the works it cites.
3D convolutional neural networks for human action recognition
Shuiwang Ji, Wei Xu, Ming Yang, and Kai Yu · 2013
Later among the works it cites.
Attribute-based classification for zero-shot learning of object categories
Christoph Lampert, Hannes Nickisch, and Stefan Harmeling · 2013
Later among the works it cites.
Video event understanding using natural language descriptions
Vignesh Ramanathan, Percy Liang, and Li Fei-Fei · 2013
Later among the works it cites.
Poselet key-framing: A model for human activity recognition
Michalis Raptis and Leonid Sigal · 2013
Later among the works it cites.
Grounding Action Descriptions in Videos
Michaela Regneri, Marcus Rohrbach, Dominikus Wetzel, Stefan Thater, Bernt Schiele, and Manfred Pinkal · 2013
Later among the works it cites.
Zero-shot learning through cross-modal transfer
Richard Socher, Milind Ganjoo, Christopher D. Manning, and Andrew Ng · 2013
Later among the works it cites.
Combining embedded accelerometers with computer vision for recognizing food preparation activities
Sebastian Stein and Stephen McKenna · 2013
Later among the works it cites.
Combining the right features for complex event recognition
Kevin Tang, Bangpeng Yao, Li Fei-Fei, and Daphne Koller · 2013
Later among the works it cites.
Action recognition with improved trajectories
Heng Wang and Cordelia Schmid · 2013
Later among the works it cites.
Articulated human detection with flexible mixtures of parts
Yi Yang and Deva Ramanan · 2013
Later among the works it cites.
Weakly supervised action labeling in videos under ordering constraints
Piotr Bojanowski, Rémi Lajugie, Francis Bach, Ivan Laptev, Jean Ponce, Cordelia Schmid, and Josef Sivic · 2014
Later among the works it cites.
Mixing Body-Part Sequences for Human Pose Estimation
Anoop Cherian, Julien Mairal, Karteek Alahari, and Cordelia Schmid · 2014
Later among the works it cites.
Efficient feature extraction, encoding and classification for action recognition
Vadim Kantorov and Ivan Laptev · 2014
Later among the works it cites.
Large-scale video classification with convolutional neural networks
Andrej Karpathy, George Toderici, Sanketh Shetty, Thomas Leung, Rahul Sukthankar, and Li Fei-Fei · 2014
Later among the works it cites.
Parsing videos of actions with segmental grammars
Hamed Pirsiavash and Deva Ramanan · 2014
Later among the works it cites.
Coherent multi-sentence video description with variable level of detail
Anna Rohrbach, Marcus Rohrbach, Wei Qiu, Annemarie Friedrich, Manfred Pinkal, and Bernt Schiele · 2014
Later among the works it cites.
Coherent multi-sentence video description with variable level of detail
Anna Senina, Marcus Rohrbach, Wei Qiu, Annemarie Friedrich, Sikandar Amin, Mykhaylo Andriluka, Manfred Pinkal, and Bernt Schiele · 2014
Later among the works it cites.
Discovering object classes from activities
Abhilash Srikantha and Juergen Gall · 2014
Later among the works it cites.
A dataset for movie description
Anna Rohrbach, Marcus Rohrbach, Niket Tandon, and Bernt Schiele · 2015
Closest in time.