Fetching the paper…
Reading the bibliography…
Scene, as the crucial unit of storytelling in movies, contains complex activities of actors and their interactions in a physical environment.
Exploring video structure beyond the shots
Yong Rui, Thomas S Huang, and Sharad Mehrotra · 1998
Earlier work this paper cites.
Fitting the mel scale
Srinivasan Umesh, Leon Cohen, and D Nelson · 1999
Earlier work this paper cites.
Scene detection in hollywood movies and tv shows
Zeeshan Rasheed and Mubarak Shah · 2003
Earlier work this paper cites.
Framewise phoneme classification with bidirectional lstm and other neural network architectures
Alex Graves and Jürgen Schmidhuber · 2005
Earlier work this paper cites.
Detection and representation of scenes in videos
Zeeshan Rasheed and Mubarak Shah · 2005
Earlier work this paper cites.
Scene detection in videos using shot clustering and sequence alignment
Vasileios T Chasanis, Aristidis C Likas, and Nikolaos P Galatsanos · 2008
Earlier work this paper cites.
A novel role-based movie scene segmentation method
Chao Liang, Yifan Zhang, Jian Cheng, Changsheng Xu, and Hanqing Lu · 2009
Earlier work this paper cites.
Video scene segmentation using a novel boundary evaluation criterion and dynamic programming
Bo Han and Weiguo Wu · 2011
Earlier work this paper cites.
Temporal video segmentation to scenes using high-level audiovisual features
Panagiotis Sidiropoulos, Vasileios Mezaris, Ioannis Kompatsiaris, Hugo Meinedo, Miguel Bugalho, and Isabel Trancoso · 2011
Earlier work this paper cites.
Finding actors and actions in movies
Piotr Bojanowski, Francis Bach, Ivan Laptev, Jean Ponce, Cordelia Schmid, and Josef Sivic · 2013
Earlier work this paper cites.
Linking people in videos with “their” names using coreference resolution
Vignesh Ramanathan, Armand Joulin, Percy Liang, and Li Fei-Fei · 2014
Earlier work this paper cites.
Storygraphs: visualizing character interactions as a timeline
Makarand Tapaswi, Martin Bauml, and Rainer Stiefelhagen · 2014
Cited alongside, same era.
A deep siamese network for scene detection in broadcast videos
Lorenzo Baraldi, Costantino Grana, and Rita Cucchiara · 2015
Cited alongside, same era.
Activitynet: A large-scale video benchmark for human activity understanding
Bernard Ghanem Fabian Caba Heilbron, Victor Escorcia and Juan Carlos Niebles · 2015
Cited alongside, same era.
Saurabh Gupta and Jitendra Malik · 2015
Cited alongside, same era.
Faster r-cnn: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2015
Cited alongside, same era.
Beyond frontal faces: Improving person recognition using multiple cues
Optimal sequential grouping for robust video scene detection using multiple modalities
Daniel Rotman, Dror Porat, and Gal Ashour · 2017
Later among the works it cites.
Pyscenedetect: Intelligent scene cut detection and video splitting tool
Brandon Castellano · 2018
Later among the works it cites.
Ava: A video dataset of spatio-temporally localized atomic visual actions
Chunhui Gu, Chen Sun, David A Ross, Carl Vondrick, Caroline Pantofaru, Yeqing Li, Sudheendra Vijayanarasimhan, George Toderici, Susanna Ricco, Rahul Sukthankar, et al · 2018
Later among the works it cites.
Unifying identification and context learning for person recognition
Qingqiu Huang, Yu Xiong, and Dahua Lin · 2018
Later among the works it cites.
Using deep features for video scene detection and annotation
Stanislav Protasov, Adil Mehmood Khan, Konstantin Sozykin, and Muhammad Ahmad · 2018
Later among the works it cites.
Moviegraphs: Towards understanding human-centric situations from videos
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ning Zhang, Manohar Paluri, Yaniv Taigman, Rob Fergus, and Lubomir Bourdev · 2015
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
Temporal segment networks: Towards good practices for deep action recognition
Limin Wang, Yuanjun Xiong, Zhe Wang, Yu Qiao, Dahua Lin, Xiaoou Tang, and Luc Val Gool · 2016
Cited alongside, same era.
Temporal segment networks: Towards good practices for deep action recognition
Limin Wang, Yuanjun Xiong, Zhe Wang, Yu Qiao, Dahua Lin, Xiaoou Tang, and Luc Van Gool · 2016
Cited alongside, same era.
Situation recognition: Visual semantic role labeling for image understanding
Mark Yatskar, Luke Zettlemoyer, and Ali Farhadi · 2016
Cited alongside, same era.
Paul Vicol, Makarand Tapaswi, Lluis Castrejon, and Sanja Fidler · 2018
Later among the works it cites.
Places: A 10 million image database for scene recognition
Bolei Zhou, Agata Lapedriza, Aditya Khosla, Aude Oliva, and Antonio Torralba · 2018
Later among the works it cites.
Naver at activitynet challenge 2019–task b active speaker detection (ava)
Joon Son Chung · 2019
Later among the works it cites.
Moments in time dataset: one million videos for event understanding
Mathew Monfort, Alex Andonian, Bolei Zhou, Kandan Ramakrishnan, Sarah Adel Bargal, Yan Yan, Lisa Brown, Quanfu Fan, Dan Gutfreund, Carl Vondrick, et al · 2019
Later among the works it cites.
Ava-activespeaker: An audio-visual dataset for active speaker detection
Joseph Roth, Sourish Chaudhuri, Ondrej Klejch, Radhika Marvin, Andrew Gallagher, Liat Kaver, Sharadh Ramaswamy, Arkadiusz Stopczynski, Cordelia Schmid, Zhonghua Xi, et al · 2019
Later among the works it cites.