Fetching the paper…
Reading the bibliography…
Understanding movies and their structural patterns is a crucial task in decoding the craft of video editing.
Arijon, D.: Grammar of the film language. Focal Press London (1976)
1976
Earlier work this paper cites.
Bordwell, D., Thompson, K., Smith, J.: Film art: An introduction, vol. 7. McGraw-Hill New York (1993)
1993
Earlier work this paper cites.
Murch, W.: In the Blink of an Eye, vol. 995. Silman-James Press Los Angeles (2001)
2001
Earlier work this paper cites.
Katz, E., Klein, F.: The film encyclopedia. Collins (2005)
2005
Earlier work this paper cites.
Everingham, M., Sivic, J., Zisserman, A.: Hello! my name is… buffy”–automatic naming of characters in tv video. In: BMVC. vol. 2, p. 6 (2006)
2006
Earlier work this paper cites.
Kozlovic, A.K.: Anatomy of film. Kinema: A Journal for Film and Audiovisual Media (2007)
2007
Earlier work this paper cites.
Laptev, I., Marszalek, M., Schmid, C., Rozenfeld, B.: Learning realistic human actions from movies. In: 2008 IEEE Conference on Computer Vision and Pattern Recognition. pp. 1–8. IEEE (2008)
2008
Earlier work this paper cites.
Smith, T.J., Henderson, J.M.: Edit blindness: The relationship between attention and global change blindness in dynamic scenes. Journal of Eye Movement Research 2
2008
Earlier work this paper cites.
Duchenne, O., Laptev, I., Sivic, J., Bach, F., Ponce, J.: Automatic annotation of human actions in video. In: 2009 IEEE 12th International Conference on Computer Vision. pp. 1491–1498. IEEE (2009)
2009
Earlier work this paper cites.
Sivic, J., Everingham, M., Zisserman, A.: “who are you?”-learning person specific classifiers from video. In: 2009 IEEE Conference on Computer Vision and Pattern Recognition. pp. 1145–1152. IEEE (2009)
2009
Earlier work this paper cites.
Thompson, R., Bowen, C.J.: Grammar of the Edit, vol. 13. Taylor & Francis (2009)
2009
Earlier work this paper cites.
Tsivian, Y.: Cinemetrics, part of the humanities’ cyberinfrastructure (2009)
2009
Earlier work this paper cites.
Wang, H.L., Cheong, L.F.: Taxonomy of directing semantics for film shot classification. IEEE transactions on circuits and systems for video technology 19
2009
Earlier work this paper cites.
Irie, G., Satou, T., Kojima, A., Yamasaki, T., Aizawa, K.: Automatic trailer generation. In: Proceedings of the 18th ACM international conference on Multimedia. pp. 839–842 (2010)
2010
Earlier work this paper cites.
Kuehne, H., Jhuang, H., Garrote, E., Poggio, T., Serre, T.: Hmdb: a large video database for human motion recognition. In: 2011 International conference on computer vision. pp. 2556–2563. IEEE (2011)
2011
Earlier work this paper cites.
Smith, T.J.: The attentional theory of cinematic continuity. Projections 6
2012
Earlier work this paper cites.
Smith, T.J., Levin, D., Cutting, J.E.: A window on reality: Perceiving edited moving images. Current Directions in Psychological Science 21
2012
Earlier work this paper cites.
Bojanowski, P., Bach, F., Laptev, I., Ponce, J., Schmid, C., Sivic, J.: Finding actors and actions in movies. In: Proceedings of the IEEE international conference on computer vision. pp. 2280–2287 (2013)
2013
Earlier work this paper cites.
Canini, L., Benini, S., Leonardi, R.: Classifying cinematographic shot types. Multimedia tools and applications 62
2013
Earlier work this paper cites.
Burch, N.: Theory of film practice. Princeton University Press (2014)
2014
Earlier work this paper cites.
Hoai, M., Zisserman, A.: Thread-safe: Towards recognizing human actions across shot boundaries. In: Asian Conference on Computer Vision. pp. 222–237. Springer (2014)
2014
Earlier work this paper cites.
Galvane, Q., Ronfard, R., Lino, C., Christie, M.: Continuity editing for 3d animation. In: Proceedings of the AAAI Conference on Artificial Intelligence. vol. 29 (2015)
2015
Cited alongside, same era.
Xu, H., Zhen, Y., Zha, H.: Trailer generation via a point process-based visual attractiveness model. In: Twenty-Fourth International Joint Conference on Artificial Intelligence (2015)
2015
Cited alongside, same era.
Benini, S., Svanera, M., Adami, N., Leonardi, R., Kovács, A.B.: Shot scale distribution in art films. Multimedia Tools and Applications 75
2016
Cited alongside, same era.
Cutting, J.E.: The evolution of pace in popular movies. Cognitive research: principles and implications 1
2016
Cited alongside, same era.
He, K., Zhang, X., Ren, S., Sun, J.: Deep residual learning for image recognition. In: Proceedings of the IEEE conference on computer vision and pattern recognition. pp. 770–778 (2016)
Tran, D., Wang, H., Torresani, L., Ray, J., LeCun, Y., Paluri, M.: A closer look at spatiotemporal convolutions for action recognition. In: Proceedings of the IEEE conference on Computer Vision and Pattern Recognition. pp. 6450–6459 (2018)
2018
Later among the works it cites.
Vicol, P., Tapaswi, M., Castrejon, L., Fidler, S.: Moviegraphs: Towards understanding human-centric situations from videos. In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2018)
2018
Later among the works it cites.
Wu, H.Y., Palù, F., Ranon, R., Christie, M.: Thinking like a director: Film editing patterns for virtual cinematographic storytelling. ACM Transactions on Multimedia Computing, Communications, and Applications (TOMM) 14
2018
Later among the works it cites.
Bost, X., Gueye, S., Labatut, V., Larson, M., Linarès, G., Malinas, D., Roth, R.: Remembering winter was coming. Multimedia Tools and Applications 78
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
Tapaswi, M., Zhu, Y., Stiefelhagen, R., Torralba, A., Urtasun, R., Fidler, S.: MovieQA: Understanding Stories in Movies through Question-Answering. In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2016)
2016
Cited alongside, same era.
Wu, H.Y., Christie, M.: Analysing cinematography with embedded constrained patterns. In: WICED-Eurographics Workshop on Intelligent Cinematography and Editing (2016)
2016
Cited alongside, same era.
Bredin, H.: pyannote.metrics: a toolkit for reproducible evaluation, diagnostic, and error analysis of speaker diarization systems. In: Interspeech 2017, 18th Annual Conference of the International Speech Communication Association. Stockholm, Sweden (August 2017), http://pyannote.github.io/pyannote-metrics
2017
Cited alongside, same era.
2017
Cited alongside, same era.
Maharaj, T., Ballas, N., Rohrbach, A., Courville, A., Pal, C.: A dataset and exploration of models for understanding video data through fill-in-the-blank question-answering. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. pp. 6884–6893 (2017)
2017
Cited alongside, same era.
Rohrbach, A., Torabi, A., Rohrbach, M., Tandon, N., Pal, C., Larochelle, H., Courville, A., Schiele, B.: Movie description. International Journal of Computer Vision 123
2017
Cited alongside, same era.
Smith, J.R., Joshi, D., Huet, B., Hsu, W., Cota, J.: Harnessing ai for augmenting creativity: Application to movie trailer creation. In: Proceedings of the 25th ACM international conference on Multimedia. pp. 1799–1808 (2017)
2017
Cited alongside, same era.
Xiong, Y., Huang, Q., Guo, L., Zhou, H., Zhou, B., Lin, D.: A graph-based framework to bridge movies and synopses. In: The IEEE International Conference on Computer Vision (ICCV) (October 2019)
2019
Later among the works it cites.
Bain, M., Nagrani, A., Brown, A., Zisserman, A.: Condensed movies: Story based retrieval with contextual embeddings (2020)
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
Chen, H., Xie, W., Vedaldi, A., Zisserman, A.: Vggsound: A large-scale audio-visual dataset. In: International Conference on Acoustics, Speech, and Signal Processing (ICASSP) (2020)
2020
Later among the works it cites.
Huang, Q., Xiong, Y., Rao, A., Wang, J., Lin, D.: Movienet: A holistic dataset for movie understanding. In: The European Conference on Computer Vision (ECCV) (2020)
2020
Later among the works it cites.
Huang, Q., Yang, L., Huang, H., Wu, T., Lin, D.: Caption-supervised face recognition: Training a state-of-the-art face model without manual annotation. In: The European Conference on Computer Vision (ECCV) (2020)
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
Rao, A., Wang, J., Xu, L., Jiang, X., Huang, Q., Zhou, B., Lin, D.: A unified framework for shot type classification based on subject centric lens. In: The European Conference on Computer Vision (ECCV) (2020)
2020
Later among the works it cites.
Rao, A., Xu, L., Xiong, Y., Xu, G., Huang, Q., Zhou, B., Lin, D.: A local-to-global approach to multi-modal movie scene segmentation. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. pp. 10146–10155 (2020)
2020
Later among the works it cites.
Wang, W., Tran, D., Feiszli, M.: What makes training multi-modal classification networks hard? In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. pp. 12695–12705 (2020)
2020
Later among the works it cites.
Wu, T., Huang, Q., Liu, Z., Wang, Y., Lin, D.: Distribution-balanced loss for multi-label classification in long-tailed datasets. In: European Conference on Computer Vision. pp. 162–178. Springer (2020)
2020
Later among the works it cites.
Xia, J., Rao, A., Xu, L., Huang, Q., Wen, J., Lin, D.: Online multi-modal person search in videos. In: The European Conference on Computer Vision (ECCV) (2020)
2020
Later among the works it cites.
Pardo, A., Caba, F., Alcazar, J.L., Thabet, A.K., Ghanem, B.: Learning to cut by watching movies. In: Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV). pp. 6858–6868 (October 2021)
2021
Closest in time.