Fetching the paper…
Reading the bibliography…
We apply a generative segmental model of task structure, guided by narration, to action segmentation in video.
Contrastive bidirectional transformer for temporal representation learning
Chen Sun, Fabien Baradel, Kevin Murphy, and Cordelia Schmid. 2019 · 1906
Earlier work this paper cites.
The Hungarian method for the assignment problem
Harold W Kuhn. 1955 · 1955
Earlier work this paper cites.
Scripts, plans, goals, and understanding: An inquiry into human knowledge structures
Roger C Schank and Robert P Abelson. 1977 · 1977
Earlier work this paper cites.
Individuation of actions from continuous motion
Tanya Sharon and Karen Wynn. 1998 · 1998
Earlier work this paper cites.
Infants parse dynamic action
Dare A Baldwin, Jodie A Baird, Megan M Saylor, and M Angela Clark. 2001 · 2001
Earlier work this paper cites.
Conditional structure versus conditional estimation in NLP models
Dan Klein and Christopher D. Manning. 2002 · 2002
Earlier work this paper cites.
Hidden semi-markov models
Kevin Murphy. 2002 · 2002
Earlier work this paper cites.
Segmenting ambiguous events
Bridgette A Martin and Barbara Tversky. 2003 · 2003
Earlier work this paper cites.
Duration modeling techniques for continuous speech recognition
Janne Pylkkonen and Mikko Kurimo. 2004 · 2004
Earlier work this paper cites.
Event segmentation
Jeffrey M Zacks and Khena M Swallow. 2007 · 2007
Earlier work this paper cites.
Unsupervised learning of narrative event chains
Nathanael Chambers and Dan Jurafsky. 2008 · 2008
Earlier work this paper cites.
Analyzing the errors of unsupervised learning
Percy Liang and Dan Klein. 2008 · 2008
Earlier work this paper cites.
Visualizing data using t-SNE
Laurens van der Maaten and Geoffrey Hinton. 2008 · 2008
Earlier work this paper cites.
Hidden semi-markov models
Shun-Zheng Yu. 2010 · 2010
Earlier work this paper cites.
A database for fine grained activity detection of cooking activities
Marcus Rohrbach, Sikandar Amin, Mykhaylo Andriluka, and Bernt Schiele. 2012 · 2012
Earlier work this paper cites.
Weakly supervised action labeling in videos under ordering constraints
Piotr Bojanowski, Rémi Lajugie, Francis Bach, Ivan Laptev, Jean Ponce, Cordelia Schmid, and Josef Sivic. 2014 · 2014
Earlier work this paper cites.
The language of actions: Recovering the syntax and semantics of goal-directed human activities
Hilde Kuehne, Ali Arslan, and Thomas Serre. 2014 · 2014
Earlier work this paper cites.
Statistical script learning with multi-argument events
Karl Pichotta and Raymond Mooney. 2014 · 2014
Earlier work this paper cites.
Instructional videos for unsupervised harvesting and learning of action examples
Shoou-I Yu, Lu Jiang, and Alexander Hauptmann. 2014 · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2015 · 2015
Cited alongside, same era.
What’s cookin’? interpreting cooking videos using text, speech and vision
Jonathan Malmaud, Jonathan Huang, Vivek Rathod, Nicholas Johnston, Andrew Rabinovich, and Kevin Murphy. 2015 · 2015
Cited alongside, same era.
Learning to predict script events from domain-specific text
Rachel Rudinger, Vera Demberg, Ashutosh Modi, Benjamin Van Durme, and Manfred Pinkal. 2015 · 2015
Cited alongside, same era.
ImageNet Large Scale Visual Recognition Challenge
Olga Russakovsky, Jia Deng, Hao Su, Jonathan Krause, Sanjeev Satheesh, Sean Ma, Zhiheng Huang, Andrej Karpathy, Aditya Khosla, Michael Bernstein, Alexander C. Berg, and Li Fei-Fei. 2015 · 2015
Cited alongside, same era.
Unsupervised semantic parsing of video collections
O. Sener, A. Zamir, S. Savarese, and A. Saxena. 2015 · 2015
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
Weakly supervised action learning with RNN based fine-to-coarse modeling
Alexander Richard, Hilde Kuehne, and Juergen Gall. 2017 · 2017
Later among the works it cites.
R-C3D: Region convolutional 3d network for temporal activity detection
Huijuan Xu, Abir Das, and Kate Saenko. 2017 · 2017
Later among the works it cites.
Temporal action detection with structured segment networks
Yue Zhao, Yuanjun Xiong, Limin Wang, Zhirong Wu, Xiaoou Tang, and Dahua Lin. 2017 · 2017
Later among the works it cites.
Weakly-supervised action segmentation with iterative soft boundary assignment
Li Ding and Chenliang Xu. 2018 · 2018
Later among the works it cites.
Advances in pre-training distributed word representations
Tomas Mikolov, Edouard Grave, Piotr Bojanowski, Christian Puhrsch, and Armand Joulin. 2018 · 2018
Later among the works it cites.
Action sets: Weakly supervised action segmentation without ordering constraints
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Karen Simonyan and Andrew Zisserman. 2015 · 2015
Cited alongside, same era.
Cross-task weakly supervised learning from instructional videos
Dimitri Zhukov, Jean-Baptiste Alayrac, Ramazan Gokberk Cinbis, David Fouhey, Ivan Laptev, and Josef Sivic. 2019 · 2015
Cited alongside, same era.
YouTube-8M: A large-scale video classification benchmark
Sami Abu-El-Haija, Nisarg Kothari, Joonseok Lee, Paul Natsev, George Toderici, Balakrishnan Varadarajan, and Sudheendra Vijayanarasimhan. 2016 · 2016
Cited alongside, same era.
Unsupervised learning from narrated instruction videos
Jean-Baptiste Alayrac, Piotr Bojanowski, Nishant Agrawal, Josef Sivic, Ivan Laptev, and Simon Lacoste-Julien. 2016 · 2016
Cited alongside, same era.
Inside-outside and forward-backward algorithms are just backprop (tutorial paper)
Jason Eisner. 2016 · 2016
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Cited alongside, same era.
Connectionist temporal modeling for weakly supervised action labeling
De-An Huang, Li Fei-Fei, and Juan Carlos Niebles. 2016 · 2016
Cited alongside, same era.
Alexander Richard, Hilde Kuehne, and Juergen Gall. 2018 · 2018
Later among the works it cites.
Unsupervised learning and segmentation of complex activities from video
Fadime Sener and Angela Yao. 2018 · 2018
Later among the works it cites.
Hierarchical quantized representations for script generation
Noah Weber, Leena Shekhar, Niranjan Balasubramanian, and Nathanael Chambers. 2018 · 2018
Later among the works it cites.
Every moment counts: Dense detailed labeling of actions in complex videos
Serena Yeung, Olga Russakovsky, Ning Jin, Mykhaylo Andriluka, Greg Mori, and Li Fei-Fei. 2018 · 2018
Later among the works it cites.
Towards automatic learning of procedures from web instructional videos
Luowei Zhou, Xu Chenliang, and Jason J. Corso. 2018 · 2018
Later among the works it cites.
Benchmarking hierarchical script knowledge
Yonatan Bisk, Jan Buys, Karl Pichotta, and Yejin Choi. 2019 · 2019
Later among the works it cites.
D3TW: Discriminative differentiable dynamic time warping for weakly supervised action alignment and segmentation
Chien-Yi Chang, De-An Huang, Yanan Sui, Li Fei-Fei, and Juan Carlos Niebles. 2019 · 2019
Later among the works it cites.
Unsupervised procedure learning via joint dynamic summarization
Ehsan Elhamifar and Zwe Naing. 2019 · 2019
Later among the works it cites.
MS-TCN: Multi-stage temporal convolutional network for action segmentation
Yazan Abu Farha and Jurgen Gall. 2019 · 2019
Later among the works it cites.
Unsupervised learning of action classes with continuous temporal embedding
Anna Kukleva, Hilde Kuehne, Fadime Sener, and Jurgen Gall. 2019 · 2019
Later among the works it cites.
Finding events in a continuous world: A developmental account
Dani Levine, Daphna Buchsbaum, Kathy Hirsh-Pasek, and Roberta M Golinkoff. 2019 · 2019
Later among the works it cites.
HowTo100M: Learning a text-video embedding by watching hundred million narrated video clips
Antoine Miech, Dimitri Zhukov, Jean-Baptiste Alayrac, Makarand Tapaswi, Ivan Laptev, and Josef Sivic. 2019 · 2019
Later among the works it cites.
COIN: A large-scale dataset for comprehensive instructional video analysis
Yansong Tang, Dajun Ding, Yongming Rao, Yu Zheng, Danyang Zhang, Lili Zhao, Jiwen Lu, and Jie Zhou. 2019 · 2019
Later among the works it cites.
From event representation to linguistic meaning
Ercenur Ünal, Yue Ji, and Anna Papafragou. 2019 · 2019
Later among the works it cites.