Fetching the paper…
Reading the bibliography…
People often watch videos on the web to learn how to cook new recipes, assemble furniture or repair a computer.
L. Wang, Z. Tong, B. Ji, and G. Wu, “Tdn: Temporal difference networks for efficient action recognition,” in
1904
Earlier work this paper cites.
K. W. Church, “A stochastic parts program and noun phrase parser for unrestricted text,” in
1989
Earlier work this paper cites.
S. M. LaValle and J. J. Kuffner, “Rapidly-exploring random trees: Progress and prospects,”
2001
Earlier work this paper cites.
S. Nikolaidis, R. Ueda, A. Hayashi, and T. Arai, “Optimal camera placement considering mobile robot trajectory,” in
2009
Earlier work this paper cites.
M. Grundmann, V. Kwatra, M. Han, and I. Essa, “Efficient hierarchical graph-based video segmentation,” in
2010
Earlier work this paper cites.
K. Pastra and Y. Aloimonos, “The minimalist grammar of action,”
2012
Earlier work this paper cites.
2013
Earlier work this paper cites.
Y. Yang, A. Guha, C. Fermuller, and Y. Aloimonos, “A cognitive system for understanding human manipulation actions,”
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
Y. Yang, Y. Li, C. Fermüller, and Y. Aloimonos, “Robot learning manipulation action plans by ”watching” unconstrained videos from the world wide web.” in
2015
Earlier work this paper cites.
Y. Yang, C. Fermuller, Y. Li, and Y. Aloimonos, “Grasp type revisited: A modern perspective on a classical feature for vision,” in
2015
Earlier work this paper cites.
C. Finn, T. Yu, T. Zhang, P. Abbeel, and S. Levine, “One-shot visual imitation learning via meta-learning,” in
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
J. Carreira and A. Zisserman, “Quo vadis, action recognition? a new model and the kinetics dataset,” in
2017
Earlier work this paper cites.
A. Salvador, N. Hynes, Y. Aytar, J. Marin, F. Ofli, I. Weber, and A. Torralba, “Learning cross-modal embeddings for cooking recipes and food images,” in
2017
Earlier work this paper cites.
2018
Cited alongside, same era.
D. Paulius, A. B. Jelodar, and Y. Sun, “Functional Object-Oriented Network: Construction & Expansion,” in
2018
Cited alongside, same era.
F. Sener and A. Yao, “Unsupervised learning and segmentation of complex activities from video,” in
2018
Cited alongside, same era.
2018
Cited alongside, same era.
L. Zhou, C. Xu, and J. J. Corso, “Towards automatic learning of procedures from web instructional videos,” in
2018
Cited alongside, same era.
C.-Y. Wu, C. Feichtenhofer, H. Fan, K. He, P. Krahenbuhl, and R. Girshick, “Long-term feature banks for detailed video understanding,” in
2019
Closest in time.
H. Zhang, P.-J. Lai, S. Paul, S. Kothawade, and S. Nikolaidis, “Learning collaborative action plans from youtube videos,” in
2019
Closest in time.
R. Holladay, T. Lozano-Pérez, and A. Rodriguez, “Force-and-motion constrained planning for tool use,”
2019
Closest in time.
2019
Closest in time.
Z. Wang, Z. Gao, L. Wang, Z. Li, and G. Wu, “Boundary-aware cascade networks for temporal action segmentation,” in
2020
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. Hallac, P. Nystrup, and S. Boyd, “Greedy gaussian segmentation of multivariate time series,”
2018
Cited alongside, same era.
A. Nguyen, D. Kanoulas, L. Muratore, D. G. Caldwell, and N. G. Tsagarakis, “Translating videos to commands for robotic manipulation with deep recurrent neural networks,” in
2018
Cited alongside, same era.
2018
Cited alongside, same era.
S.-H. Sun, H. Noh, S. Somasundaram, and J. Lim, “Neural program synthesis from diverse demonstration videos,” in
2018
Cited alongside, same era.
2018
Cited alongside, same era.
C. Gu, C. Sun, D. A. Ross, C. Vondrick, C. Pantofaru, Y. Li, S. Vijayanarasimhan, G. Toderici, S. Ricco, R. Sukthankar,
2018
Cited alongside, same era.
D.-A. Huang, S. Buch, L. Dery, A. Garg, L. Fei-Fei, and J. Carlos Niebles, “Finding it: Weakly-supervised reference-aware visual grounding in instructional videos,” in
2018
Cited alongside, same era.
Y. Huang, Y. Sugano, and Y. Sato, “Improving action segmentation via graph-based temporal reasoning,” in
2020
Closest in time.
J. Tang, J. Xia, X. Mu, B. Pang, and C. Lu, “Asynchronous interaction aggregation for action detection,” in
2020
Closest in time.
2021
Closest in time.
2021
Closest in time.
Z. Li, Y. Abu Farha, and J. Gall, “Temporal action segmentation from timestamp supervision,” in
2021
Closest in time.
J. Li and S. Todorovic, “Action shuffle alternating learning for unsupervised action segmentation,” in
2021
Closest in time.
P. Tirupattur, K. Duarte, Y. S. Rawat, and M. Shah, “Modeling multi-label action dependencies for temporal action localization,” in
2021
Closest in time.
Y. Chen, Z. Zhang, C. Yuan, B. Li, Y. Deng, and W. Hu, “Channel-wise topology refinement graph convolution for skeleton-based action recognition,” in
2021
Closest in time.
Z. Wang, H. Chen, X. Li, C. Liu, Y. Xiong, J. Tighe, and C. Fowlkes, “Sscap: Self-supervised co-occurrence action parsing for unsupervised temporal action segmentation,” in
2022
Closest in time.