Fetching the paper…
Reading the bibliography…
Learning actions from human demonstration video is promising for intelligent robotic systems.
T. Flash and N. Hogan, “The coordination of arm movements: an experimentally confirmed mathematical model,” Journal of Neuroscience
1985
Earlier work this paper cites.
K. Ikeuchi and T. Suchiro, “Towards an assembly plan from observation. i. assembly task recognition using face-contact relations (polyhedral objects),” in ICRA
1992
Earlier work this paper cites.
R. Kohavi, “A study of cross-validation and bootstrap for accuracy estimation and model selection,” in IJCAI
1995
Earlier work this paper cites.
S. Schaal, “Is imitation learning the route to humanoid robots?,” Trends in Cognitive Sciences
1999
Earlier work this paper cites.
Y. Kuniyoshi, Y. Yorozu, M. Inaba, and H. Inoue, “From visuo-motor self learning to early imitation – a neural architecture for humanoid learning,” in ICRA
2003
Earlier work this paper cites.
A. J. Ijspeert, J. Nakanishi, and S. Schaal, “Learning attractor landscapes for learning motor primitives,” in NIPS
2003
Earlier work this paper cites.
T. Shiratori, A. Nakazawa, and K. Ikeuchi, “Detecting dance motion structure through music analysis,” in FG
2004
Earlier work this paper cites.
S. Nakaoka, A. Nakazawa, F. Kanehiro, K. Kaneko, M. Morisawa, H. Hirukawa, and K. Ikeuchi, “Learning from observation paradigm: Leg task models for enabling a biped humanoid robot to imitate human dances,” IJRR
2007
Earlier work this paper cites.
B. D. Argall, S. Chernova, M. Veloso, and B. Browning, “A survey of robot learning from demonstration,” Robotics and Autonomous Systems
2009
Earlier work this paper cites.
D. Lee and C. Ott, “Incremental kinesthetic teaching of motion primitives using the motion refinement tube,” Autonomous Robots
2011
Earlier work this paper cites.
B. Akgun, M. Cakmak, J. W. Yoo, and A. L. Thomaz, “Trajectories and keyframes for kinesthetic teaching: A human-robot interaction perspective,” in HRI
2012
Earlier work this paper cites.
Y. Kawahara and M. Sugiyama, “Sequential change-point detection based on direct density-ratio estimation,” Statistical Analysis and Data Mining: The ASA Data Science Journal
2012
Earlier work this paper cites.
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean, “Distributed representations of words and phrases and their compositionality,” in NIPS
2013
Earlier work this paper cites.
S. Liu, M. Yamada, N. Collier, and M. Sugiyama, “Change-point detection in time-series data by relative density-ratio estimation,” Neural Networks
2013
Earlier work this paper cites.
J. Koenemann, F. Burget, and M. Bennewitz, “Real-time imitation of human whole-body motions by humanoids,” in ICRA
2014
Earlier work this paper cites.
J. Donahue, L. Anne Hendricks, S. Guadarrama, M. Rohrbach, S. Venugopalan, K. Saenko, and T. Darrell, “Long-term recurrent convolutional networks for visual recognition and description,” in CVPR
2015
Cited alongside, same era.
S. Venugopalan, M. Rohrbach, J. Donahue, R. Mooney, T. Darrell, and K. Saenko, “Sequence to sequence-video to text,” in ICCV
2015
Cited alongside, same era.
S. Ioffe and C. Szegedy, “Batch normalization: Accelerating deep network training by reducing internal covariate shift,” vol. 37 of JMLR
2015
Cited alongside, same era.
R. Blanco, G. Ottaviano, and E. Meij, “Fast and space-efficient entity linking for queries,” in WSDM
2015
Cited alongside, same era.
M. Kusner, Y. Sun, N. Kolkin, and K. Weinberger, “From word embeddings to document distances,” in ICML
2015
Cited alongside, same era.
A. Nguyen, D. Kanoulas, L. Muratore, D. G. Caldwell, and N. G. Tsagarakis, “Translating videos to commands for robotic manipulation with deep recurrent neural networks,” in ICRA
2018
Later among the works it cites.
M. Plappert, C. Mandery, and T. Asfour, “Learning a bidirectional mapping between human whole-body motion and natural language using deep recurrent neural networks,” Robotics and Autonomous Systems
2018
Later among the works it cites.
K. Ikeuchi, Z. Ma, Z. Yan, S. Kudoh, and M. Nakamura, “Describing upper-body motions based on labanotation for learning-from-observation robots,” IJCV
2018
Later among the works it cites.
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Z. Shou, D. Wang, and S.-F. Chang, “Temporal action localization in untrimmed videos via multi-stage cnns,” in CVPR
2016
Cited alongside, same era.
J. Xu, T. Mei, T. Yao, and Y. Rui, “Msr-vtt: A large video description dataset for bridging video and language,” in CVPR
2016
Cited alongside, same era.
C. Lea, R. Vidal, and G. D. Hager, “Learning convolutional action primitives for fine-grained action recognition,” in ICRA
2016
Cited alongside, same era.
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in CVPR
2016
Cited alongside, same era.
C. Pérez-D’Arpino and J. A. Shah, “C-learn: Learning geometric constraints from demonstrations for multi-step manipulation in shared autonomy,” in ICRA
2017
Cited alongside, same era.
E. Ilg, N. Mayer, T. Saikia, M. Keuper, A. Dosovitskiy, and T. Brox, “Flownet 2.0: Evolution of optical flow estimation with deep networks,” in CVPR
2017
Cited alongside, same era.
2019
Later among the works it cites.
2019
Later among the works it cites.
T. Lin, X. Liu, X. Li, E. Ding, and S. Wen, “Bmn: Boundary-matching network for temporal action proposal generation,” in ICCV
2019
Later among the works it cites.
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
“Azure kinect dk – develop ai models — microsoft azure.” https://azure.microsoft.com/en-us/services/kinect-dk/
2020
Closest in time.
J. Lei, L. Wang, Y. Shen, D. Yu, T. L. Berg, and M. Bansal, “Mart: Memory-augmented recurrent transformer for coherent video paragraph captioning,” in ACL
2020
Closest in time.
M. Xu, C. Zhao, D. S. Rojas, A. Thabet, and B. Ghanem, “G-tad: Sub-graph localization for temporal action detection,” in CVPR
2020
Closest in time.