Fetching the paper…
Reading the bibliography…
We propose V2CNet, a new deep learning framework to automatically translate the demonstration videos to commands that can be directly used in robotic applications.
In: International Conference on Robotics and Automation (ICRA)
Zhang T, McCarthy Z, Jow O, Lee D, Goldberg K and Abbeel P (2018) Deep imitation learning for complex manipulation tasks from virtual reality teleoperation · 1903
Earlier work this paper cites.
In: Conference of the North American Chapter of the Association for Computational Linguistics on Human Language Technology
Toutanova K, Klein D, Manning CD and Singer Y (2003) Feature-rich part-of-speech tagging with a cyclic dependency network · 2003
Earlier work this paper cites.
Robotics and Autonomous Systems
Argall BD, Chernova S, Veloso M and Browning B (2009) A survey of robot learning from demonstration · 2009
Earlier work this paper cites.
In: International Conference on Advanced Robotics (ICAR)
Calinon S, Evrard P, Gribovskaya E, Billard A and Kheddar A (2009) Learning collaborative manipulation tasks by demonstration using a haptic interface · 2009
Earlier work this paper cites.
Cambridge University Press
Chrystopher L Nehaniv KD (2009) Imitation and Social Learning in Robots, Humans and Animals Behavioural, Social and Communicative Dimensions · 2009
Earlier work this paper cites.
In: International Conference on Robotics and Automation (ICRA)
Pastor P, Hoffmann H, Asfour T and Schaal S (2009) Learning and generalization of motor skills by learning from demonstration · 2009
Earlier work this paper cites.
Robotics & Automation Magazine
Kruger V, Herzog DL, Baby S, Ude A and Kragic D (2010) Learning actions from observations · 2010
Earlier work this paper cites.
In: International Conference on Intelligent Robots and Systems (IROS)
Pastor P, Righetti L, Kalakrishnan M and Schaal S (2011) Online movement adaptation based on previous sensor experiences · 2011
Earlier work this paper cites.
In: AAAI Conference on Artificial Intelligence
Tellex S, Kollar T, Dickerson S, Walter MR, Banerjee AG, Teller SJ and Roy N (2011) Understanding natural language commands for robotic navigation and mobile manipulation · 2011
Earlier work this paper cites.
In: International Conference on Human-Robot Interaction (HRI)
Akgun B, Cakmak M, Yoo JW and Thomaz AL (2012) Trajectories and keyframes for kinesthetic teaching: A human-robot interaction perspective · 2012
Earlier work this paper cites.
In: Advances in Neural Information Processing Systems (NIPS)
Krizhevsky A, Sutskever I and Hinton GE (2012) Imagenet classification with deep convolutional neural networks · 2012
Earlier work this paper cites.
In: International Conference on Intelligent Robots and Systems (IROS)
Guadarrama S, Riano L, Golland D, Go D, Jia Y, Klein D, Abbeel P, Darrell T et al. (2013) Grounding spatial relations for human-robot interaction · 2013
Earlier work this paper cites.
Robotics and Autonomous Systems
Lee K, Su Y, Kim TK and Demiris Y (2013) A syntactic approach to robot imitation learning using probabilistic activity grammars · 2013
Earlier work this paper cites.
In: Advances in Neural Information Processing Systems (NIPS)
Mikolov T, Sutskever I, Chen K, Corrado GS and Dean J (2013) Distributed representations of words and phrases and their compositionality · 2013
Earlier work this paper cites.
Cho K, van Merrienboer B, Gülçehre Ç, Bougares F, Schwenk H and Bengio Y (2014) Learning phrase representations using RNN encoder-decoder for statistical machine translation · 2014
Earlier work this paper cites.
In: Computer Vision and Pattern Recognition (CVPR)
Donahue J, Hendricks LA, Guadarrama S, Rohrbach M, Venugopalan S, Saenko K and Darrell T (2014) Long-term recurrent convolutional networks for visual recognition and description · 2014
Earlier work this paper cites.
In: Conference on Computer Vision and Pattern Recognition (CVPR)
Karpathy A, Toderici G, Shetty S, Leung T, Sukthankar R and Fei-Fei L (2014) Large-scale video classification with convolutional neural networks · 2014
Earlier work this paper cites.
In: International Conference for Learning Representations (ICLR)
Kingma D and Ba J (2014) Adam: A method for stochastic optimization · 2014
Earlier work this paper cites.
In: International Conference on Robotics and Automation (ICRA)
Koenemann J, Burget F and Bennewitz M (2014) Real-time imitation of human whole-body motions by humanoids · 2014
Cited alongside, same era.
In: Conference on Computer Vision and Pattern Recognition (CVPR)
Kuehne H, Arslan A and Serre T (2014) The language of actions: Recovering the syntax and semantics of goal-directed human activities · 2014
Cited alongside, same era.
In: Advances in Neural Information Processing Systems (NIPS)
Sutskever I, Vinyals O and Le QV (2014) Sequence to sequence learning with neural networks · 2014
Cited alongside, same era.
arXiv preprint arXiv:1412.4729
Venugopalan S, Xu H, Donahue J, Rohrbach M, Mooney R and Saenko K (2014) Translating videos to natural language using deep recurrent neural networks · 2014
Cited alongside, same era.
Software available from tensorflow.org
Abadi M, Agarwal A, Barham P, Brevdo E, Chen Z, Citro C, Corrado GS, Davis A, Dean J, Devin M, Ghemawat S, Goodfellow I, Harp A, Irving G, Isard M, Jia Y, Jozefowicz R, Kaiser L, Kudlur M, Levenberg J, Mané D, Monga R, Moore S, Murray D, Olah C, Schuster M, Shlens J, Steiner B, Sutskever I, Talwar K, Tucker P, Vanhoucke V, Vasudevan V, Viégas F, Vinyals O, Warden P, Wattenberg M, Wicke M, Yu Y and Zheng X (2015) TensorFlow: Large-scale machine learning on heterogeneous systems · 2015
In: Conference on Computer Vision and Pattern Recognition (CVPR)
Venugopalan S, Rohrbach M, Donahue J, Mooney R, Darrell T and Saenko K (2016) Sequence to sequence - Video to text · 2016
Later among the works it cites.
In: International Conference on Intelligent Robots and Systems (IROS)
Welschehold T, Dornhege C and Burgard W (2016) Learning manipulation actions from human demonstrations · 2016
Later among the works it cites.
In: Conference on Computer Vision and Pattern Recognition (CVPR)
Xu J, Mei T, Yao T and Rui Y (2016) Msr-vtt: A large video description dataset for bridging video and language · 2016
Later among the works it cites.
In: Conference on Computer Vision and Pattern Recognition (CVPR)
You Q, Jin H, Wang Z, Fang C and Luo J (2016) Image captioning with semantic attention · 2016
Later among the works it cites.
In: Conference on Robot Learning (CoRL)
Finn C, Yu T, Zhang T, Abbeel P and Levine S (2017) One-shot visual imitation learning via meta-learning · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Artificial Intelligence 10.1016/j.artint.2015.08.009
Ramirez-Amaro K, Beetz M and Cheng G (2015) Transferring skills to humanoid robots by extracting semantic representations from observations of human activities · 2015
Cited alongside, same era.
In: International Conference on Robotics and Automation (ICRA)
Rocchi A, Mingo Hoffman E, Caldwell D and Tsagarakis N (2015) OpenSoT: A Whole-Body Control Library for the Compliant Humanoid Robot COMAN · 2015
Cited alongside, same era.
International Journal of Computer Vision (IJCV)
Russakovsky O, Deng J, Su H, Krause J, Satheesh S, Ma S, Huang Z, Karpathy A, Khosla A, Bernstein M, Berg A and Fei-Fei L (2015) ImageNet Large Scale Visual Recognition Challenge · 2015
Cited alongside, same era.
In: International Conference on Computer Vision (CVPR)
Sun L, Jia K, Yeung DY and Shi BE (2015) Human action recognition using factorized spatio-temporal convolutional networks · 2015
Cited alongside, same era.
In: International Conference on Computer Vision (ICCV)
Tran D, Bourdev L, Fergus R, Torresani L and Paluri M (2015) Learning spatiotemporal features with 3d convolutional networks · 2015
Cited alongside, same era.
In: AAAI Conference on Artificial Intelligence
Yang Y, Li Y, Fermuller C and Aloimonos Y (2015) Robot learning manipulation action plans by ”watching” unconstrained videos from the world wide web · 2015
Cited alongside, same era.
In: International Conference on Computer Vision (CVPR)
Yao L, Torabi A, Cho K, Ballas N, Pal C, Larochelle H and Courville A (2015) Describing videos by exploiting temporal structure · 2015
Cited alongside, same era.
Gan Z, Gan C, He X, Pu Y, Tran K, Gao J, Carin L and Deng L (2017) Semantic compositional networks for visual captioning · 2017
Later among the works it cites.
In: International Conference on Computer Vision (CVPR)
Lea C, Flynn MD, Vidal R, Reiter A and Hager GD (2017) Temporal convolutional networks for action segmentation and detection · 2017
Later among the works it cites.
In: International Conference on Robotic Computing
Muratore L, Laurenzi A, Mingo Hoffman E, Rocchi A, Caldwell DG and Tsagarakis NG (2017) Xbotcore: A real-time cross-robot software platform · 2017
Later among the works it cites.
In: International Conference on Computer Vision (CVPR)
Pan Y, Yao T, Li H and Mei T (2017) Video captioning with transferred semantic attributes · 2017
Later among the works it cites.
Plappert M, Mandery C and Asfour T (2017) Learning a bidirectional mapping between human whole-body motion and natural language using deep recurrent neural networks · 2017
Later among the works it cites.
In: Conference on Computer Vision and Pattern Recognition (CVPR)
Ramanishka V, Das A, Zhang J and Saenko K (2017) Top-down visual saliency guided by captions · 2017
Later among the works it cites.
In: Conference on Robot Learning (CoRL)
Rana MA, Mukadam M, Ahmadzadeh SR, Chernova S and Boots B (2017) Towards robust skill generalization: Unifying learning from demonstration and motion planning · 2017
Later among the works it cites.
In: International Conference on Computer Vision (ICCV)
Rhinehart N and Kitani KM (2017) First-person activity forecasting with online inverse reinforcement learning · 2017
Later among the works it cites.
In: International Conference on Learning Representations (ICLR)
Stadie BC, Abbeel P and Sutskever I (2017) Third-person imitation learning · 2017
Later among the works it cites.
Journal of Field Robotics
Tsagarakis NG, Caldwell DG, Negrello F, Choi W, Baccelliere L, Loc V, Noorden J, Muratore L, Margan A, Cardellino A, Natale L, Mingo Hoffman E, Dallali H, Kashiri N, Malzahn J, Lee J, Kryczka P, Kanoulas D, Garabini M, Catalano M, Ferrati M, Varricchio V, Pallottino L, Pavan C, Bicchi A, Settimi A, Rocchi A and Ajoudani A (2017) Walk-man: A high-performance humanoid platform for realistic environments · 2017
Later among the works it cites.
In: International Conference Robotics and Automation (ICRA)
Do TT, Nguyen A and Reid I (2018) Affordancenet: An end-to-end deep learning approach for object affordance detection · 2018
Later among the works it cites.
arXiv preprint arXiv:1802.01557
Yu T, Finn C, Xie A, Dasari S, Zhang T, Abbeel P and Levine S (2018) One-shot imitation from observing humans via domain-adaptive meta-learning · 2018
Later among the works it cites.