Fetching the paper…
Reading the bibliography…
"How can we animate 3D-characters from a movie script or move robots by simply telling them what we would like them to do?" "How unstructured and complex can we make a sentence and still generate plausible movements from it?" These are questions that need to be answered in the long-run, as the field is still in its infancy.
Long short-term memory
Jürgen Schmidhuber and Sepp Hochreiter · 1997
Earlier work this paper cites.
Interactive topology formation of linguistic space and motion space
Wataru Takano, Dana Kulic, and Yoshihiko Nakamura · 2007
Earlier work this paper cites.
Scenemaker: Intelligent multimodal visualisation of natural language scripts
Eva Hanser, Paul Mc Kevitt, Tom Lunney, and Joan Condell · 2009
Earlier work this paper cites.
Bigram-based natural language model and statistical motion symbol model for scalable language of humanoid robots
Wataru Takano and Yoshihiko Nakamura · 2012
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg Corrado, and Jeffrey Dean · 2013
Earlier work this paper cites.
Zero-shot learning through cross-modal transfer
Richard Socher, Milind Ganjoo, Hamsa Sridhar, Osbert Bastani, Christopher D Manning, and Andrew Y Ng · 2013
Earlier work this paper cites.
On the properties of neural machine translation: Encoder-decoder approaches
Kyunghyun Cho, Bart Van Merriënboer, Dzmitry Bahdanau, and Yoshua Bengio · 2014
Earlier work this paper cites.
Generative adversarial networks
Ian J Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Master motor map (mmm)—framework and toolkit for capturing, representing, and reproducing human motion on humanoid robots
Ömer Terlemez, Stefan Ulbrich, Christian Mandery, Martin Do, Nikolaus Vahrenkamp, and Tamim Asfour · 2014
Earlier work this paper cites.
Hierarchical recurrent neural network for skeleton based action recognition
Yong Du, Wei Wang, and Liang Wang · 2015
Earlier work this paper cites.
Recurrent network models for human dynamics
Katerina Fragkiadaki, Sergey Levine, Panna Felsen, and Jitendra Malik · 2015
Earlier work this paper cites.
Texture synthesis using convolutional neural networks
Leon A Gatys, Alexander S Ecker, and Matthias Bethge · 2015
Earlier work this paper cites.
Learning motion manifolds with convolutional autoencoders
Daniel Holden, Jun Saito, Taku Komura, and Thomas Joyce · 2015
Earlier work this paper cites.
Statistical mutual conversion between whole body motion primitives and linguistic sentences for human motions
Wataru Takano and Yoshihiko Nakamura · 2015
Earlier work this paper cites.
Symbolically structured database for human whole body motions based on association between motion symbols and motion words
Wataru Takano and Yoshihiko Nakamura · 2015
Earlier work this paper cites.
Image style transfer using convolutional neural networks
Leon A Gatys, Alexander S Ecker, and Matthias Bethge · 2016
Earlier work this paper cites.
A deep learning framework for character motion synthesis and editing
Daniel Holden, Jun Saito, and Taku Komura · 2016
Earlier work this paper cites.
Unifying representations and large-scale whole-body motion databases for studying human motion
Christian Mandery, Ömer Terlemez, Martin Do, Nikolaus Vahrenkamp, and Tamim Asfour · 2016
Earlier work this paper cites.
Listen, attend, and walk: Neural mapping of navigational instructions to action sequences
Hongyuan Mei, Mohit Bansal, and Matthew Walter · 2016
Earlier work this paper cites.
The KIT motion-language dataset
Matthias Plappert, Christian Mandery, and Tamim Asfour · 2016
Earlier work this paper cites.
Tatsuro Yamada, Shingo Murata, Hiroaki Arie, and Tetsuya Ogata · 2016
Earlier work this paper cites.
Forecasting human dynamics from static images
Yu-Wei Chao, Jimei Yang, Brian Price, Scott Cohen, and Jia Deng · 2017
Earlier work this paper cites.
Learning human motion models for long-term predictions
Partha Ghosh, Jie Song, Emre Aksan, and Otmar Hilliges · 2017
Earlier work this paper cites.
A recurrent variational autoencoder for human motion synthesis
Ikhsanul Habibie, Daniel Holden, Jonathan Schwarz, Joe Yearsley, and Taku Komura · 2017
Earlier work this paper cites.
Grounded language learning in a simulated 3d world
Karl Moritz Hermann, Felix Hill, Simon Green, Fumin Wang, Ryan Faulkner, Hubert Soyer, David Szepesvari, Wojciech Marian Czarnecki, Max Jaderberg, Denis Teplyashin, et al · 2017
Cited alongside, same era.
Phase-functioned neural networks for character control
Daniel Holden, Taku Komura, and Jun Saito · 2017
Cited alongside, same era.
Temporal convolutional networks for action segmentation and detection
Colin Lea, Michael D Flynn, Rene Vidal, Austin Reiter, and Gregory D Hager · 2017
Cited alongside, same era.
Auto-conditioned recurrent networks for extended complex human motion synthesis
Zimo Li, Yi Zhou, Shuangjiu Xiao, Chong He, Zeng Huang, and Hao Li · 2017
Cited alongside, same era.
On human motion prediction using recurrent neural networks
Julieta Martinez, Michael J Black, and Javier Romero · 2017
Cited alongside, same era.
Coalescing narrative and dialogue for grounded pose forecasting
Chaitanya Ahuja · 2019
Later among the works it cites.
Language2pose: Natural language grounded pose forecasting
C. Ahuja and L. Morency · 2019
Later among the works it cites.
Structured prediction helps 3d human motion modelling
Emre Aksan, Manuel Kaufmann, and Otmar Hilliges · 2019
Later among the works it cites.
Action-agnostic human pose forecasting
Hsu-kuang Chiu, Ehsan Adeli, Borui Wang, De-An Huang, and Juan Carlos Niebles · 2019
Later among the works it cites.
What does bert look at? an analysis of bert’s attention
Kevin Clark, Urvashi Khandelwal, Omer Levy, and Christopher D Manning · 2019
Later among the works it cites.
Stylistic locomotion modeling with conditional variational autoencoder
Han Du, Erik Herrmann, Janis Sprenger, Noshaba Cheema, Somayeh Hosseini, Klaus Fischer, and Philipp Slusallek · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Speech-to-gesture generation: A challenge in deep learning approach with bi-directional lstm
Kenta Takeuchi, Dai Hasegawa, Shinichi Shirakawa, Naoshi Kaneko, Hiroshi Sakuta, and Kazuhiko Sumi · 2017
Cited alongside, same era.
Learning to generate long-term future via hierarchical prediction
Ruben Villegas, Jimei Yang, Yuliang Zou, Sungryull Sohn, Xunyu Lin, and Honglak Lee · 2017
Cited alongside, same era.
The pose knows: Video forecasting by generating pose futures
Jacob Walker, Kenneth Marino, Abhinav Gupta, and Martial Hebert · 2017
Cited alongside, same era.
Text2action: Generative adversarial synthesis from language to action
H. Ahn, T. Ha, Y. Choi, H. Yoo, and S. Oh · 2018
Cited alongside, same era.
Deep motifs and motion signatures
Andreas Aristidou, Daniel Cohen-Or, Jessica K Hodgins, Yiorgos Chrysanthou, and Ariel Shamir · 2018
Cited alongside, same era.
Multimodal machine learning: A survey and taxonomy
Tadas Baltrušaitis, Chaitanya Ahuja, and Louis-Philippe Morency · 2018
Cited alongside, same era.
Hp-gan: Probabilistic 3d human motion prediction via gan
Emad Barsoum, John Kender, and Zicheng Liu · 2018
Cited alongside, same era.
Later among the works it cites.
Forecasting people trajectories and head poses by jointly reasoning on tracklets and vislets
Irtiza Hasan, Francesco Setti, Theodore Tsesmelis, Vasileios Belagiannis, Sikandar Amin, Alessio Del Bue, Marco Cristani, and Fabio Galasso · 2019
Later among the works it cites.
Bihmp-gan: Bidirectional 3d human motion prediction gan
Jogendra Nath Kundu, Maharshi Gor, and R Venkatesh Babu · 2019
Later among the works it cites.
Ian Tenney, Patrick Xia, Berlin Chen, Alex Wang, Adam Poliak, R Thomas McCoy, Najoung Kim, Benjamin Van Durme, Samuel R Bowman, Dipanjan Das, et al · 2019
Later among the works it cites.
A multiscale visualization of attention in the transformer model
Jesse Vig · 2019
Later among the works it cites.
Combining recurrent neural networks and adversarial training for human motion synthesis and control
Zhiyong Wang, Jinxiang Chai, and Shihong Xia · 2019
Later among the works it cites.
Real-time human motion forecasting using a rgb camera
Erwin Wu and Hideki Koike · 2019
Later among the works it cites.
Style transfer for co-speech gesture animation: A multi-speaker conditional-mixture approach
Chaitanya Ahuja, Dong Won Lee, Yukiko I Nakano, and Louis-Philippe Morency · 2020
Later among the works it cites.
Learning dynamic relationships for 3d human motion prediction
Qiongjie Cui, Huaijiang Sun, and Fei Yang · 2020
Later among the works it cites.
Learning to dance: A graph convolutional adversarial network to generate realistic dance motions from audio
Joao P Ferreira, Thiago M Coutinho, Thiago L Gomes, José F Neto, Rafael Azevedo, Renato Martins, and Erickson R Nascimento · 2020
Later among the works it cites.
Moglow: Probabilistic and controllable motion synthesis using normalising flows
Gustav Eje Henter, Simon Alexanderson, and Jonas Beskow · 2020
Later among the works it cites.
Constructing human motion manifold with sequential networks
Deok-Kyeong Jang and Sung-Hee Lee · 2020
Later among the works it cites.
Cross-conditioned recurrent networks for long-term synthesis of inter-person human motion interactions
Jogendra Nath Kundu, Himanshu Buckchash, Priyanka Mandikal, Anirudh Jamkhandi, Venkatesh Babu RADHAKRISHNAN, et al · 2020
Later among the works it cites.
Disentangling human dynamics for pedestrian locomotion forecasting with noisy supervision
Karttikeya Mangalam, Ehsan Adeli, Kuan-Hui Lee, Adrien Gaidon, and Juan Carlos Niebles · 2020
Later among the works it cites.
Social-stgcnn: A social spatio-temporal graph convolutional neural network for human trajectory prediction
Abduallah Mohamed, Kun Qian, Mohamed Elhoseiny, and Christian Claudel · 2020
Later among the works it cites.
Physcap: Physically plausible monocular 3d motion capture in real time
Soshi Shimada, Vladislav Golyanik, Weipeng Xu, and Christian Theobalt · 2020
Later among the works it cites.
Hierarchical style-based networks for motion synthesis
Jingwei Xu, Huazhe Xu, Bingbing Ni, Xiaokang Yang, Xiaolong Wang, and Trevor Darrell · 2020
Later among the works it cites.
Efficient human motion prediction using temporal convolutional generative adversarial network
Qiongjie Cui, Huaijiang Sun, Yue Kong, Xiaoqian Zhang, and Yanmeng Li · 2021
Closest in time.