Fetching the paper…
Reading the bibliography…
Our goal is to synthesize 3D human motions given textual inputs describing simultaneous actions, for example 'waving hand' while 'walking' at the same time.
Animating rotation with quaternion curves
Ken Shoemake · 1985
Earlier work this paper cites.
Motion synthesis from annotations
Okan Arikan, David A. Forsyth, and James F. O’Brien · 2003
Earlier work this paper cites.
Representing cyclic human motion using functional analysis
Dirk Ormoneit, Michael J. Black, T. Hastie, and H. Kjellström · 2005
Earlier work this paper cites.
Learning people detection models from few training samples
Leonid Pishchulin, Arjun Jain, Christian Wojek, Mykhaylo Andriluka, Thorsten Thormählen, and Bernt Schiele · 2011
Earlier work this paper cites.
Auto-encoding variational bayes
Diederik P. Kingma and Max Welling · 2014
Earlier work this paper cites.
SMPL: A skinned multi-person linear model
Matthew Loper, Naureen Mahmood, Javier Romero, Gerard Pons-Moll, and Michael J. Black · 2015
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics
Jascha Sohl-Dickstein, Eric Weiss, Niru Maheswaranathan, and Surya Ganguli · 2015
Earlier work this paper cites.
Synthesizing training images for boosting human 3D pose estimation
W. Chen, H. Wang, Y. Li, H. Su, Z. Wang, C. Tu, D. Lischinski, D. Cohen-Or, and B. Chen · 2016
Earlier work this paper cites.
A deep learning framework for character motion synthesis and editing
Daniel Holden, Jun Saito, and Taku Komura · 2016
Earlier work this paper cites.
The KIT motion-language dataset
Matthias Plappert, Christian Mandery, and Tamim Asfour · 2016
Earlier work this paper cites.
MoCap-guided data augmentation for 3D pose estimation in the wild
Grégory Rogez and Cordelia Schmid · 2016
Earlier work this paper cites.
HP-GAN: probabilistic 3D human motion prediction via GAN
Emad Barsoum, John Kender, and Zicheng Liu · 2017
Earlier work this paper cites.
A recurrent variational autoencoder for human motion synthesis
I. Habibie, Daniel Holden, Jonathan Schwarz, J. Yearsley, and T. Komura · 2017
Earlier work this paper cites.
FlowNet 2.0: Evolution of optical flow estimation with deep networks
E. Ilg, N. Mayer, T. Saikia, M. Keuper, A. Dosovitskiy, and T. Brox · 2017
Earlier work this paper cites.
On human motion prediction using recurrent neural networks
Julieta Martinez, Michael J. Black, and Javier Romero · 2017
Earlier work this paper cites.
From red wine to red tomato: Composition with context
Ishan Misra, Abhinav Gupta, and Martial Hebert · 2017
Earlier work this paper cites.
Learning from synthetic humans
Gül Varol, Javier Romero, Xavier Martin, Naureen Mahmood, Michael J. Black, Ivan Laptev, and Cordelia Schmid · 2017
Earlier work this paper cites.
Visual relationship detection with internal and external linguistic knowledge distillation
Ruichi Yu, Ang Li, Vlad I. Morariu, and Larry S. Davis · 2017
Earlier work this paper cites.
Text2Action: Generative adversarial synthesis from language to action
Hyemin Ahn, Timothy Ha, Yunho Choi, Hwiyeon Yoo, and Songhwai Oh · 2018
Earlier work this paper cites.
Compositional learning for human object interaction
Keizo Kato, Yin Li, and Abhinav Gupta · 2018
Earlier work this paper cites.
Generating animated videos of human activities from natural language descriptions
Angela S Lin, Lemeng Wu, Rodolfo Corona, Kevin Tai, Qixing Huang, and Raymond J Mooney · 2018
Earlier work this paper cites.
Shuffle-Then-Assemble: Learning object-agnostic visual relationship features
Xu Yang, Hanwang Zhang, and Jianfei Cai · 2018
Earlier work this paper cites.
Language2Pose: Natural language grounded pose forecasting
Chaitanya Ahuja and Louis-Philippe Morency · 2019
Earlier work this paper cites.
Learning joint reconstruction of hands and manipulated objects
Yana Hasson, Gül Varol, Dimitrios Tzionas, Igor Kalevatykh, Michael J. Black, Ivan Laptev, and Cordelia Schmid · 2019
Earlier work this paper cites.
AMASS: Archive of motion capture as surface shapes
Naureen Mahmood, Nima Ghorbani, Nikolaus F. Troje, Gerard Pons-Moll, and Michael J. Black · 2019
Earlier work this paper cites.
DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter
Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf · 2019
Cited alongside, same era.
Neural state machine for character-scene interactions
Sebastian Starke, He Zhang, Taku Komura, and Jun Saito · 2019
Cited alongside, same era.
Compositional video prediction
Yufei Ye, Maneesh Singh, Abhinav Gupta, and Shubham Tulsiani · 2019
Cited alongside, same era.
On the continuity of rotation representations in neural networks
Yi Zhou, Connelly Barnes, Jingwan Lu, Jimei Yang, and Hao Li · 2019
Cited alongside, same era.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel Ziegler, Jeffrey Wu, Clemens Winter, Chris Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei · 2020
Cited alongside, same era.
Sai Shashank Kalakonda, Shubh Maheshwari, and Ravi Kiran Sarvadevabhatla · 2022
Later among the works it cites.
Conditional motion in-betweening
Jihoon Kim, Taehyun Byun, Seungyoung Shin, Jungdam Won, and Sungjoon Choi · 2022
Later among the works it cites.
GANimator: Neural motion synthesis from a single sequence
Peizhuo Li, Kfir Aberman, Zihan Zhang, Rana Hanocka, and Olga Sorkine-Hornung · 2022
Later among the works it cites.
Towards robust and adaptive motion forecasting: A causal representation perspective
Yuejiang Liu, Riccardo Cadei, Jonas Schweizer, Sherwin Bahmani, and Alexandre Alahi · 2022
Later among the works it cites.
PoseGPT: Quantizing human motion for large scale generative modeling
Thomas Lucas, Fabien Baradel, Philippe Weinzaepfel, and Grégory Rogez · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Context-aware human motion prediction
Enric Corona, Albert Pumarola, G. Alenyà, and F. Moreno-Noguer · 2020
Cited alongside, same era.
Inferring temporal compositions of actions using probabilistic automata
Rodrigo Santa Cruz, Anoop Cherian, Basura Fernando, Dylan Campbell, and Stephen Gould · 2020
Cited alongside, same era.
Action2Motion: Conditioned generation of 3d human motions
Chuan Guo, Xinxin Zuo, Sen Wang, Shihao Zou, Qingyao Sun, Annan Deng, Minglun Gong, and Li Cheng · 2020
Cited alongside, same era.
Robust motion in-betweening
Félix G. Harvey, Mike Yurick, Derek Nowrouzezahrai, and Cristopher Joseph Pal · 2020
Cited alongside, same era.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel · 2020
Cited alongside, same era.
DLow: Diversifying latent flows for diverse human motion prediction
Ye Yuan and Kris M. Kitani · 2020
Cited alongside, same era.
BERTScore: Evaluating text generation with BERT
Tianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger, and Yoav Artzi · 2020
Cited alongside, same era.
BRACE: The breakdancing competition dataset for dance motion synthesis
Davide Moltisanti, Jinyi Wu, Bo Dai, and Chen Change Loy · 2022
Later among the works it cites.
TEMOS: Generating diverse human motions from textual descriptions
Mathis Petrovich, Michael J. Black, and Gül Varol · 2022
Later among the works it cites.
Motron: Multimodal probabilistic human motion forecasting
Tim Salzmann, Marco Pavone, and Markus Ryll · 2022
Later among the works it cites.
Real-time controllable motion transition for characters
Xiangjun Tang, He Wang, Bo Hu, Xu Gong, Ruifan Yi, Qilong Kou, and Xiaogang Jin · 2022
Later among the works it cites.
MotionCLIP: Exposing human motion generation to CLIP space
Guy Tevet, Brian Gordon, Amir Hertz, Amit H. Bermano, and Daniel Cohen-Or · 2022
Later among the works it cites.
Towards diverse and natural scene-aware 3D human motion synthesis
Jingbo Wang, Yu Rong, Jingyuan Liu, Sijie Yan, Dahua Lin, and Bo Dai · 2022
Later among the works it cites.
Reconstructing action-conditioned human-object interactions using commonsense knowledge priors
Xi Wang, Gen Li, Yen-Ling Kuo, Muhammed Kocabas, Emre Aksan, and Otmar Hilliges · 2022
Later among the works it cites.
HUMANISE: Language-conditioned human motion generation in 3D scenes
Zan Wang, Yixin Chen, Tengyu Liu, Yixin Zhu, Wei Liang, and Siyuan Huang · 2022
Later among the works it cites.
MotionDiffuse: Text-driven human motion generation with diffusion model
Mingyuan Zhang, Zhongang Cai, Liang Pan, Fangzhou Hong, Xinying Guo, Lei Yang, and Ziwei Liu · 2022
Later among the works it cites.
The wanderings of Odysseus in 3D scenes
Yan Zhang and Siyu Tang · 2022
Later among the works it cites.
Compositional human-scene interaction synthesis with semantic control
Kaifeng Zhao, Shaofei Wang, Yan Zhang, Thabo Beeler, and Siyu Tang · 2022
Later among the works it cites.
Spatio-temporal gating-adjacency gcn for human motion prediction
Chongyang Zhong, Lei Hu, Zihao Zhang, Yongjing Ye, and Shi hong Xia · 2022
Later among the works it cites.
InstructPix2Pix: Learning to follow image editing instructions
Tim Brooks, Aleksander Holynski, and Alexei A Efros · 2023
Closest in time.
Executing your commands via motion diffusion in latent space
Xin Chen, Biao Jiang, Wen Liu, Zilong Huang, Bin Fu, Tao Chen, Jingyi Yu, and Gang Yu · 2023
Closest in time.
MoFusion: A framework for denoising-diffusion-based motion synthesis
Rishabh Dabral, Muhammad Hamza Mughal, Vladislav Golyanik, and Christian Theobalt · 2023
Closest in time.
FLAME: Free-form language-based motion synthesis & editing
Jihoon Kim, Jiseob Kim, and Sungjoon Choi · 2023
Closest in time.
MultiAct: Long-term 3D human motion generation from multiple action labels
Taeryung Lee, Gyeongsik Moon, and Kyoung Mu Lee · 2023
Closest in time.
Human motion diffusion model
Guy Tevet, Sigal Raab, Brian Gordon, Yonatan Shafir, Daniel Cohen-Or, and Amit H. Bermano · 2023
Closest in time.
T2M-GPT: Generating human motion from textual descriptions with discrete representations
Jianrong Zhang, Yangsong Zhang, Xiaodong Cun, Shaoli Huang, Yong Zhang, Hongwei Zhao, Hongtao Lu, and Xi Shen · 2023
Closest in time.