Fetching the paper…
Reading the bibliography…
Text-to-motion generation has gained increasing attention, but most existing methods are limited to generating short-term motions that correspond to a single sentence describing a single action.
“Auto-encoding variational bayes”
Diederik Kingma and Max Welling · 2013
Earlier work this paper cites.
“Generative Adversarial Nets”
Ian Goodfellow et al · 2014
Earlier work this paper cites.
“Adam: A method for stochastic optimization”
Diederik Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
“Recurrent network models for human dynamics”
Katerina Fragkiadaki, Sergey Levine, Panna Felsen and Jitendra Malik · 2015
Earlier work this paper cites.
“SMPL: A skinned multi-person linear model”
Matthew Loper et al · 2015
Earlier work this paper cites.
“Deep unsupervised learning using nonequilibrium thermodynamics”
Jascha Sohl-Dickstein, Eric Weiss, Niru Maheswaranathan and Surya Ganguli · 2015
Earlier work this paper cites.
“Learning structured output representation using deep conditional generative models”
Kihyuk Sohn, Honglak Lee and Xinchen Yan · 2015
Earlier work this paper cites.
“Decoupled weight decay regularization”
Ilya Loshchilov and Frank Hutter · 2017
Earlier work this paper cites.
“On human motion prediction using recurrent neural networks”
Julieta Martinez, Michael Black and Javier Romero · 2017
Earlier work this paper cites.
“Embodied hands: modeling and capturing hands and bodies together”
Javier Romero, Dimitrios Tzionas and Michael Black · 2017
Earlier work this paper cites.
“Glow: Generative flow with invertible 1x1 convolutions”
Durk Kingma and Prafulla Dhariwal · 2018
Earlier work this paper cites.
“Language2pose: Natural language grounded pose forecasting”
Chaitanya Ahuja and Louis-Philippe Morency · 2019
Earlier work this paper cites.
“Implicit generation and modeling with energy based models”
Yilun Du and Igor Mordatch · 2019
Earlier work this paper cites.
“Human motion prediction via spatio-temporal inpainting”
Alejandro Hernandez, Jurgen Gall and Francesc Moreno-Noguer · 2019
Earlier work this paper cites.
“Convolutional sequence generation for skeleton-based action synthesis”
Sijie Yan et al · 2019
Earlier work this paper cites.
“Compositional visual generation with energy based models”
Yilun Du, Shuang Li and Igor Mordatch · 2020
Earlier work this paper cites.
“Robust motion in-betweening”
Félix Harvey, Mike Yurick, Derek Nowrouzezahrai and Christopher Pal · 2020
Earlier work this paper cites.
“Denoising diffusion probabilistic models”
Jonathan Ho, Ajay Jain and Pieter Abbeel · 2020
Cited alongside, same era.
“Convolutional autoencoders for human motion infilling”
Manuel Kaufmann et al · 2020
Cited alongside, same era.
“Bayesian adversarial human motion synthesis”
Rui Zhao, Hui Su and Qiang Ji · 2020
Cited alongside, same era.
“WenLan: Bridging vision and language by large-scale multi-modal pre-training”
Yuqi Huo et al · 2021
Cited alongside, same era.
“Glide: Towards photorealistic image generation and editing with text-guided diffusion models”
Alex Nichol et al · 2021
Cited alongside, same era.
“Action-conditioned 3d human motion synthesis with transformer vae”
“MultiAct: Long-Term 3D Human Motion Generation from Multiple Action Labels”
Taeryung Lee, Gyeongsik Moon and Kyoung Lee · 2022
Later among the works it cites.
“Danceformer: Music conditioned 3d dance generation with parametric motion transformer”
Buyu Li, Yongchi Zhao, Shi Zhelun and Lu Sheng · 2022
Later among the works it cites.
“Compositional visual generation with composable diffusion models”
Nan Liu et al · 2022
Later among the works it cites.
“Action-conditioned On-demand Motion Generation”
Qiujing Lu, Yipeng Zhang, Mingjian Lu and Vwani Roychowdhury · 2022
Later among the works it cites.
“Repaint: Inpainting using denoising diffusion probabilistic models”
Andreas Lugmayr et al · 2022
Later among the works it cites.
“Synthesizing Coherent Story with Auto-Regressive Latent Diffusion Models”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Mathis Petrovich, Michael Black and Gül Varol · 2021
Cited alongside, same era.
“BABEL: Bodies, action and behavior with english labels”
Abhinanda Punnakkal et al · 2021
Cited alongside, same era.
“Learning transferable visual models from natural language supervision”
Alec Radford et al · 2021
Cited alongside, same era.
“TEACH: Temporal Action Composition for 3D Humans”
Nikos Athanasiou, Mathis Petrovich, Michael Black and Gül Varol · 2022
Cited alongside, same era.
“Implicit neural representations for variable length human motion generation”
Pablo Cervantes, Yusuke Sekikawa, Ikuro Sato and Koichi Shinoda · 2022
Cited alongside, same era.
“Diffuseq: Sequence to sequence text generation with diffusion models”
Shansan Gong et al · 2022
Cited alongside, same era.
“Generating Diverse and Natural 3D Human Motions From Text”
Chuan Guo et al · 2022
Cited alongside, same era.
Xichen Pan et al · 2022
Later among the works it cites.
“TEMOS: Generating diverse human motions from textual descriptions”
Mathis Petrovich, Michael Black and Gül Varol · 2022
Later among the works it cites.
“Hierarchical text-conditional image generation with clip latents”
Aditya Ramesh et al · 2022
Later among the works it cites.
“High-resolution image synthesis with latent diffusion models”
Robin Rombach et al · 2022
Later among the works it cites.
“Palette: Image-to-image diffusion models”
Chitwan Saharia et al · 2022
Later among the works it cites.
“Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding”
Chitwan Saharia et al · 2022
Later among the works it cites.
“Human motion diffusion model”
Guy Tevet et al · 2022
Later among the works it cites.
“PhysDiff: Physics-Guided Human Motion Diffusion Model”
Ye Yuan et al · 2022
Later among the works it cites.
“Motiondiffuse: Text-driven human motion generation with diffusion model”
Mingyuan Zhang et al · 2022
Later among the works it cites.
“Reduce, Reuse, Recycle: Compositional Generation with Energy-Based Diffusion Models and MCMC”
Yilun Du et al · 2023
Closest in time.
“Action2motion: Conditioned generation of 3d human motions”
Chuan Guo et al · 2029
Closest in time.