Fetching the paper…
Reading the bibliography…
We present a GAN-based Transformer for general action-conditioned 3D human motion generation, including not only single-person actions but also multi-person interactive actions.
Efficient and robust annotation of motion capture data
Meinard Müller, Andreas Baak, and Hans-Peter Seidel · 2009
Earlier work this paper cites.
Microsoft kinect sensor and its effect
Zhengyou Zhang · 2012
Earlier work this paper cites.
Rectifier nonlinearities improve neural network acoustic models
Andrew L Maas, Awni Y Hannun, Andrew Y Ng, et al · 2013
Earlier work this paper cites.
Generative adversarial nets
Ian J. Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron C. Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Human3.6m: Large scale datasets and predictive methods for 3d human sensing in natural environments
Catalin Ionescu, Dragos Papava, Vlad Olaru, and Cristian Sminchisescu · 2014
Earlier work this paper cites.
Conditional generative adversarial nets
Mehdi Mirza and Simon Osindero · 2014
Earlier work this paper cites.
Pose-conditioned joint angle limits for 3d human pose reconstruction
Ijaz Akhter and Michael J. Black · 2015
Earlier work this paper cites.
Recurrent network models for human dynamics
Katerina Fragkiadaki, Sergey Levine, Panna Felsen, and Jitendra Malik · 2015
Earlier work this paper cites.
SMPL: a skinned multi-person linear model
Matthew Loper, Naureen Mahmood, Javier Romero, Gerard Pons-Moll, and Michael J. Black · 2015
Earlier work this paper cites.
The KIT whole-body human motion database
Christian Mandery, Ömer Terlemez, Martin Do, Nikolaus Vahrenkamp, and Tamim Asfour · 2015
Earlier work this paper cites.
Bridging nonlinearities and stochastic regularizers with gaussian error linear units
Dan Hendrycks and Kevin Gimpel · 2016
Earlier work this paper cites.
Structural-rnn: Deep learning on spatio-temporal graphs
Ashesh Jain, Amir Roshan Zamir, Silvio Savarese, and Ashutosh Saxena · 2016
Earlier work this paper cites.
Martín Arjovsky, Soumith Chintala, and Léon Bottou · 2017
Earlier work this paper cites.
Modulating early visual processing by language
Harm de Vries, Florian Strub, Jérémie Mary, Hugo Larochelle, Olivier Pietquin, and Aaron C. Courville · 2017
Earlier work this paper cites.
A recurrent variational autoencoder for human motion synthesis
Ikhsanul Habibie, Daniel Holden, Jonathan Schwarz, Joe Yearsley, and Taku Komura · 2017
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter · 2017
Earlier work this paper cites.
On human motion prediction using recurrent neural networks
Julieta Martinez, Michael J. Black, and Javier Romero · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Text2action: Generative adversarial synthesis from language to action
Hyemin Ahn, Timothy Ha, Yunho Choi, Hwiyeon Yoo, and Songhwai Oh · 2018
Earlier work this paper cites.
HP-GAN: probabilistic 3d human motion prediction via GAN
Emad Barsoum, John Kender, and Zicheng Liu · 2018
Earlier work this paper cites.
Learning to detect and track visible and occluded body joints in a virtual world
Matteo Fabbri, Fabio Lanzi, Simone Calderara, Andrea Palazzi, Roberto Vezzani, and Rita Cucchiara · 2018
Cited alongside, same era.
Social GAN: socially acceptable trajectories with generative adversarial networks
Agrim Gupta, Justin Johnson, Li Fei-Fei, Silvio Savarese, and Alexandre Alahi · 2018
Cited alongside, same era.
Generating animated videos of human activities from natural language descriptions
Angela S. Lin, Lemeng Wu, Rodolfo Corona, Kevin Tai, Qixing Huang, and Raymond J. Mooney · 2018
Cited alongside, same era.
cgans with projection discriminator
Takeru Miyato and Masanori Koyama · 2018
Cited alongside, same era.
Spatial temporal graph convolutional networks for skeleton-based action recognition
Sijie Yan, Yuanjun Xiong, and Dahua Lin · 2018
Cited alongside, same era.
MT-VAE: learning motion transformations to generate multimodal human dynamics
NTU RGB+D 120: A large-scale benchmark for 3d human activity understanding
Jun Liu, Amir Shahroudy, Mauricio Perez, Gang Wang, Ling-Yu Duan, and Alex C. Kot · 2020
Later among the works it cites.
Unicon: Universal neural controller for physics-based character motion
Tingwu Wang, Yunrong Guo, Maria Shugrina, and Sanja Fidler · 2020
Later among the works it cites.
A scalable approach to control diverse behaviors for physically simulated characters
Jungdam Won, Deepak Gopinath, and Jessica K. Hodgins · 2020
Later among the works it cites.
Perpetual motion: Generating unbounded human motion
Yan Zhang, Michael J. Black, and Siyu Tang · 2020
Later among the works it cites.
Tripod: Human trajectory and pose dynamics forecasting in the wild
Vida Adeli, Mahsa Ehsanpour, Ian D. Reid, Juan Carlos Niebles, Silvio Savarese, Ehsan Adeli, and Hamid Rezatofighi · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Xinchen Yan, Akash Rastogi, Ruben Villegas, Kalyan Sunkavalli, Eli Shechtman, Sunil Hadap, Ersin Yumer, and Honglak Lee · 2018
Cited alongside, same era.
Language2pose: Natural language grounded pose forecasting
Chaitanya Ahuja and Louis-Philippe Morency · 2019
Cited alongside, same era.
Structured prediction helps 3d human motion modelling
Emre Aksan, Manuel Kaufmann, and Otmar Hilliges · 2019
Cited alongside, same era.
Dancing to music
Hsin-Ying Lee, Xiaodong Yang, Ming-Yu Liu, Ting-Chun Wang, Yu-Ding Lu, Ming-Hsuan Yang, and Jan Kautz · 2019
Cited alongside, same era.
AMASS: archive of motion capture as surface shapes
Naureen Mahmood, Nima Ghorbani, Nikolaus F. Troje, Gerard Pons-Moll, and Michael J. Black · 2019
Cited alongside, same era.
Pytorch: An imperative style, high-performance deep learning library
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, Alban Desmaison, Andreas Köpf, Edward Z. Yang, Zachary DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala · 2019
Cited alongside, same era.
Neural state machine for character-scene interactions
Sebastian Starke, He Zhang, Taku Komura, and Jun Saito · 2019
Cited alongside, same era.
Later among the works it cites.
Combining transformer generators with convolutional discriminators
Ricard Durall, Stanislav Frolov, Andreas Dengel, and Janis Keuper · 2021
Later among the works it cites.
Synthesis of compositional animations from textual descriptions
Anindita Ghosh, Noshaba Cheema, Cennet Oguz, Christian Theobalt, and Philipp Slusallek · 2021
Later among the works it cites.
Dance revolution: Long-term dance generation with music via curriculum learning
Ruozi Huang, Huang Hu, Wei Wu, Kei Sawada, Mi Zhang, and Daxin Jiang · 2021
Later among the works it cites.
Transgan: Two transformers can make one strong GAN
Yifan Jiang, Shiyu Chang, and Zhangyang Wang · 2021
Later among the works it cites.
Vitgan: Training gans with vision transformers
Kwonjoon Lee, Huiwen Chang, Lu Jiang, Han Zhang, Zhuowen Tu, and Ce Liu · 2021
Later among the works it cites.
Learn to dance with AIST++: music conditioned 3d dance generation
Ruilong Li, Shan Yang, David A. Ross, and Angjoo Kanazawa · 2021
Later among the works it cites.
Action-conditioned 3D human motion synthesis with transformer VAE
Mathis Petrovich, Michael J. Black, and Gül Varol · 2021
Later among the works it cites.
BABEL: bodies, action and behavior with english labels
Abhinanda R. Punnakkal, Arjun Chandrasekaran, Nikos Athanasiou, Alejandra Quiros-Ramirez, and Michael J. Black · 2021
Later among the works it cites.
A scalable approach to predict multi-agent motion for human-robot collaboration
Mohammad Samin Yasar and Tariq Iqbal · 2021
Later among the works it cites.
Teach: Temporal action composition for 3d humans
Nikos Athanasiou, Mathis Petrovich, Michael J Black, and Gül Varol · 2022
Closest in time.
Flame: Free-form language-based motion synthesis & editing
Jihoon Kim, Jiseob Kim, and Sungjoon Choi · 2022
Closest in time.
Temos: Generating diverse human motions from textual descriptions
Mathis Petrovich, Michael J Black, and Gül Varol · 2022
Closest in time.
Guy Tevet, Sigal Raab, Brian Gordon, Yonatan Shafir, Daniel Cohen-Or, and Amit H Bermano · 2022
Closest in time.
Motiondiffuse: Text-driven human motion generation with diffusion model
Mingyuan Zhang, Zhongang Cai, Liang Pan, Fangzhou Hong, Xinying Guo, Lei Yang, and Ziwei Liu · 2022
Closest in time.