Fetching the paper…
Reading the bibliography…
Generating conversational gestures from speech audio is challenging due to the inherent one-to-many mapping between audio and body motions.
A scale for the measurement of the psychological magnitude pitch
Stanley Smith Stevens, John Volkmann, and Edwin B Newman · 1937
Earlier work this paper cites.
Real-Time Prosody-Driven Synthesis of Body Language
Sergey Levine, Christian Theobalt, and Vladlen Koltun · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Xavier Glorot and Yoshua Bengio · 2010
Earlier work this paper cites.
Gesture controllers
Sergey Levine, Philipp Krähenbühl, Sebastian Thrun, and Vladlen Koltun · 2010
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Auto-encoding variational bayes
Diederik P Kingma and Max Welling · 2014
Earlier work this paper cites.
Recurrent network models for human dynamics
Katerina Fragkiadaki, Sergey Levine, Panna Felsen, and Jitendra Malik · 2015
Earlier work this paper cites.
Infogan: Interpretable representation learning by information maximizing generative adversarial nets
Xi Chen, Yan Duan, Rein Houthooft, John Schulman, Ilya Sutskever, and Pieter Abbeel · 2016
Earlier work this paper cites.
Tutorial on variational autoencoders
CARL DOERSCH · 2016
Earlier work this paper cites.
Jeff Donahue, Philipp Krähenbühl, and Trevor Darrell · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Structural-rnn: Deep learning on spatio-temporal graphs
Ashesh Jain, Amir R Zamir, Silvio Savarese, and Ashutosh Saxena · 2016
Earlier work this paper cites.
Autoencoding beyond pixels using a learned similarity metric
Anders Boesen Lindbo Larsen, Søren Kaae Sønderby, Hugo Larochelle, and Ole Winther · 2016
Earlier work this paper cites.
Realtime multi-person 2d pose estimation using part affinity fields
Zhe Cao, Tomas Simon, Shih-En Wei, and Yaser Sheikh · 2017
Earlier work this paper cites.
On human motion prediction using recurrent neural networks
Julieta Martinez, Michael J Black, and Javier Romero · 2017
Cited alongside, same era.
Automatic differentiation in pytorch
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer · 2017
Cited alongside, same era.
Toward multimodal image-to-image translation
Jun-Yan Zhu, Richard Zhang, Deepak Pathak, Trevor Darrell, Alexei A Efros, Oliver Wang, and Eli Shechtman · 2017
Cited alongside, same era.
Deep learning using rectified linear units (relu)
Abien Fred Agarap · 2018
Cited alongside, same era.
An empirical evaluation of generic convolutional and recurrent networks for sequence modeling
Shaojie Bai, J Zico Kolter, and Vladlen Koltun · 2018
Cited alongside, same era.
Human motion prediction via spatio-temporal inpainting
Alejandro Hernandez, Jurgen Gall, and Francesc Moreno-Noguer · 2019
Later among the works it cites.
Learning to reconstruct 3D human pose and shape via model-fitting in the loop
Nikos Kolotouros, Georgios Pavlakos, Michael Black, and Kostas Daniilidis · 2019
Later among the works it cites.
Mode seeking generative adversarial networks for diverse image synthesis
Qi Mao, Hsin-Ying Lee, Hung-Yu Tseng, Siwei Ma, and Ming-Hsuan Yang · 2019
Later among the works it cites.
Expressive body capture: 3d hands, face, and body from a single image
Georgios Pavlakos, Vasileios Choutas, Nima Ghorbani, Timo Bolkart, Ahmed A. A. Osman, Dimitrios Tzionas, and Michael J. Black · 2019
Later among the works it cites.
Predicting 3D human dynamics from video
Jason Zhang, Panna Felsen, Angjoo Kanazawa, and Jitendra Malik · 2019
Later among the works it cites.
On the continuity of rotation representations in neural networks
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Iva: Investigating the use of recurrent motion modelling for speech gesture generation
Ylva Ferstl and Rachel McDonnell · 2018
Cited alongside, same era.
Multimodal unsupervised image-to-image translation
Xun Huang, Ming-Yu Liu, Serge Belongie, and Jan Kautz · 2018
Cited alongside, same era.
End-to-End Recovery of Human Shape and Pose
Angjoo Kanazawa, Michael J. Black, David W. Jacobs, and Jitendra Malik · 2018
Cited alongside, same era.
Glow: Generative flow with invertible 1x1 convolutions
Durk P Kingma and Prafulla Dhariwal · 2018
Cited alongside, same era.
Quaternet: A quaternion-based recurrent model for human motion
Dario Pavllo, David Grangier, and Michael Auli · 2018
Cited alongside, same era.
Audio to body dynamics
Eli Shlizerman, Lucio Dery, Hayden Schoen, and Ira Kemelmacher-Shlizerman · 2018
Cited alongside, same era.
Mocogan: Decomposing motion and content for video generation
Sergey Tulyakov, Ming-Yu Liu, Xiaodong Yang, and Jan Kautz · 2018
Cited alongside, same era.
Yi Zhou, Connelly Barnes, Lu Jingwan, Yang Jimei, and Li Hao · 2019
Later among the works it cites.
Style-controllable speech-driven gesture synthesis using normalising flows
Simon Alexanderson, Gustav Eje Henter, Taras Kucherenko, and Jonas Beskow · 2020
Later among the works it cites.
Stargan v2: Diverse image synthesis for multiple domains
Yunjey Choi, Youngjung Uh, Jaejun Yoo, and Jung-Woo Ha · 2020
Later among the works it cites.
Moglow: Probabilistic and controllable motion synthesis using normalising flows
Gustav Eje Henter, Simon Alexanderson, and Jonas Beskow · 2020
Later among the works it cites.
Normalizing flows: An introduction and review of current methods
Ivan Kobyzev, Simon Prince, and Marcus Brubaker · 2020
Later among the works it cites.
Character controllers using motion vaes
Hung Yu Ling, Fabio Zinno, George Cheng, and Michiel Van De Panne · 2020
Later among the works it cites.
librosa/librosa: 0.8.0, July 2020
Brian McFee, Vincent Lostanlen, Alexandros Metsai, Matt McVicar, Stefan Balke, Carl Thomé, Colin Raffel, Frank Zalkow, Ayoub Malek, Dana, Kyungyun Lee, Oriol Nieto, Jack Mason, Dan Ellis, Eric Battenberg, Scott Seyfarth, Ryuichi Yamamoto, Keunwoo Choi, viktorandreevichmorozov, Josh Moore, Rachel Bittner, Shunsuke Hidaka, Ziyao Wei, nullmightybofo, Darío Hereñú, Fabian-Robert Stöter, Pius Friesch, Adam Weiss, Matt Vollrath, and Taewoon Kim · 2020
Later among the works it cites.
S3vae: Self-supervised sequential vae for representation disentanglement and data generation
Yizhe Zhu, Martin Renqiang Min, Asim Kadav, and Hans Peter Graf · 2020
Later among the works it cites.
https://www.mixamo.com
Mixamo · 2021
Closest in time.