Fetching the paper…
Reading the bibliography…
We present a learning-based approach with pose perceptual loss for automatic music video generation.
Audio-visual integration in multimodal communication
Chen T. and Rao R.R · 1998
Earlier work this paper cites.
Image quality assessment: from error visibility to structural similarity
Zhou Wang, Alan C. Bovik, Hamid R. Sheikh, and Eero P. Simoncelli · 2004
Earlier work this paper cites.
Making them dance
Jae Woo Kim, Hesham Fouad, and James K. Hahn · 2006
Earlier work this paper cites.
Dancing-to-music character animation
Takaaki Shiratori, Atsushi Nakazawa, and Katsushi Ikeuchi · 2006
Earlier work this paper cites.
HMDB: a large video database for human motion recognition
H. Kuehne, H. Jhuang, E. Garrote, T. Poggio, and T. Serre · 2011
Earlier work this paper cites.
No-reference image quality assessment in the spatial domain
Anish Mittal, Anush Krishna Moorthy, and Alan Conrad Bovik · 2012
Earlier work this paper cites.
UCF101: A dataset of 101 human actions classes from videos in the wild
Khurram Soomro, Amir Roshan Zamir, and Mubarak Shah · 2012
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman · 2015
Earlier work this paper cites.
Super-resolution with deep convolutional sufficient statistics
Joan Bruna, Pablo Sprechmann, and Yann LeCun · 2016
Earlier work this paper cites.
Generating images with perceptual similarity metrics based on deep networks
Alexey Dosovitskiy and Thomas Brox · 2016
Earlier work this paper cites.
Perceptual losses for real-time style transfer and super-resolution
Justin Johnson, Alexandre Alahi, and Li Fei-Fei · 2016
Earlier work this paper cites.
Deep multi-scale video prediction beyond mean square error
Michaël Mathieu, Camille Couprie, and Yann LeCun · 2016
Earlier work this paper cites.
Synthesizing the preferred inputs for neurons in neural networks via deep generator networks
Anh Mai Nguyen, Alexey Dosovitskiy, Jason Yosinski, Thomas Brox, and Jeff Clune · 2016
Earlier work this paper cites.
NTU RGB+D: A large scale dataset for 3d human activity analysis
Amir Shahroudy, Jun Liu, Tian-Tsong Ng, and Gang Wang · 2016
Earlier work this paper cites.
Generating videos with scene dynamics
Carl Vondrick, Hamed Pirsiavash, and Antonio Torralba · 2016
Cited alongside, same era.
Convolutional pose machines
Shih-En Wei, Varun Ramakrishna, Takeo Kanade, and Yaser Sheikh · 2016
Cited alongside, same era.
Groovenet: Real-time music-driven dance movement generation using artificial neural networks
Omid Alemi, Jules Françoise, and Philippe Pasquier · 2017
Cited alongside, same era.
Realtime multi-person 2d pose estimation using part affinity fields
Zhe Cao, Tomas Simon, Shih-En Wei, and Yaser Sheikh · 2017
Cited alongside, same era.
Photographic image synthesis with cascaded refinement networks
Qifeng Chen and Vladlen Koltun · 2017
Cited alongside, same era.
Convolutional recurrent neural networks for music classification
Keunwoo Choi, György Fazekas, Mark B. Sandler, and Kyunghyun Cho · 2017
Cited alongside, same era.
OpenPose: realtime multi-person 2D pose estimation using Part Affinity Fields
Zhe Cao, Gines Hidalgo, Tomas Simon, Shih-En Wei, and Yaser Sheikh · 2018
Later among the works it cites.
Let’s dance: Learning from online dance videos
Daniel Castro, Steven Hickson, Patsorn Sangkloy, Bhavishya Mittal, Sean Dai, James Hays, and Irfan Essa · 2018
Later among the works it cites.
Human motion analysis with deep metric learning
Huseyin Coskun, David Joseph Tan, Sailesh Conjeti, Nassir Navab, and Federico Tombari · 2018
Later among the works it cites.
Listen to dance: Music-driven choreography generation using autoregressive encoder-decoder network
Juheon Lee, Seohyun Kim, and Kyogu Lee · 2018
Later among the works it cites.
Co-occurrence feature learning from skeleton data for action recognition and detection with hierarchical aggregation
Chao Li, Qiaoyong Zhong, Di Xie, and Shiliang Pu · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Improved training of wasserstein gans
Ishaan Gulrajani, Faruk Ahmed, Martín Arjovsky, Vincent Dumoulin, and Aaron C. Courville · 2017
Cited alongside, same era.
Image-to-image translation with conditional adversarial networks
Phillip Isola, Jun-Yan Zhu, Tinghui Zhou, and Alexei A Efros · 2017
Cited alongside, same era.
Video generation from text
Yitong Li, Martin Renqiang Min, Dinghan Shen, David E. Carlson, and Lawrence Carin · 2017
Cited alongside, same era.
A structured self-attentive sentence embedding
Zhouhan Lin, Minwei Feng, Cícero Nogueira dos Santos, Mo Yu, Bing Xiang, Bowen Zhou, and Yoshua Bengio · 2017
Cited alongside, same era.
On human motion prediction using recurrent neural networks
Julieta Martinez, Michael J. Black, and Javier Romero · 2017
Cited alongside, same era.
Automatic differentiation in pytorch
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer · 2017
Cited alongside, same era.
Later among the works it cites.
Dance with melody: An lstm-autoencoder approach to music-oriented dance synthesis
Taoran Tang, Jia Jia, and Hanyang Mao · 2018
Later among the works it cites.
MoCoGAN: Decomposing motion and content for video generation
Sergey Tulyakov, Ming-Yu Liu, Xiaodong Yang, and Jan Kautz · 2018
Later among the works it cites.
End-to-end speech-driven facial animation with temporal gans
Konstantinos Vougioukas, Stavros Petridis, and Maja Pantic · 2018
Later among the works it cites.
High-resolution image synthesis and semantic manipulation with conditional gans
Ting-Chun Wang, Ming-Yu Liu, Jun-Yan Zhu, Andrew Tao, Jan Kautz, and Bryan Catanzaro · 2018
Later among the works it cites.
Video-to-video synthesis
Ting-Chun Wang, Ming-Yu Liu, Jun-Yan Zhu, Nikolai Yakovenko, Andrew Tao, Jan Kautz, and Bryan Catanzaro · 2018
Later among the works it cites.
Spatial temporal graph convolutional networks for skeleton-based action recognition
Sijie Yan, Yuanjun Xiong, and Dahua Lin · 2018
Later among the works it cites.
The unreasonable effectiveness of deep features as a perceptual metric
Richard Zhang, Phillip Isola, Alexei A. Efros, Eli Shechtman, and Oliver Wang · 2018
Later among the works it cites.
Everybody dance now
Caroline Chan, Shiry Ginosar, Tinghui Zhou, and Alexei A Efros · 2019
Closest in time.
Liquid warping GAN: A unified framework for human motion imitation, appearance transfer and novel view synthesis
Wen Liu, Zhixin Piao, Jie Min, Wenhan Luo, Lin Ma, and Shenghua Gao · 2019
Closest in time.
Dance dance generation: Motion transfer for internet videos
Yipin Zhou, Zhaowen Wang, Chen Fang, Trung Bui, and Tamara L. Berg · 2019
Closest in time.