Fetching the paper…
Reading the bibliography…
We present AIST++, a new multi-modal dataset of 3D dance motion and music, along with FACT, a Full-Attention Cross-modal Transformer network for generating 3D dance motion conditioned on music.
A hierarchical approach to interactive motion editing for human-like figures
Jehee Lee and Sung Yong Shin · 1999
Earlier work this paper cites.
Learning statistical models of human motion
Richard Bowden · 2000
Earlier work this paper cites.
Style machines
Matthew Brand and Aaron Hertzmann · 2000
Earlier work this paper cites.
Animating by multi-level sampling
Katherine Pullen and Christoph Bregler · 2000
Earlier work this paper cites.
Learning variable-length markov models of behavior
Aphrodite Galata, Neil Johnson, and David Hogg · 2001
Earlier work this paper cites.
Interactive motion generation from examples
Okan Arikan and David A Forsyth · 2002
Earlier work this paper cites.
Dance music, movement and tempo preferences
Dirk Moelants · 2003
Earlier work this paper cites.
Efficient content-based retrieval of motion capture data
Meinard Müller, Tido Röder, and Michael Clausen · 2005
Earlier work this paper cites.
Dancing-to-music character animation
Takaaki Shiratori, Atsushi Nakazawa, and Katsushi Ikeuchi · 2006
Earlier work this paper cites.
Parametric motion graphs
Rachel Heck and Michael Gleicher · 2007
Earlier work this paper cites.
Motion graphs
Lucas Kovar, Michael Gleicher, and Frédéric Pighin · 2008
Earlier work this paper cites.
Fmdistance: A fast and effective distance function for motion capture data
Kensuke Onuma, Christos Faloutsos, and Jessica K Hodgins · 2008
Earlier work this paper cites.
Factored conditional restricted boltzmann machines for modeling motion style
Graham W Taylor and Geoffrey E Hinton · 2009
Earlier work this paper cites.
Example-based automatic music-driven conventional dance motion synthesis
Rukun Fan, Songhua Xu, and Weidong Geng · 2011
Earlier work this paper cites.
2d human pose estimation: New benchmark and state of the art analysis
Mykhaylo Andriluka, Leonid Pishchulin, Peter Gehler, and Bernt Schiele · 2014
Earlier work this paper cites.
Human3.6m: Large scale datasets and predictive methods for 3d human sensing in natural environments
Catalin Ionescu, Dragos Papava, Vlad Olaru, and Cristian Sminchisescu · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Translating videos to natural language using deep recurrent neural networks
Subhashini Venugopalan, Huijuan Xu, Jeff Donahue, Marcus Rohrbach, Raymond Mooney, and Kate Saenko · 2014
Earlier work this paper cites.
Scheduled sampling for sequence prediction with recurrent neural networks
Samy Bengio, Oriol Vinyals, Navdeep Jaitly, and Noam Shazeer · 2015
Earlier work this paper cites.
Recurrent network models for human dynamics
Katerina Fragkiadaki, Sergey Levine, Panna Felsen, and Jitendra Malik · 2015
Earlier work this paper cites.
Learning motion manifolds with convolutional autoencoders
Daniel Holden, Jun Saito, Taku Komura, and Thomas Joyce · 2015
Earlier work this paper cites.
Deep visual-semantic alignments for generating image descriptions
Andrej Karpathy and Li Fei-Fei · 2015
Earlier work this paper cites.
SMPL: A skinned multi-person linear model
Matthew Loper, Naureen Mahmood, Javier Romero, Gerard Pons-Moll, and Michael J. Black · 2015
Earlier work this paper cites.
librosa: Audio and music signal analysis in python
Brian McFee, Colin Raffel, Dawen Liang, Daniel PW Ellis, Matt McVicar, Eric Battenberg, and Oriol Nieto · 2015
Earlier work this paper cites.
Sequence to sequence-video to text
Subhashini Venugopalan, Marcus Rohrbach, Jeffrey Donahue, Raymond Mooney, Trevor Darrell, and Kate Saenko · 2015
Earlier work this paper cites.
Describing videos by exploiting temporal structure
Li Yao, Atousa Torabi, Kyunghyun Cho, Nicolas Ballas, Christopher Pal, Hugo Larochelle, and Aaron Courville · 2015
Earlier work this paper cites.
A deep learning framework for character motion synthesis and editing
Daniel Holden, Jun Saito, and Taku Komura · 2016
Earlier work this paper cites.
Structural-rnn: Deep learning on spatio-temporal graphs
Ashesh Jain, Amir R Zamir, Silvio Savarese, and Ashutosh Saxena · 2016
Earlier work this paper cites.
Video paragraph captioning using hierarchical recurrent neural networks
Haonan Yu, Jiang Wang, Zhiheng Huang, Yi Yang, and Wei Xu · 2016
Earlier work this paper cites.
Groovenet: Real-time music-driven dance movement generation using artificial neural networks
Omid Alemi, Jules Françoise, and Philippe Pasquier · 2017
Earlier work this paper cites.
Deep representation learning for human motion prediction and classification
Judith Bütepage, Michael J Black, Danica Kragic, and Hedvig Kjellström · 2017
Earlier work this paper cites.
Learning human motion models for long-term predictions
Partha Ghosh, Jie Song, Emre Aksan, and Otmar Hilliges · 2017
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter · 2017
Cited alongside, same era.
Phase-functioned neural networks for character control
Daniel Holden, Taku Komura, and Jun Saito · 2017
Cited alongside, same era.
Dense-captioning events in videos
Ranjay Krishna, Kenji Hata, Frederic Ren, Li Fei-Fei, and Juan Carlos Niebles · 2017
Cited alongside, same era.
Benchmarking and error diagnosis in multi-instance pose estimation
Matteo Ruggero Ronchi and Pietro Perona · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
OpenPose: realtime multi-person 2D pose estimation using Part Affinity Fields
Cross view fusion for 3d human pose estimation
Haibo Qiu, Chunyu Wang, Jingdong Wang, Naiyan Wang, and Wenjun Zeng · 2019
Later among the works it cites.
Music-oriented dance video synthesis with pose perceptual loss
Xuanchi Ren, Haoran Li, Zijian Huang, and Qifeng Chen · 2019
Later among the works it cites.
Fastspeech: Fast, robust and controllable text to speech
Yi Ren, Yangjun Ruan, Xu Tan, Tao Qin, Sheng Zhao, Zhou Zhao, and Tie-Yan Liu · 2019
Later among the works it cites.
Neural state machine for character-scene interactions
Sebastian Starke, He Zhang, Taku Komura, and Jun Saito · 2019
Later among the works it cites.
Learning video representations using contrastive bidirectional transformer
Chen Sun, Fabien Baradel, Kevin Murphy, and Cordelia Schmid · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhe Cao, Gines Hidalgo, Tomas Simon, Shih-En Wei, and Yaser Sheikh · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Cited alongside, same era.
Weakly supervised dense event captioning in videos
Xuguang Duan, Wenbing Huang, Chuang Gan, Jingdong Wang, Wenwu Zhu, and Junzhou Huang · 2018
Cited alongside, same era.
Investigating the use of recurrent motion modelling for speech gesture generation
Ylva Ferstl and Rachel McDonnell · 2018
Cited alongside, same era.
Deep inertial poser learning to reconstruct human pose from sparseinertial measurements in real time
Yinghao Huang, Manuel Kaufmann, Emre Aksan, Michael J. Black, Otmar Hilliges, and Gerard Pons-Moll · 2018
Cited alongside, same era.
Listen to dance: Music-driven choreography generation using autoregressive encoder-decoder network
Juheon Lee, Seohyun Kim, and Kyogu Lee · 2018
Cited alongside, same era.
Auto-conditioned recurrent networks for extended complex human motion synthesis
Zimo Li, Yi Zhou, Shuangjiu Xiao, Chong He, Zeng Huang, and Hao Li · 2018
Cited alongside, same era.
Videobert: A joint model for video and language representation learning
Chen Sun, Austin Myers, Carl Vondrick, Kevin Murphy, and Cordelia Schmid · 2019
Later among the works it cites.
Aist dance video database: Multi-genre, multi-dancer, and multi-camera database for dance information processing
Shuhei Tsuchida, Satoru Fukayama, Masahiro Hamasaki, and Masataka Goto · 2019
Later among the works it cites.
Lpcnet: Improving neural speech synthesis through linear prediction
Jean-Marc Valin and Jan Skoglund · 2019
Later among the works it cites.
Imitation learning for human pose prediction
Borui Wang, Ehsan Adeli, Hsu-kuang Chiu, De-An Huang, and Juan Carlos Niebles · 2019
Later among the works it cites.
Weakly-supervised deep recurrent neural networks for basic dance step generation
Nelson Yalta, Shinji Watanabe, Kazuhiro Nakadai, and Tetsuya Ogata · 2019
Later among the works it cites.
Generative autoregressive networks for 3d dancing move synthesis from music
Hyemin Ahn, Jaehun Kim, Kihyun Kim, and Songhwai Oh · 2020
Later among the works it cites.
Attention, please: A spatio-temporal transformer for 3d human motion prediction
Emre Aksan, Peng Cao, Manuel Kaufmann, and Otmar Hilliges · 2020
Later among the works it cites.
Adversarial gesture generation with realistic gesture phasing
Ylva Ferstl, Michael Neff, and Rachel McDonnell · 2020
Later among the works it cites.
Foley music: Learning to generate music from videos
Chuang Gan, Deng Huang, Peihao Chen, Joshua B Tenenbaum, and Antonio Torralba · 2020
Later among the works it cites.
fairmotion - tools to load, process and visualize motion capture data
Deepak Gopinath and Jungdam Won · 2020
Later among the works it cites.
Multi-modal dense video captioning
Vladimir Iashin and Esa Rahtu · 2020
Later among the works it cites.
Temporally guided music-to-body-movement generation
Hsuan-Kai Kao and Li Su · 2020
Later among the works it cites.
Cross-conditioned recurrent networks for long-term synthesis of inter-person human motion interactions
Jogendra Nath Kundu, Himanshu Buckchash, Priyanka Mandikal, Anirudh Jamkhandi, Venkatesh Babu RADHAKRISHNAN, et al · 2020
Later among the works it cites.
Learning to generate diverse dance motions with transformer
Jiaman Li, Yihang Yin, Hang Chu, Yi Zhou, Tingwu Wang, Sanja Fidler, and Hao Li · 2020
Later among the works it cites.
Self-supervised dance video synthesis conditioned on music
Xuanchi Ren, Haoran Li, Zijian Huang, and Qifeng Chen · 2020
Later among the works it cites.
Local motion phases for learning multi-contact character movements
Sebastian Starke, Yiwei Zhao, Taku Komura, and Kazi Zaman · 2020
Later among the works it cites.
https://www.statista.com/statistics/249396/top-youtube-videos-views/ , 2020
Statista · 2020
Later among the works it cites.
Deepdance: Music-to-dance motion choreography with adversarial learning
Guofei Sun, Yongkang Wong, Zhiyong Cheng, Mohan S Kankanhalli, Weidong Geng, and Xiangdong Li · 2020
Later among the works it cites.
Feel the music: Automatically generating a dance for an input song
Purva Tendulkar, Abhishek Das, Aniruddha Kembhavi, and Devi Parikh · 2020
Later among the works it cites.
Cross-modal relation-aware networks for audio-visual event localization
Haoming Xu, Runhao Zeng, Qingyao Wu, Mingkui Tan, and Chuang Gan · 2020
Later among the works it cites.
Choreonet: Towards music to dance synthesis with choreographic action unit
Zijie Ye, Haozhe Wu, Jia Jia, Yaohua Bu, Wei Chen, Fanbo Meng, and Yanfeng Wang · 2020
Later among the works it cites.
Music2dance: Music-driven dance generation using wavenet
Wenlin Zhuang, Congyi Wang, Siyu Xia, Jinxiang Chai, and Yangang Wang · 2020
Later among the works it cites.
Towards 3d dance motion synthesis and control
Wenlin Zhuang, Yangang Wang, Joseph Robinson, Congyi Wang, Ming Shao, Yun Fu, and Siyu Xia · 2020
Later among the works it cites.
Text2gestures: A transformer-based network for generating emotive body gestures for virtual agents
Uttaran Bhattacharya, Nicholas Rewkowski, Abhishek Banerjee, Pooja Guhan, Aniket Bera, and Dinesh Manocha · 2021
Closest in time.
Choreomaster: choreography-oriented music-driven dance synthesis
Kang Chen, Zhipeng Tan, Jin Lei, Song-Hai Zhang, Yuan-Chen Guo, Weidong Zhang, and Shi-Min Hu · 2021
Closest in time.
Dance revolution: Long-term dance generation with music via curriculum learning
Ruozi Huang, Huang Hu, Wei Wu, Kei Sawada, Mi Zhang, and Daxin Jiang · 2021
Closest in time.