Fetching the paper…
Reading the bibliography…
This study aims to improve the generation of 3D gestures by utilizing multimodal information from human speech.
Animated conversation: rule-based generation of facial expression, gesture & spoken intonation for multiple conversational agents
Cassell, J.; Pelachaud, C.; Badler, N.; Steedman, M.; Achorn, B.; Becket, T.; Douville, B.; Prevost, S.; and Stone, M. 1994 · 1994
Earlier work this paper cites.
Synthesizing multimodal utterances for conversational agents
Kopp, S.; and Wachsmuth, I. 2004 · 2004
Earlier work this paper cites.
Towards natural gesture synthesis: Evaluating gesture units in a data-driven approach to gesture synthesis
Kipp, M.; Neff, M.; Kipp, K. H.; and Albrecht, I. 2007 · 2007
Earlier work this paper cites.
Real-time prosody-driven synthesis of body language
Levine, S.; Theobalt, C.; and Koltun, V. 2009 · 2009
Earlier work this paper cites.
Gesture controllers
Levine, S.; Krähenbühl, P.; Thrun, S.; and Koltun, V. 2010 · 2010
Earlier work this paper cites.
Gesture and speech in interaction: An overview
Wagner, P.; Malisz, Z.; and Kopp, S. 2014 · 2014
Earlier work this paper cites.
Enriching word vectors with subword information
Bojanowski, P.; Grave, E.; Joulin, A.; and Mikolov, T. 2017 · 2017
Earlier work this paper cites.
Arbitrary style transfer in real-time with adaptive instance normalization
Huang, X.; and Belongie, S. 2017 · 2017
Earlier work this paper cites.
Learning individual styles of conversational gesture
Ginosar, S.; Bar, A.; Kohavi, G.; Chan, C.; Owens, A.; and Malik, J. 2019 · 2019
Earlier work this paper cites.
Robots learn social skills: End-to-end learning of co-speech gesture generation for humanoid robots
Yoon, Y.; Ko, W.-R.; Jang, M.; Lee, J.; Kim, J.; and Lee, G. 2019 · 2019
Earlier work this paper cites.
wav2vec 2.0: A framework for self-supervised learning of speech representations
Baevski, A.; Zhou, Y.; Mohamed, A.; and Auli, M. 2020 · 2020
Cited alongside, same era.
A simple framework for contrastive learning of visual representations
Chen, T.; Kornblith, S.; Norouzi, M.; and Hinton, G. 2020 · 2020
Cited alongside, same era.
A large, crowdsourced evaluation of gesture generation systems on common data: The GENEA Challenge 2020
Kucherenko, T.; Jonell, P.; Yoon, Y.; Wolfert, P.; and Henter, G. E. 2021 · 2020
Cited alongside, same era.
Speech gesture generation from the trimodal context of text, audio, and speaker identity
Yoon, Y.; Cha, B.; Lee, J.-H.; Jang, M.; Lee, J.; Kim, J.; and Lee, G. 2020 · 2020
Cited alongside, same era.
Speech drives templates: Co-speech gesture synthesis with learned templates
Qian, S.; Tu, Z.; Zhi, Y.; Liu, W.; and Gao, S. 2021 · 2021
Cited alongside, same era.
Gesture2Vec: Clustering gestures using representation learning methods for co-speech gesture generation
Yazdian, P. J.; Chen, M.; and Lim, A. 2022 · 2022
Later among the works it cites.
Generating Holistic 3D Human Motion from Speech
Yi, H.; Liang, H.; Liu, Y.; Cao, Q.; Wen, Y.; Bolkart, T.; Tao, D.; and Black, M. J. 2022 · 2022
Later among the works it cites.
GestureDiffuCLIP: Gesture Diffusion Model with CLIP Latents
Ao, T.; Zhang, Z.; and Liu, L. 2023 · 2023
Closest in time.
FineDance: A Fine-grained Choreography Dataset for 3D Full Body Dance Generation
Li, R.; Zhao, J.; Zhang, Y.; Su, M.; Ren, Z.; Zhang, H.; Tang, Y.; and Li, X. 2023 · 2023
Closest in time.
Consistent123: One image to highly consistent 3d asset using case-aware diffusion priors
Lin, Y.; Han, H.; Gong, C.; Xu, Z.; Zhang, Y.; and Li, X. 2023 · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Perturbed Self-Distillation: Weakly supervised large-scale point cloud semantic segmentation
Zhang, Y.; Qu, Y.; Xie, Y.; Li, Z.; Zheng, S.; and Li, C. 2021 · 2021
Cited alongside, same era.
An uncertainty analysis on finite difference time-domain computations with artificial neural networks: improving accuracy while maintaining low computational costs
Hu, R.; Monebhurrun, V.; Himeno, R.; Yokota, H.; and Costen, F. 2022 · 2022
Cited alongside, same era.
BEAT: A Large-Scale Semantic and Emotional Multi-Modal Dataset for Conversational Gestures Synthesis
Liu, H.; Zhu, Z.; Iwamoto, N.; Peng, Y.; Li, Z.; Zhou, Y.; Bozkurt, E.; and Zheng, B. 2022a · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Wei, J.; Wang, X.; Schuurmans, D.; Bosma, M.; Xia, F.; Chi, E.; Le, Q. V.; Zhou, D.; et al. 2022 · 2022
Cited alongside, same era.
The ReprGesture entry to the GENEA Challenge 2022
Yang, S.; Wu, Z.; Li, M.; Zhao, M.; Lin, J.; Chen, L.; and Bao, W. 2022 · 2022
Cited alongside, same era.
Audio2gestures: Generating diverse gestures from speech audio with conditional variational autoencoders
Li, J.; Kang, D.; Pei, W.; Zhe, X.; Zhang, Y.; He, Z.; and Bao, L. 2021a
Cited in the paper.
Ai choreographer: Music conditioned 3d dance generation with aist++
Li, R.; Yang, S.; Ross, D. A.; and Kanazawa, A. 2021b
Cited in the paper.
Closest in time.
A Comprehensive Review of Data-Driven Co-Speech Gesture Generation
Nyatsanga, S.; Kucherenko, T.; Ahuja, C.; Henter, G. E.; and Neff, M. 2023 · 2023
Closest in time.
EmotionGesture: Audio-Driven Diverse Emotional Co-Speech 3D Gesture Generation
Qi, X.; Liu, C.; Li, L.; Hou, J.; Xin, H.; and Yu, X. 2023 · 2023
Closest in time.
Bridging vision and language encoders: Parameter-efficient tuning for referring image segmentation
Xu, Z.; Chen, Z.; Zhang, Y.; Song, Y.; Wan, X.; and Li, G. 2023 · 2023
Closest in time.
The DiffuseStyleGesture+ entry to the GENEA Challenge 2023
Yang, S.; Xue, H.; Zhang, Z.; Li, M.; Wu, Z.; Wu, X.; Xu, S.; and Dai, Z. 2023d · 2023
Closest in time.
EMoG: Synthesizing Emotive Co-speech 3D Gesture with Diffusion Model
Yin, L.; Wang, Y.; He, T.; Liu, J.; Zhao, W.; Li, B.; Jin, X.; and Lin, J. 2023 · 2023
Closest in time.