Fetching the paper…
Reading the bibliography…
This paper reports on the second GENEA Challenge to benchmark data-driven automatic co-speech gesture generation.
Processing language in face-to-face conversation: Questions with gestures get faster responses
Judith Holler, Kobin H. Kendrick, and Stephen C. Levinson. 2018 · 1908
Earlier work this paper cites.
A new test for 2 × \times 2 tables
George Alfred Barnard. 1945 · 1945
Earlier work this paper cites.
Estimates of the regression coefficient based on Kendall’s tau
Pranab Kumar Sen. 1968 · 1968
Earlier work this paper cites.
Rank Correlation Methods (4 ed.)
Maurice G. Kendall. 1970 · 1970
Earlier work this paper cites.
Hearing lips and seeing voices
Harry McGurk and John MacDonald. 1976 · 1976
Earlier work this paper cites.
A simple sequentially rejective multiple test procedure
Sture Holm. 1979 · 1979
Earlier work this paper cites.
Comparison of parametric representations for monosyllabic word recognition in continuously spoken sentences
Steven B. Davis and Paul Mermelstein. 1980 · 1980
Earlier work this paper cites.
Spatial control of arm movements
Pietro Morasso. 1981 · 1981
Earlier work this paper cites.
Canonical Correlation Analysis: Uses and Interpretation . Vol. 47
Bruce Thompson. 1984 · 1984
Earlier work this paper cites.
Formation and control of optimal trajectory in human multijoint arm movement
Yoji Uno, Mitsuo Kawato, and Rika Suzuki. 1989 · 1989
Earlier work this paper cites.
Statistical Intervals: A Guide for Practitioners . Vol. 92
Gerald J. Hahn and William Q. Meeker. 1991 · 1991
Earlier work this paper cites.
Hand and Mind: What Gestures Reveal about Thought
David McNeill. 1992 · 1992
Earlier work this paper cites.
A rank-invariant method of linear and polynomial regression analysis
Henri Theil. 1992 · 1992
Earlier work this paper cites.
Methods for subjective determination of transmission quality
International Telecommunication Union, Telecommunication Standardisation Sector. 1996 · 1996
Earlier work this paper cites.
Practical parameterization of rotations using the exponential map
F. Sebastian Grassia. 1998 · 1998
Earlier work this paper cites.
BEAT: The behavior expression animation toolkit. In Proceedings of the Annual Conference on Computer Graphics and Interactive Techniques (SIGGRAPH ’01) . 477–486
Justine Cassell, Hannes Högni Vilhjálmsson, and Timothy Bickmore. 2001 · 2001
Earlier work this paper cites.
Interactive motion generation from examples
Okan Arikan and David A. Forsyth. 2002 · 2002
Earlier work this paper cites.
Motion graphs
Lucas Kovar, Michael Gleicher, and Frédéric Pighin. 2002 · 2002
Earlier work this paper cites.
Interactive control of avatars animated with human motion data
Jehee Lee, Jinxiang Chai, Paul S. A. Reitsma, Jessica K. Hodgins, and Nancy S. Pollard. 2002 · 2002
Earlier work this paper cites.
Perception and comprehension of synthetic speech
Stephen J. Winters and David B. Pisoni. 2004 · 2004
Earlier work this paper cites.
The Blizzard Challenge – 2005: Evaluating corpus-based speech synthesis on common datasets. In Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech ’05) . 77–80
Alan W. Black and Keiichi Tokuda. 2005 · 2005
Earlier work this paper cites.
GNetIc – Using Bayesian decision networks for iconic gesture generation. In Proceedings of the International Conference on Intelligent Virtual Agents (IVA ’09) . 76–89
Kirsten Bergmann and Stefan Kopp. 2009 · 2009
Earlier work this paper cites.
Recommended tests for association in 2 × \times 2 tables
Stian Lydersen, Morten W. Fagerland, and Petter Laake. 2009 · 2009
Earlier work this paper cites.
SynFace—Speech-driven facial animation for virtual speech-reading support
Giampiero Salvi, Jonas Beskow, Samer Al Moubayed, and Björn Granström. 2009 · 2009
Earlier work this paper cites.
Individualized gesturing outperforms average gesturing – Evaluating gesture production in virtual humans. In Proceedings of the International Conference on Intelligent Virtual Agents (ICA ’10) . 104–117
Kirsten Bergmann, Stefan Kopp, and Friederike Eyssel. 2010 · 2010
Earlier work this paper cites.
Gesture controllers
Sergey Levine, Philipp Krähenbühl, Sebastian Thrun, and Vladlen Koltun. 2010 · 2010
Earlier work this paper cites.
Comparison of approaches for instrumentally predicting the quality of text-to-speech systems. In Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech ’10) . 1325–1328
Sebastian Möller, Florian Hinterleitner, Tiago H. Falk, and Tim Polzehl. 2010 · 2010
Earlier work this paper cites.
The relation of speech and gestures: Temporal synchrony follows semantic synchrony. In Proceedings of the Workshop on Gesture and Speech in Interaction (GeSpIn ’11)
Kirsten Bergmann, Volkan Aksu, and Stefan Kopp. 2011 · 2011
Earlier work this paper cites.
A friendly gesture: Investigating the effect of multimodal robot behavior in human-robot interaction. In Proceedings of the IEEE International Symposium on Robot and Human Interactive Communication (RO-MAN ’11) . 247–252
Maha Salem, Katharina Rohlfing, Stefan Kopp, and Frank Joublin. 2011 · 2011
Earlier work this paper cites.
Evaluating an expressive gesture model for a humanoid robot: Experimental results
Quoc Anh Le and Catherine Pelachaud. 2012 · 2012
Earlier work this paper cites.
Generation and evaluation of communicative robot gesture
Maha Salem, Stefan Kopp, Ipke Wachsmuth, Katharina Rohlfing, and Frank Joublin. 2012 · 2012
Earlier work this paper cites.
Evaluating expressive speech synthesis from audiobooks in conversational phrases. In Proceedings of the International Conference on Language Resources and Evaluation (LREC ’12) . 3335–3339
Éva Székely, João P. Cabral, Mohamed Abou-Zleikha, Peter Cahill, and Julie Carson-Berndsen. 2012 · 2012
Earlier work this paper cites.
Expressive speech synthesis in MARY TTS using audiobook data and EmotionML. In Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech ’13) . 1564–1568
Marcela Charfuelan and Ingmar Steiner. 2013 · 2013
Earlier work this paper cites.
To err is human(-like): Effects of robot gesture on perceived anthropomorphism and likability
Maha Salem, Friederike Eyssel, Katharina Rohlfing, Stefan Kopp, and Frank Joublin. 2013 · 2013
Earlier work this paper cites.
Gesture and speech in interaction: An overview
Petra Wagner, Zofia Malisz, and Stefan Kopp. 2014 · 2013
Earlier work this paper cites.
Measuring a decade of progress in text-to-speech
Simon King. 2014 · 2014
Earlier work this paper cites.
GloVe: Global vectors for word representation. In Proceedings of the Conference on Empirical Methods in Natural Language Processing (EMNLP ’14) . 1532–1543
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Earlier work this paper cites.
Expectations and speech intelligibility
Molly Babel and Jamie Russell. 2015 · 2015
Earlier work this paper cites.
Affect-expressive hand gestures synthesis and animation. In Proceedings of the International Conference on Multimedia and Expo (ICME ’15) . 1–6
Elif Bozkurt, Engin Erzin, and Yücel Yemez. 2015 · 2015
Earlier work this paper cites.
Motion matching – the road to next gen animation. In Proceedings of Nucl.ai
Michael Büttner and Simon Clavet. 2015 · 2015
Earlier work this paper cites.
Predicting co-verbal gestures: A deep and temporal modeling approach. In Proceedings of the International Conference on Intelligent Virtual Agents (IVA ’15) . 152–166
Chung-Cheng Chiu, Louis-Philippe Morency, and Stacy Marsella. 2015 · 2015
Earlier work this paper cites.
A perceptual investigation of wavelet-based decomposition of f f 0 for text-to-speech synthesis. In Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech ’15) . 1586–1590
Manuel Sam Ribeiro, Junichi Yamagishi, and Robert A. J. Clark. 2015 · 2015
Earlier work this paper cites.
ExpressGesture: Expressive gesture generation from speech through database matching
Ylva Ferstl, Michael Neff, and Rachel McDonnell. 2021 · 2016
Earlier work this paper cites.
Synthetic speech detection using phase information
Ibon Saratxaga, Jon Sanchez, Zhizheng Wu, Inma Hernaez, and Eva Navas. 2016 · 2016
Earlier work this paper cites.
WaveNet: A generative model for raw audio
Aäron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew Senior, and Koray Kavukcuoglu. 2016 · 2016
Cited alongside, same era.
A hierarchical predictor of synthetic speech naturalness using neural networks. In Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech ’16) . 342–346
Takenori Yoshimura, Gustav Eje Henter, Oliver Watts, Mirjam Wester, Junichi Yamagishi, and Keiichi Tokuda. 2016 · 2016
Cited alongside, same era.
Enriching word vectors with subword information
Piotr Bojanowski, Édouard Grave, Armand Joulin, and Tomáš Mikolov. 2017 · 2017
Cited alongside, same era.
GANs trained by a two time-scale update rule converge to a local Nash equilibrium. In Advances in Neural Information Processing Systems (NIPS ’17)
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter. 2017 · 2017
Cited alongside, same era.
Double-DCCCAE: Estimation of body gestures from speech waveform. In Proceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP ’21) . 900–904
JinHong Lu, TianHang Liu, ShuZhuang Xu, and Hiroshi Shimodaira. 2021 · 2021
Later among the works it cites.
NeRF: Representing scenes as neural radiance fields for view synthesis
Ben Mildenhall, Pratul P Srinivasan, Matthew Tancik, Jonathan T. Barron, Ravi Ramamoorthi, and Ren Ng. 2021 · 2021
Later among the works it cites.
Hellinger distance
Mikhail S. Nikulin. 2001 · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision. In Proceedings of the International Conference on Machine Learning (ICML ’21) . 8748–8763
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever. 2021 · 2021
Later among the works it cites.
Passing a non-verbal Turing test: Evaluating gesture animations generated from speech. In Proceedings of the IEEE Conference on Virtual Reality and 3D User Interfaces (VR ’21) . 573–581
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
The perception-distortion tradeoff. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR ’18) . 6228–6237
Yochai Blau and Tomer Michaeli. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of deep bidirectional Transformers for language understanding. In Proceedings of the Conference of the North American Chapter of the Association for Computational Linguistics (NAACL ’18) . 4171–4186
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Cited alongside, same era.
Investigating the use of recurrent motion modelling for speech gesture generation. In Proceedings of the ACM International Conference on Intelligent Virtual Agents (IVA ’18) . 93–98
Ylva Ferstl and Rachel McDonnell. 2018 · 2018
Cited alongside, same era.
A speech-driven hand gesture generation method and evaluation in android robots
Carlos T. Ishi, Daichi Machiyashiki, Ryusuke Mikata, and Hiroshi Ishiguro. 2018 · 2018
Cited alongside, same era.
Generating body motions using spoken language in dialogue. In Proceedings of the International Conference on Intelligent Virtual Agents (IVA ’18) . 87–92
Ryo Ishii, Taichi Katayama, Ryuichiro Higashinaka, and Junji Tomita. 2018 · 2018
Cited alongside, same era.
Intergroup anxiety and willingness to accommodate: Exploring the effects of accent stereotyping and social attraction
Gretchen Montgomery and Yan Bing Zhang. 2018 · 2018
Cited alongside, same era.
Natural TTS synthesis by conditioning WaveNet on mel spectrogram predictions. In Proceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP’18) . 4799–4783
Jonathan Shen, Ruoming Pang, Ron J. Weiss, Mike Schuster, Navdeep Jaitly, Zongheng Yang, Zhifeng Chen, Yu Zhang, Yuxuan Wang, RJ Skerry-Ryan, Rif A. Saurous, Yannis Agiomyrgiannakis, and Yonghui Wu. 2018 · 2018
Cited alongside, same era.
Generation of gestures during presentation for humanoid robots. In Proceedings of the IEEE International Symposium on Robot and Human Interactive Communication (RO-MAN ’18) . 961–968
Akihito Shimazu, Chie Hieida, Takayuki Nagai, Tomoaki Nakamura, Yuki Takeda, Takenori Hara, Osamu Nakagawa, and Tsuyoshi Maeda. 2018 · 2018
Cited alongside, same era.
Manuel Rebol, Christian Güti, and Krzysztof Pietroszek. 2021 · 2021
Later among the works it cites.
The importance of qualitative elements in subjective evaluation of semantic gestures. In Proceedings of the IEEE International Conference on Automatic Face and Gesture Recognition (FG ’21) . 1–8
Carolyn Saund and Stacy Marsella. 2021 · 2021
Later among the works it cites.
Speech gesture generation from acoustic and textual information using LSTMs. In Proceedings of the International Conference on Electrical Engineering/Electronics, Computer, Telecommunications and Information Technology (ECTI-CON ’21) . 718–723
Ausdang Thangthai, Kwanchiva Thangthai, Arnon Namsanit, Sumonmas Thatphithakkul, and Sittipong Saychum. 2021 · 2021
Later among the works it cites.
Integrated speech and gesture synthesis. In Proceedings of the ACM International Conference on Multimodal Interaction (ICMI ’21) . 177–185
Siyang Wang, Simon Alexanderson, Joakim Gustafson, Jonas Beskow, Gustav Eje Henter, and Éva Székely. 2021 · 2021
Later among the works it cites.
To rate or not to rate: Investigating evaluation methods for generated co-speech gestures. In Proceedings of the ACM International Conference on Multimodal Interaction (ICMI ’21) . 494–502
Pieter Wolfert, Jeffrey M. Girard, Taras Kucherenko, and Tony Belpaeme. 2021 · 2021
Later among the works it cites.
Development of an interactive human/agent loop using multimodal recurrent neural networks. In Proceedings of the ACM International Conference on Multimodal Interaction (ICMI ’21) . 822–826
Jieyeon Woo. 2021 · 2021
Later among the works it cites.
SGToolkit: An interactive gesture authoring toolkit for embodied conversational agents. In Proceedings of the Annual ACM Symposium on User Interface Software and Technology (UIST ’21) . ACM, 826–840
Youngwoo Yoon, Keunwoo Park, Minsu Jang, Jaehong Kim, and Geehyuk Lee. 2021 · 2021
Later among the works it cites.
Low-resource adaptation for personalized co-speech gesture generation. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR ’22) . 20566–20576
Chaitanya Ahuja, Dong Won Lee, and Louis-Philippe Morency. 2022 · 2022
Later among the works it cites.
Rhythmic Gesticulator: Rhythm-aware co-speech gesture synthesis with hierarchical neural embeddings
Tenglong Ao, Qingzhe Gao, Yuke Lou, Baoquan Chen, and Libin Liu. 2022 · 2022
Later among the works it cites.
The IVI Lab entry to the GENEA Challenge 2022 – A Tacotron2 based method for co-speech gesture generation with locality-constraint attention mechanism. In Proceedings of the ACM International Conference on Multimodal Interaction (ICMI ’22) . 784–789
Che-Jui Chang, Sen Zhang, and Mubbasir Kapadia. 2022 · 2022
Later among the works it cites.
WavLM: Large-scale self-supervised pre-training for full stack speech processing
Sanyuan Chen, Chengyi Wang, Zhengyang Chen, Yu Wu, Shujie Liu, Zhuo Chen, Jinyu Li, Naoyuki Kanda, Takuya Yoshioka, Xiong Xiao, Jian Wu, Long Zhou, Shuo Ren, Yanmin Qian, Yao Qian, Jian Wu, Michael Zeng, Xiangzhan Yu, and Furu Wei. 2022 · 2022
Later among the works it cites.
Exemplar-based stylized gesture generation from speech: An entry to the GENEA Challenge 2022. In Proceedings of the ACM International Conference on Multimodal Interaction (ICMI ’22) . 778–783
Saeed Ghorbani, Ylva Ferstl, and Marc-André Carbonneau. 2022 · 2022
Later among the works it cites.
Evaluating data-driven co-speech gestures of embodied conversational agents through real-time interaction. In Proceedings of the ACM International Conference on Intelligent Virtual Agents (IVA ’22) . Article 8, 8 pages
Yuan He, André Pereira, and Taras Kucherenko. 2022 · 2022
Later among the works it cites.
Automatic quality assessment of speech-driven synthesized gestures
Zhiyuan He. 2022 · 2022
Later among the works it cites.
The VoiceMOS Challenge 2022. In Proceedings of the Annual Conference of the International Speech Communication Association (Interspeech ’22) . 4536–4540
Wen Chin Huang, Erica Cooper, Yu Tsao, Hsin-Min Wang, Tomoki Toda, and Junichi Yamagishi. 2022 · 2022
Later among the works it cites.
TransGesture: Autoregressive gesture generation with RNN-transducer. In Proceedings of the ACM International Conference on Multimodal Interaction (ICMI ’22) . 753–757
Naoshi Kaneko, Yuna Mitsubayashi, and Geng Mu. 2022 · 2022
Later among the works it cites.
ReCell: replicating recurrent cell for auto-regressive pose generation. In Companion publication of the ACM International Conference on Multimodal Interaction (ICMI ’22 Companion) . 94–97
Vladislav Korzun, Anna Beloborodva, and Arkady Ilin. 2022 · 2022
Later among the works it cites.
Multimodal analysis of the predictability of hand-hesture properties. In Proceedings of the International Conference on Autonomous Agents and Multiagent Systems (AAMAS ’22) . 770–779
Taras Kucherenko, Rajmund Nagy, Michael Neff, Hedvig Kjellström, and Gustav Eje Henter. 2022 · 2022
Later among the works it cites.
SEEG: Semantic energized co-speech gesture generation. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR ’22) . 10473–10482
Yuanzhi Liang, Qianyu Feng, Linchao Zhu, Li Hu, Pan Pan, and Yi Yang. 2022 · 2022
Later among the works it cites.
Learning hierarchical cross-modal association for co-speech gesture generation. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR ’22) . 10462–10472
Xian Liu, Qianyi Wu, Hang Zhou, Yinghao Xu, Rui Qian, Xinyi Lin, Xiaowei Zhou, Wayne Wu, Bo Dai, and Bolei Zhou. 2022a · 2022
Later among the works it cites.
The DeepMotion entry to the GENEA Challenge 2022. In Proceedings of the ACM International Conference on Multimodal Interaction (ICMI ’22) . ACM, 790–796
Shuhong Lu and Andrew Feng. 2022 · 2022
Later among the works it cites.
Hybrid seq2seq architecture for 3D co-speech gesture generation. In Proceedings of the ACM International Conference on Multimodal Interaction (ICMI ’22) . 748–752
Khaled Saleh. 2022 · 2022
Later among the works it cites.
Deep gesture generation for social robots using type-specific libraries. In Proceedings of the IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS ’22) . 8286–8291
Hitoshi Teshima, Naoki Wake, Diego Thomas, Yuta Nakashima, Hiroshi Kawasaki, and Katsushi Ikeuchi. 2022 · 2022
Later among the works it cites.
UEA Digital Humans entry to the GENEA Challenge 2022. In Proceedings of the ACM International Conference on Multimodal Interaction (ICMI ’22) . 802–810
Jonathan Windle, David Greenwood, and Sarah Taylor. 2022 · 2022
Later among the works it cites.
A review of evaluation practices of gesture generation in embodied conversational agents
Pieter Wolfert, Nicole Robinson, and Tony Belpaeme. 2022 · 2022
Later among the works it cites.
The ReprGesture entry to the GENEA Challenge 2022. In Proceedings of the ACM International Conference on Multimodal Interaction (ICMI ’22) . 758–763
Sicheng Yang, Zhiyong Wu, Minglei Li, Mengchen Zhao, Jiuxin Lin, Liyang Chen, and Weihong Bao. 2022 · 2022
Later among the works it cites.
Gesture2Vec: Clustering gestures using representation learning methods for co-speech gesture generation. In Proceedings of the IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS ’22) . 3100–3107
Payam Jome Yazdian, Mo Chen, and Angelica Lim. 2022 · 2022
Later among the works it cites.
Audio-driven stylized gesture generation with flow-based model. In Proceedings of the European Conference on Computer Vision (ECCV ’22) . 712–728
Sheng Ye, Yu-Hui Wen, Yanan Sun, Ying He, Ziyang Zhang, Yaoyuan Wang, Weihua He, and Yong-Jin Liu. 2022 · 2022
Later among the works it cites.
The GENEA Challenge 2022: A large evaluation of data-driven co-speech gesture generation. In Proceedings of the ACM International Conference on Multimodal Interaction (ICMI ’22) . 736–747
Youngwoo Yoon, Pieter Wolfert, Taras Kucherenko, Carla Viegas, Teodor Nikolov, Mihail Tsakov, and Gustav Eje Henter. 2022 · 2022
Later among the works it cites.
GestureMaster: Graph-based speech-driven gesture generation. In Proceedings of the ACM International Conference on Multimodal Interaction (ICMI ’22) . 764–770
Chi Zhou, Tengyue Bian, and Kang Chen. 2022 · 2022
Later among the works it cites.
Listen, Denoise, Action! Audio-Driven Motion Synthesis with Diffusion Models
Simon Alexanderson, Rajmund Nagy, Jonas Beskow, and Gustav Eje Henter. 2023 · 2023
Closest in time.
The GENEA Challenge 2023: A large-scale evaluation of gesture generation models in monadic and dyadic settings. In Proceedings of the ACM International Conference on Multimodal Interaction (ICMI ’23) . 792–801
Taras Kucherenko, Rajmund Nagy, Youngwoo Yoon, Jieyeon Woo, Teodor Nikolov, Mihail Tsakov, and Gustav Eje Henter. 2023 · 2023
Closest in time.
Objective evaluation metric for motion generative models: Validating Fréchet motion distance on foot skating and over-smoothing artifacts. In Proceedings of the ACM SIGGRAPH Conference on Motion, Interaction and Games (MIG ’23) . Article 2, 11 pages
Antoine Maiorca, Hugo Bohy, Youngwoo Yoon, and Thierry Dutoit. 2023 · 2023
Closest in time.
Diff-TTSG: Denoising probabilistic integrated speech and gesture synthesis. In Proceedings of the ISCA Speech Synthesis Workshop (SSW ’23)
Shivam Mehta, Siyang Wang, Simon Alexanderson, Jonas Beskow, Éva Székely, and Gustav Eje Henter. 2023 · 2023
Closest in time.
A comprehensive review of data-driven co-speech gesture generation
Simbarashe Nyatsanga, Taras Kucherenko, Chaitanya Ahuja, Gustav Eje Henter, and Michael Neff. 2023 · 2023
Closest in time.
“Am I listening?”, Evaluating the quality of generated data-driven listening motion. In Companion publication of the ACM International Conference on Multimodal Interaction (ICMI ’23 Companion) . 6–10
Pieter Wolfert, Gustav Eje Henter, and Tony Belpaeme. 2023 · 2023
Closest in time.
DiffMotion: Speech-driven gesture synthesis using denoising diffusion model. In Proceedings of the International Conference on Multimedia Modeling (MMM ’23) . 231–242
Fan Zhang, Naye Ji, Fuxing Gao, and Yongping Li. 2023 · 2023
Closest in time.
CLIC 2020: Overview, and analysis of the competition results
George Toderici, Lucas Theis, Nick Johnston, Eirikur Agustsson, Johannes Ballé, Fabian Mentzer, Wenzhe Shi, and Radu Timofte. 2020 · 2024
Closest in time.
Speech2AffectiveGestures: Synthesizing co-speech gestures with generative adversarial affective expression learning. In Proceedings of the ACM International Conference on Multimedia (MM ’21) . 2027–2036
Uttaran Bhattacharya, Elizabeth Childs, Nicholas Rewkowski, and Dinesh Manocha. 2021 · 2036
Closest in time.