Fetching the paper…
Reading the bibliography…
We present a framework for generating full-bodied photorealistic avatars that gesture according to the conversational dynamics of a dyadic interaction.
Nonverbal leakage and clues to deception
Paul Ekman and Wallace V Friesen · 1969
Earlier work this paper cites.
Animated conversation: Rule-based generation of facial expression, gesture and spoken intonation for multiple conversational agents
Justine Cassell, Catherine Pelachaud, Norman Badler, Mark Steedman, Brett Achorn, Tripp Becket, Brett Douville, Scott Prevost, and Matthew Stone · 1994
Earlier work this paper cites.
Virtual rapport
Jonathan Gratch, Anna Okhmatovskaia, Francois Lamothe, Stacy Marsella, Mathieu Morales, Rick J van der Werf, and Louis-Philippe Morency · 2006
Earlier work this paper cites.
A review of vector quantization techniques
A Vasuki and PT Vanathi · 2006
Earlier work this paper cites.
Anthropomorphism influences perception of computer-animated characters’ actions
Thierry Chaminade, Jessica Hodgins, and Mitsuo Kawato · 2007
Earlier work this paper cites.
Facilitating multiparty dialog with gaze, gesture, and speech
Dan Bohus and Eric Horvitz · 2010
Earlier work this paper cites.
Virtual rapport 2.0
Lixing Huang, Louis-Philippe Morency, and Jonathan Gratch · 2011
Earlier work this paper cites.
Render me real? investigating the effect of render style on the perception of animated virtual humans
Rachel McDonnell, Martin Breidt, and Heinrich H Bülthoff · 2012
Earlier work this paper cites.
Learn2smile: Learning non-verbal interaction through observation
Will Feng, Anitha Kannan, Georgia Gkioxari, and Larry Zitnick · 2017
Earlier work this paper cites.
Dyadgan: Generating facial expressions in dyadic interactions
Yuchi Huang and Saad M Khan · 2017
Earlier work this paper cites.
Neural discrete representation learning
Aaron Van Den Oord, Oriol Vinyals, et al · 2017
Earlier work this paper cites.
Deep appearance models for face rendering
Stephen Lombardi, Jason Saragih, Tomas Simon, and Yaser Sheikh · 2018
Earlier work this paper cites.
Interactive generative adversarial networks for facial expression generation in dyadic interactions
Behnaz Nojavanasghari, Yuchi Huang, and Saad Khan · 2018
Earlier work this paper cites.
To react or not to react: End-to-end visual pose forecasting for personalized avatar during dyadic conversations
Chaitanya Ahuja, Shugao Ma, Louis-Philippe Morency, and Yaser Sheikh · 2019
Earlier work this paper cites.
Capture, learning, and synthesis of 3D speaking styles
Daniel Cudeiro, Timo Bolkart, Cassidy Laidlaw, Anurag Ranjan, and Michael Black · 2019
Earlier work this paper cites.
Learning individual styles of conversational gesture
Shiry Ginosar, Amir Bar, Gefen Kohavi, Caroline Chan, Andrew Owens, and Jitendra Malik · 2019
Cited alongside, same era.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Li Du, Maxwell Forbes, and Yejin Choi · 2019
Cited alongside, same era.
Learning non-verbal behavior for a social robot from youtube videos
Patrik Jonell, Taras Kucherenko, Erik Ekstedt, and Jonas Beskow · 2019
Cited alongside, same era.
Towards social artificial intelligence: Nonverbal social signal prediction in a triadic interaction
Hanbyul Joo, Tomas Simon, Mina Cikara, and Yaser Sheikh · 2019
Cited alongside, same era.
Talking with hands 16.2 m: A large-scale dataset of synchronized body-finger motion and audio for conversational motion analysis and synthesis
Gilwoo Lee, Zhiwei Deng, Shugao Ma, Takaaki Shiratori, Siddhartha S Srinivasa, and Yaser Sheikh · 2019
Cited alongside, same era.
Improved denoising diffusion probabilistic models
Alexander Quinn Nichol and Prafulla Dhariwal · 2021
Later among the works it cites.
Meshtalk: 3d face animation from speech using cross-modality disentanglement
Alexander Richard, Michael Zollhoefer, Yandong Wen, Fernando de la Torre, and Yaser Sheikh · 2021
Later among the works it cites.
Soundstream: An end-to-end neural audio codec
Neil Zeghidour, Alejandro Luebs, Ahmed Omran, Jan Skoglund, and Marco Tagliasacchi · 2021
Later among the works it cites.
Beat: A large-scale semantic and emotional multi-modal dataset for conversational gestures synthesis
Haiyang Liu, Zihao Zhu, Naoya Iwamoto, Yichen Peng, Zhengqing Li, You Zhou, Elif Bozkurt, and Bo Zheng · 2022
Later among the works it cites.
Learning to listen: Modeling non-deterministic dyadic facial motion
Evonne Ng, Hanbyul Joo, Liwen Hu, Hao Li, Trevor Darrell, Angjoo Kanazawa, and Shiry Ginosar · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
wav2vec 2.0: A framework for self-supervised learning of speech representations
Alexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, and Michael Auli · 2020
Cited alongside, same era.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel · 2020
Cited alongside, same era.
Neural voice puppetry: Audio-driven facial reenactment
Justus Thies, Mohamed Elgharib, Ayush Tewari, Christian Theobalt, and Matthias Nießner · 2020
Cited alongside, same era.
Realistic speech-driven facial animation with gans
Konstantinos Vougioukas, Stavros Petridis, and Maja Pantic · 2020
Cited alongside, same era.
Apb2face: Audio-guided face reenactment with auxiliary pose and blink signals
Jiangning Zhang, Liang Liu, Zhucun Xue, and Yong Liu · 2020
Cited alongside, same era.
Driving-signal aware full-body avatars
Timur Bagautdinov, Chenglei Wu, Tomas Simon, Fabián Prada, Takaaki Shiratori, Shih-En Wei, Weipeng Xu, Yaser Sheikh, and Jason Saragih · 2021
Cited alongside, same era.
Taming transformers for high-resolution image synthesis
Patrick Esser, Robin Rombach, and Bjorn Ommer · 2021
Cited alongside, same era.
Guy Tevet, Sigal Raab, Brian Gordon, Yonatan Shafir, Daniel Cohen-Or, and Amit H Bermano · 2022
Later among the works it cites.
Responsive listening head generation: A benchmark dataset and baseline
Mohan Zhou, Yalong Bai, Wei Zhang, Ting Yao, Tiejun Zhao, and Tao Mei · 2022
Later among the works it cites.
Listen, denoise, action! audio-driven motion synthesis with diffusion models
Simon Alexanderson, Rajmund Nagy, Jonas Beskow, and Gustav Eje Henter · 2023
Later among the works it cites.
Gesturediffuclip: Gesture diffusion model with clip latents
Tenglong Ao, Zeyi Zhang, and Libin Liu · 2023
Later among the works it cites.
Can language models learn to listen?
Evonne Ng, Sanjay Subramanian, Dan Klein, Angjoo Kanazawa, Trevor Darrell, and Shiry Ginosar · 2023
Later among the works it cites.
Social diffusion: Long-term multiple human motion anticipation
Julian Tanke, Linguang Zhang, Amy Zhao, Chengcheng Tang, Yujun Cai, Lezi Wang, Po-Chen Wu, Juergen Gall, and Cem Keskin · 2023
Later among the works it cites.
Edge: Editable dance generation from music
Jonathan Tseng, Rodrigo Castellon, and Karen Liu · 2023
Later among the works it cites.
Generating holistic 3d human motion from speech
Hongwei Yi, Hualin Liang, Yifei Liu, Qiong Cao, Yandong Wen, Timo Bolkart, Dacheng Tao, and Michael J Black · 2023
Later among the works it cites.
Talking head generation with probabilistic audio-to-visual diffusion priors
Zhentao Yu, Zixin Yin, Deyu Zhou, Duomin Wang, Finn Wong, and Baoyuan Wang · 2023
Later among the works it cites.
Livelyspeaker: Towards semantic-aware co-speech gesture generation
Yihao Zhi, Xiaodong Cun, Xuelin Chen, Xi Shen, Wen Guo, Shaoli Huang, and Shenghua Gao · 2023
Later among the works it cites.