Fetching the paper…
Reading the bibliography…
We present ZeroEGGS, a neural network framework for speech-driven gesture generation with zero-shot style control by example.
BEAT: the Behavior Expression Animation Toolkit
Cassell J., Vilhjálmsson H. H., Bickmore T · 2001
Earlier work this paper cites.
Max-a multimodal assistant in virtual reality construction
Kopp S., Jung B., Lessmann N., Wachsmuth I · 2003
Earlier work this paper cites.
Gesture and the communicative intention of the speaker
Melinger A., Levelt W. J · 2004
Earlier work this paper cites.
Nonverbal Behavior Generator for Embodied Conversational Agents
Lee J., Marsella S · 2006
Earlier work this paper cites.
Gesture modeling and animation based on a probabilistic re-creation of speaker style
Neff M., Kipp M., Albrecht I., Seidel H.-P · 2008
Earlier work this paper cites.
The interplay between gesture and speech in the production of referring expressions: Investigating the tradeoff hypothesis
De Ruiter J. P., Bangerter A., Dings P · 2012
Earlier work this paper cites.
Auto-encoding variational bayes
Kingma D. P., Welling M · 2013
Earlier work this paper cites.
Virtual character performance from speech
Marsella S., Xu Y., Lhommet M., Feng A., Scherer S., Shapiro A · 2013
Earlier work this paper cites.
Deep speech: Scaling up end-to-end speech recognition
Hannun A., Case C., Casper J., Catanzaro B., Diamos G., Elsen E., Prenger R., Satheesh S., Sengupta S., Coates A., et al · 2014
Earlier work this paper cites.
Method for the subjective assessment of intermediate quality level of audio systems (mushra)
ITU · 2015
Earlier work this paper cites.
Generating sentences from a continuous space
Bowman S. R., Vilnis L., Vinyals O., Dai A. M., Józefowicz R., Bengio S · 2016
Earlier work this paper cites.
Attention is all you need
Vaswani A., Shazeer N., Parmar N., Uszkoreit J., Jones L., Gomez A. N., Kaiser Ł., Polosukhin I · 2017
Earlier work this paper cites.
Investigating the use of recurrent motion modelling for speech gesture generation
Ferstl Y., McDonnell R · 2018
Cited alongside, same era.
Recurrent transition networks for character locomotion
Harvey F. G., Pal C · 2018
Cited alongside, same era.
Mode-adaptive neural networks for quadruped motion control
Zhang H., Starke S., Komura T., Saito J · 2018
Cited alongside, same era.
Multi-objective adversarial gesture generation
Ferstl Y., Neff M., McDonnell R · 2019
Cited alongside, same era.
Learning individual styles of conversational gesture
Ginosar S., Bar A., Kohavi G., Chan C., Owens A., Malik J · 2019
Cited alongside, same era.
Hierarchical Generative Modeling for Controllable Speech Synthesis
Hsu W.-N., Zhang Y., Weiss R. J., Zen H., Wu Y., Wang Y., Cao Y., Jia Y., Chen Z., Shen J., Nguyen P., Pang R · 2019
Learned motion matching
Holden D., Kanoun O., Perepichka M., Popa T · 2020
Later among the works it cites.
Gesticulator: A framework for semantically-aware speech-driven gesture generation
Kucherenko T., Jonell P., van Waveren S., Henter G. E., Alexandersson S., Leite I., Kjellström H · 2020
Later among the works it cites.
Speech gesture generation from the trimodal context of text, audio, and speaker identity
Yoon Y., Cha B., Lee J.-H., Jang M., Lee J., Kim J., Lee G · 2020
Later among the works it cites.
Statistics-based motion synthesis for social conversations
Yang Y., Yang J., Hodgins J · 2020
Later among the works it cites.
Hemvip: Human evaluation of multiple videos in parallel
Jonell P., Yoon Y., Wolfert P., Kucherenko T., Henter G. E · 2021
Later among the works it cites.
A large, crowdsourced evaluation of gesture generation systems on common data: The genea challenge 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
On the variance of the adaptive learning rate and beyond
Liu L., Jiang H., He P., Chen W., Liu X., Gao J., Han J · 2019
Cited alongside, same era.
Variational autoencoders pursue pca directions (by accident)
Rolinek M., Zietlow D., Martius G · 2019
Cited alongside, same era.
Style-controllable speech-driven gesture synthesis using normalising flows
Alexanderson S., Henter G. E., Kucherenko T., Beskow J · 2020
Cited alongside, same era.
Style transfer for co-speech gesture animation: A multi-speaker conditional-mixture approach
Ahuja C., Lee D. W., Nakano Y. I., Morency L.-P · 2020
Cited alongside, same era.
Unpaired motion style transfer from video to animation
Aberman K., Weng Y., Lischinski D., Cohen-Or D., Chen B · 2020
Cited alongside, same era.
Style Tokens: Unsupervised Style Modeling, Control and Transfer in End-to-End Speech Synthesis
Wang Y., Stanton D., Zhang Y., Skerry-Ryan R., Battenberg E., Shor J., Xiao Y., Ren F., Jia Y., Saurous R. A
Cited in the paper.
Kucherenko T., Jonell P., Yoon Y., Wolfert P., Henter G. E · 2021
Later among the works it cites.
Passing a non-verbal turing test: Evaluatina gesture animations generated from speech
Rebol M., Güti C., Pietroszek K · 2021
Later among the works it cites.
A conversational agent framework with multi-modal personality expression
Sonlu S., Güdükbay U., Durupinar F · 2021
Later among the works it cites.
Transflower: probabilistic autoregressive dance generation with multimodal attention
Valle-Pérez G., Henter G. E., Beskow J., Holzapfel A., Oudeyer P.-Y., Alexanderson S · 2021
Later among the works it cites.
Sgtoolkit: An interactive gesture authoring toolkit for embodied conversational agents
Yoon Y., Park K., Jang M., Kim J., Lee G · 2021
Later among the works it cites.
Daft-exprt: Robust prosody transfer across speakers for expressive speech synthesis
Zaïdi J., Seuté H., van Niekerk B., Carbonneau M.-A · 2021
Later among the works it cites.