Fetching the paper…
Reading the bibliography…
In this paper, we propose a dual-condition diffusion pre-training model named SignDiff that can generate human sign language speakers from a skeleton pose.
Sign Language Structure
W. C. Stokoe · 1980
Earlier work this paper cites.
Using Dynamic Time Warping to Find Patterns in Time Series
D. J. Berndt and J. Clifford · 1994
Earlier work this paper cites.
Isolated Sign Language Recognition using Hidden Markov Models
K. Grobel and M. Assan · 1997
Earlier work this paper cites.
TESSA, a System to Aid Communication with Deaf People
S. Cox, M. Lincoln, J. Tryggvason, M. Nakisa, M. Wells, M. Tutt, and S. Abbott · 2002
Earlier work this paper cites.
Minimal Training, Large Lexicon, Unconstrained Sign Language Recognition
T. Kadir, R. Bowden, E.-J. Ong, and A. Zisserman · 2004
Earlier work this paper cites.
Image Quality Assessment: From Error Visibility to Structural Similarity
Z. Wang, A. C. Bovik, H. R. Sheikh, and E. P. Simoncelli · 2004
Earlier work this paper cites.
Large Lexicon Detection of Sign Language
H. Cooper and R. Bowden · 2007
Earlier work this paper cites.
Educational Resources and Implementation of a Greek Sign Language Synthesis Architecture
K. Karpouzis, G. Caridakis, S.-E. Fotinea, and E. Efthimiou · 2007
Earlier work this paper cites.
A Study of Sign Language Coarticulation
J. Segouat · 2009
Earlier work this paper cites.
RWTH-PHOENIX-Weather: A Large Vocabulary Sign Language Recognition and Translation Corpus
J. Forster, C. Schmidt, T. Hoyoux, O. Koller, U. Zelle, J. H. Piater, and H. Ney · 2012
Earlier work this paper cites.
Generative Adversarial Nets
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2014
Earlier work this paper cites.
Continuous Sign Language Recognition: Towards Large Vocabulary Statistical Recognition Systems Handling Multiple Signers
O. Koller, J. Forster, and H. Ney · 2015
Earlier work this paper cites.
Unsupervised Representation Learning with Deep Convolutional Generative Adversarial Networks
A. Radford, L. Metz, and S. Chintala · 2015
Earlier work this paper cites.
U-net: Convolutional Networks for Biomedical Image Segmentation
O. Ronneberger, P. Fischer, and T. Brox · 2015
Earlier work this paper cites.
Going deeper with convolutions
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich · 2015
Earlier work this paper cites.
Automated Technique for Real-Time Production of Lifelike Animations of American Sign Language
J. McDonald, R. Wolfe, J. Schnepp, J. Hochgesang, D. G. Jamrozik, M. Stumbo, L. Berke, M. Bialek, and F. Thomas · 2016
Earlier work this paper cites.
Generating Videos with Scene Dynamics
C. Vondrick, H. Pirsiavash, and A. Torralba · 2016
Earlier work this paper cites.
OpenPose: Realtime Multi-Person 2D Pose Estimation using Part Affinity Fields
Z. Cao, G. Hidalgo, T. Simon, S.-E. Wei, and Y. Sheikh · 2017
Earlier work this paper cites.
Recurrent Convolutional Neural Networks for Continuous Sign Language Recognition by Staged Optimization
R. Cui, H. Liu, and C. Zhang · 2017
Earlier work this paper cites.
Sign Language Interpreting in the Workplace
J. Dickinson · 2017
Earlier work this paper cites.
GANs Trained by a Two Time-Scale Update Rule Converge to a Local Nash Equilibrium
M. Heusel, H. Ramsauer, T. Unterthiner, B. Nessler, and S. Hochreiter · 2017
Earlier work this paper cites.
Image-to-Image Translation with Conditional Adversarial Networks
P. Isola, J.-Y. Zhu, T. Zhou, and A. A. Efros · 2017
Earlier work this paper cites.
Pose Guided Person Image Generation
L. Ma, X. Jia, Q. Sun, B. Schiele, T. Tuytelaars, and L. Van Gool · 2017
Earlier work this paper cites.
Unpaired Image-to-Image Translation using Cycle-Consistent Adversarial Networks
J.-Y. Zhu, T. Park, P. Isola, and A. A. Efros · 2017
Earlier work this paper cites.
Synthesizing Images of Humans in Unseen Poses
G. Balakrishnan, A. Zhao, A. V. Dalca, F. Durand, and J. Guttag · 2018
Earlier work this paper cites.
Production and Comprehension of Prosodic Markers in Sign Language Imperatives
D. Brentari, J. Falk, A. Giannakidou, A. Herrmann, E. Volk, and M. Steinbach · 2018
Earlier work this paper cites.
Neural Sign Language Translation
N. C. Camgöz, S. Hadfield, O. Koller, H. Ney, and R. Bowden · 2018
Cited alongside, same era.
Densepose: Dense human pose estimation in the wild
R. A. Güler, N. Neverova, and I. Kokkinos · 2018
Cited alongside, same era.
Deformable GANs for Pose-Based Human Image Generation
A. Siarohin, E. Sangineto, S. Lathuiliere, and N. Sebe · 2018
Cited alongside, same era.
Sign Language Production using Neural Machine Translation and Generative Adversarial Networks
S. Stoll, N. C. Camgöz, S. Hadfield, and R. Bowden · 2018
Cited alongside, same era.
GestureGAN for Hand Gesture-to-Gesture Translation in the wild
H. Tang, W. Wang, D. Xu, Y. Yan, and N. Sebe · 2018
Cited alongside, same era.
MoCoGAN: Decomposing Motion and Content for Video Generation
S. Tulyakov, M.-Y. Liu, X. Yang, and J. Kautz · 2018
Cited alongside, same era.
Progressive Transformers for End-to-End Sign Language Production
B. Saunders, N. C. Camgöz, and R. Bowden · 2020
Later among the works it cites.
Text2Sign: Towards Sign Language Production using Neural Machine Translation and Generative Adversarial Networks
S. Stoll, N. C. Camgöz, S. Hadfield, and R. Bowden · 2020
Later among the works it cites.
XingGAN for Person Image Generation
H. Tang, S. Bai, L. Zhang, P. H. Torr, and N. Sebe · 2020
Later among the works it cites.
Neural voice puppetry: Audio-driven facial reenactment
J. Thies, M. Elgharib, A. Tewari, C. Theobalt, and M. Nießner · 2020
Later among the works it cites.
Can Everybody Sign Now? Exploring Sign Language Video Generation from 2D Poses
L. Ventura, A. Duarte, and X. Giró-i Nieto · 2020
Later among the works it cites.
GAC-GAN: A General Method for Appearance-Controllable Human Video Motion Transfer
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Video-to-Video Synthesis
T.-C. Wang, M.-Y. Liu, J.-Y. Zhu, G. Liu, A. Tao, J. Kautz, and B. Catanzaro · 2018
Cited alongside, same era.
High-Resolution Image Synthesis and Semantic Manipulation with Conditional GANs
T.-C. Wang, M.-Y. Liu, J.-Y. Zhu, A. Tao, J. Kautz, and B. Catanzaro · 2018
Cited alongside, same era.
Everybody Dance Now
C. Chan, S. Ginosar, T. Zhou, and A. A. Efros · 2019
Cited alongside, same era.
Deep Gesture Video Generation With Learning on Regions of Interest
R. Cui, Z. Cao, W. Pan, C. Zhang, and J. Wang · 2019
Cited alongside, same era.
3D Hand Shape and Pose Estimation from a Single RGB Image
L. Ge, Z. Ren, Y. Li, Z. Xue, Y. Wang, J. Cai, and J. Yuan · 2019
Cited alongside, same era.
Neural Sign Language Translation based on Human Keypoint Estimation
S.-K. Ko, C. J. Kim, H. Jung, and C. Cho · 2019
Cited alongside, same era.
D. Wei, X. Xu, H. Shen, and K. Huang · 2020
Later among the works it cites.
MM-Hand: 3D-Aware Multi-Modal Guided Hand Generative Network for 3D Hand Pose Synthesis
Z. Wu, D. Hoang, S.-Y. Lin, Y. Xie, L. Chen, Y.-Y. Lin, Z. Wang, and W. Fan · 2020
Later among the works it cites.
Neural Sign Language Synthesis: Words Are Our Glosses
J. Zelinka and J. Kanis · 2020
Later among the works it cites.
How2Sign: A Large-Scale Multimodal Dataset for Continuous American Sign Language
A. Duarte, S. Palaskar, L. Ventura, D. Ghadiyaram, K. DeHaan, F. Metze, J. Torres, and X. Giro-i Nieto · 2021
Later among the works it cites.
Towards Fast and High-Quality Sign Language Production
W. Huang, W. Pan, Z. Zhao, and Q. Tian · 2021
Later among the works it cites.
Translating sign language videos to talking faces
S. Mazumder, R. Mukhopadhyay, V. P. Namboodiri, and C. V. Jawahar · 2021
Later among the works it cites.
AnonySign: Novel Human Appearance Synthesis for Sign Language Video Anonymisation
B. Saunders, N. C. Camgöz, and R. Bowden · 2021
Later among the works it cites.
Continuous 3D Multi-Channel Sign Language Production via Progressive Transformers and Mixture Density Networks
B. Saunders, N. C. Camgöz, and R. Bowden · 2021
Later among the works it cites.
Mixed SIGNals: Sign Language Production via a Mixture of Motion Primitives
B. Saunders, N. C. Camgöz, and R. Bowden · 2021
Later among the works it cites.
Skeletal Graph Self-Attention: Embedding a Skeleton Inductive Bias into Sign Language Production
B. Saunders, N. C. Camgoz, and R. Bowden · 2021
Later among the works it cites.
Multimodal inputs driven talking face generation with spatial–temporal dependency
L. Yu, J. Yu, M. Li, and Q. Ling · 2021
Later among the works it cites.
Sign pose-based transformer for word-level sign language recognition
M. Boháček and M. Hrúz · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer · 2022
Later among the works it cites.
Photorealistic text-to-image diffusion models with deep language understanding
C. Saharia, W. Chan, S. Saxena, L. Li, J. Whang, E. L. Denton, K. Ghasemipour, R. Gontijo Lopes, B. Karagol Ayan, T. Salimans, J. Ho, D. J. Fleet, and M. Norouzi · 2022
Later among the works it cites.
Signing at scale: Learning to co-articulate signs for large-scale photo-realistic sign language production
B. Saunders, N. C. Camgoz, and R. Bowden · 2022
Later among the works it cites.
Text2video: Text-driven talking-head video synthesis with personalized phoneme - pose dictionary
S. Zhang, J. Yuan, M. Liao, and L. Zhang · 2022
Later among the works it cites.
Adding conditional control to text-to-image diffusion models, 2023
L. Zhang and M. Agrawala · 2023
Closest in time.
Neural sign actors: A diffusion model for 3d sign language production from text
V. Baltatzis, R. A. Potamias, E. Ververas, G. Sun, J. Deng, and S. Zafeiriou · 2024
Closest in time.
Signllm: Sign languages production large language models, 2024
S. Fang, L. Wang, C. Zheng, Y. Tian, and C. Chen · 2024
Closest in time.
Signx: The foundation model for sign recognition, 2025
S. Fang, C. Sui, H. Yi, C. Neidle, and D. N. Metaxas · 2025
Closest in time.