arXiv · 2105.02742
Pose-Guided Sign Language Video GAN with Dynamic Lambda
Abstract
We propose a novel approach for the synthesis of sign language videos using GANs. We extend the previous work of Stoll et al. by using the human semantic parser of the Soft-Gated Warping-GAN from to produce photorealistic videos guided by region-level spatial layouts. Synthesizing target poses improves performance on independent and contrasting signers. Therefore, we have evaluated our system with the highly heterogeneous MS-ASL dataset with over 200 signers resulting in a SSIM of 0.893. Furthermore, we introduce a periodic weighting approach to the generator that reactivates the training and leads to quantitatively better results.
Explore related subjects
Keep this discovery
Christopher Kissel, Christopher Kümmel, Dennis Ritter, Kristian Hildebrand. 2021-05-06. Pose-Guided Sign Language Video GAN with Dynamic Lambda. https://arxiv.org/abs/2105.02742
Cite the original work for its findings. Save a collection to share your selection of sources.