9 papers
CanonSLR: Canonical-View Guided Multi-View Continuous Sign Language Recognition
Xu Wang, Shengeng Tang, Wan Jiang +3
Continuous Sign Language Recognition (CSLR) has achieved remarkable progress in recent years; however, most existing methods are developed under single-view settings and thus remai…
Text2Lip: Progressive Lip-Synced Talking Face Generation from Text via Viseme-Guided Rendering
Xu Wang, Shengeng Tang, Fei Wang +4
Generating semantically coherent and visually accurate talking faces requires bridging the gap between linguistic meaning and facial articulation. Although audio-driven methods rem…
Moderating the Generalization of Score-based Generative Model
Wan Jiang, He Wang, Xin Zhang +4
Score-based Generative Models (SGMs) have demonstrated remarkable generalization abilities, e.g. generating unseen, but natural data. However, the greater the generalization power,…
StgcDiff: Spatial-Temporal Graph Condition Diffusion for Sign Language Transition Generation
Jiashu He, Jiayi He, Shengeng Tang +3
Sign language transition generation seeks to convert discrete sign language segments into continuous sign videos by synthesizing smooth transitions. However,most existing methods m…
SignAligner: Harmonizing Complementary Pose Modalities for Coherent Sign Language Generation
Xu Wang, Shengeng Tang, Lechao Cheng +3
Sign language generation aims to produce diverse sign representations based on spoken language. However, achieving realistic and naturalistic generation remains a significant chall…
Learning Speaker-Invariant Visual Features for Lipreading
Yu Li, Feng Xue, Shujie Li +4
Lipreading is a challenging cross-modal task that aims to convert visual lip movements into spoken text. Existing lipreading methods often extract visual features that include spea…