activity
20242026
collaborators

9 papers

cs.CV2026

CanonSLR: Canonical-View Guided Multi-View Continuous Sign Language Recognition

Xu Wang, Shengeng Tang, Wan Jiang +3

Continuous Sign Language Recognition (CSLR) has achieved remarkable progress in recent years; however, most existing methods are developed under single-view settings and thus remai…

cs.CV2025

Text2Lip: Progressive Lip-Synced Talking Face Generation from Text via Viseme-Guided Rendering

Xu Wang, Shengeng Tang, Fei Wang +4

Generating semantically coherent and visually accurate talking faces requires bridging the gap between linguistic meaning and facial articulation. Although audio-driven methods rem…

cs.LG2025

Moderating the Generalization of Score-based Generative Model

Wan Jiang, He Wang, Xin Zhang +4

Score-based Generative Models (SGMs) have demonstrated remarkable generalization abilities, e.g. generating unseen, but natural data. However, the greater the generalization power,…

cs.CV2025

StgcDiff: Spatial-Temporal Graph Condition Diffusion for Sign Language Transition Generation

Jiashu He, Jiayi He, Shengeng Tang +3

Sign language transition generation seeks to convert discrete sign language segments into continuous sign videos by synthesizing smooth transitions. However,most existing methods m…

cs.CV2025

SignAligner: Harmonizing Complementary Pose Modalities for Coherent Sign Language Generation

Xu Wang, Shengeng Tang, Lechao Cheng +3

Sign language generation aims to produce diverse sign representations based on spoken language. However, achieving realistic and naturalistic generation remains a significant chall…

cs.CV2025

Learning Speaker-Invariant Visual Features for Lipreading

Yu Li, Feng Xue, Shujie Li +4

Lipreading is a challenging cross-modal task that aims to convert visual lip movements into spoken text. Existing lipreading methods often extract visual features that include spea…