3 papers
cs.SD2024
DiM-Gestor: Co-Speech Gesture Generation with Adaptive Layer Normalization Mamba-2
Fan Zhang, Siyuan Zhao, Naye Ji +12
Speech-driven gesture generation using transformer-based generative models represents a rapidly advancing area within virtual human creation. However, existing models face signific…
cs.GR2024
DiM-Gesture: Co-Speech Gesture Generation with Adaptive Layer Normalization Mamba-2 framework
Fan Zhang, Naye Ji, Fuxing Gao +9
Speech-driven gesture generation is an emerging domain within virtual human creation, where current methods predominantly utilize Transformer-based architectures that necessitate e…
cs.SD2024
Audio is all in one: speech-driven gesture synthetics using WavLM pre-trained model
Fan Zhang, Naye Ji, Fuxing Gao +3
The generation of co-speech gestures for digital humans is an emerging area in the field of virtual human creation. Prior research has made progress by using acoustic and semantic…