1 citations · 1 across the 12 of their papers we have counts for
Showing cs.MMShow all
3 papers · 1 filter
cs.MM2026
Ges-QA: A Multidimensional Quality Assessment Dataset for Audio-to-3D Gesture Generation
Zhilin Gao, Yunhao Li, Sijing Wu +3
The Audio-to-3D-Gesture (A2G) task has enormous potential for various applications in virtual reality and computer graphics, etc. However, current evaluation metrics, such as Fréc…
cs.MM2026
SFQA: A Comprehensive Perceptual Quality Assessment Dataset for Singing Face Generation
Zhilin Gao, Yunhao Li, Sijing Wu +3
The Talking Face Generation task has enormous potential for various applications in digital humans and agents, etc. Singing, as a common facial movement second only to talking, can…
cs.MM2024
MVBIND: Self-Supervised Music Recommendation For Videos Via Embedding Space Binding
Jiajie Teng, Huiyu Duan, Yucheng Zhu +2
Recent years have witnessed the rapid development of short videos, which usually contain both visual and audio modalities. Background music is important to the short videos, which…