Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
Soul: Breathe Life into Digital Human for High-fidelity Long-term Multimodal Animation
Jiangning Zhang, Junwei Zhu, Zhenye Gan +14
We propose a multimodal-driven framework for high-fidelity long-term digital human animation termed , which generates semantically coherent videos from a single-fram…
cs.CV2024
MDT-A2G: Exploring Masked Diffusion Transformers for Co-Speech Gesture Generation
Xiaofeng Mao, Zhengkai Jiang, Qilin Wang +7
Recent advancements in the field of Diffusion Transformers have substantially improved the generation of high-quality 2D images, 3D videos, and 3D shapes. However, the effectivenes…
cs.CV2024
A Coarse-to-Fine Place Recognition Approach using Attention-guided Descriptors and Overlap Estimation
Chencan Fu, Lin Li, Jianbiao Mei +4
Place recognition is a challenging but crucial task in robotics. Current description-based methods may be limited by representation capabilities, while pairwise similarity-based me…