3 papers
cs.CV2025
Spatial-Temporal Graph Mamba for Music-Guided Dance Video Synthesis
Hao Tang, Ling Shao, Zhenyu Zhang +2
We propose a novel spatial-temporal graph Mamba (STG-Mamba) for the music-guided dance video synthesis task, i.e., to translate the input music to a dance video. STG-Mamba consists…
cs.CV2025
Enhanced Multi-Scale Cross-Attention for Person Image Generation
Hao Tang, Ling Shao, Nicu Sebe +1
In this paper, we propose a novel cross-attention-based generative adversarial network (GAN) for the challenging person image generation task. Cross-attention is a novel and intuit…
cs.CV2024
Hierarchical Cross-Attention Network for Virtual Try-On
Hao Tang, Bin Ren, Pingping Wu +1
In this paper, we present an innovative solution for the challenges of the virtual try-on task: our novel Hierarchical Cross-Attention Network (HCANet). HCANet is crafted with two…