8 citations · 15 across the 19 of their papers we have counts for
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2024
VidMusician: Video-to-Music Generation with Semantic-Rhythmic Alignment via Hierarchical Visual Features
Sifei Li, Binxin Yang, Chunji Yin +4
Video-to-music generation presents significant potential in video production, requiring the generated music to be both semantically and rhythmically aligned with the video. Achievi…
cs.SD2024
Music Style Transfer with Time-Varying Inversion of Diffusion Models
Sifei Li, Yuxin Zhang, Fan Tang +3
With the development of diffusion models, text-guided image style transfer has demonstrated high-quality controllable synthesis results. However, the utilization of text for divers…
cs.SD2024
Dance-to-Music Generation with Encoder-based Textual Inversion
Sifei Li, Weiming Dong, Yuxin Zhang +5
The seamless integration of music with dance movements is essential for communicating the artistic intent of a dance piece. This alignment also significantly improves the immersive…