1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CV2024
UniVG: Towards UNIfied-modal Video Generation
Ludan Ruan, Lei Tian, Chuanwei Huang +2
Diffusion based video generation has received extensive attention and achieved considerable success within both the academic and industrial communities. However, current efforts ar…
cs.CV2023★ 1 cited
Accommodating Audio Modality in CLIP for Multimodal Processing
Ludan Ruan, Anwen Hu, Yuqing Song +3
Multimodal processing has attracted much attention lately especially with the success of pre-training. However, the exploration has mainly focused on vision-language pre-training,…