1 citations · 1 across the 8 of their papers we have counts for
1 paper · 1 filter
Yusheng Dai, Zehua Chen, Yuxuan Jiang +4
Training a unified model integrating video-to-audio (V2A), text-to-audio (T2A), and joint video-text-to-audio (VT2A) generation offers significant application flexibility, yet face…