1 paper
Hanmo Chen, Chengcheng Liu, Tianxiao Chen +7
Recent joint video-audio generation models have achieved strong semantic correspondence and temporal synchronization. However, applications such as AR/VR and interactive gaming fur…