1 citations · 1 across the 21 of their papers we have counts for
1 paper · 1 filter
Yuchen Sun, Qian Yang, Jun Wang +4
Text-to-audio-video (T2AV) generation has advanced rapidly, but its evaluation still underestimates the audio modality. Existing benchmarks either treat audio as an auxiliary compo…