3 citations · 5 across the 21 of their papers we have counts for
1 paper · 1 filter
Ali Vosoughi, Jing Bi, Pinxin Liu +2
What happens when we push audio-visual alignment to its absolute limits? To systematically investigate this question, we needed datasets with granular alignment quality annotations…