1 citations · 1 across the 6 of their papers we have counts for
1 paper · 1 filter
Yuke Lin, Ming Cheng, Fulin Zhang +3
In this paper, we provide a large audio-visual speaker recognition dataset, VoxBlink2, which includes approximately 10M utterances with videos from 110K+ speakers in the wild. This…