papers

Publications (15)

eess.AS2024

VoxBlink2: A 100K+ Speaker Recognition Corpus and the Open-Set Speaker-Identification Benchmark

Yuke Lin, Ming Cheng, Fulin Zhang +3

In this paper, we provide a large audio-visual speaker recognition dataset, VoxBlink2, which includes approximately 10M utterances with videos from 110K+ speakers in the wild. This…

eess.AS2023

VoxBlink: A Large Scale Speaker Verification Dataset on Camera

Yuke Lin, Xiaoyi Qin, Guoqing Zhao +4

In this paper, we introduce a large-scale and high-quality audio-visual speaker verification dataset, named VoxBlink. We propose an innovative and robust automatic audio-visual dat…

cs.CL2025

Easy Turn: Integrating Acoustic and Linguistic Modalities for Robust Turn-Taking in Full-Duplex Spoken Dialogue Systems

Guojian Li, Chengyou Wang, Hongfei Xue +8

Full-duplex interaction is crucial for natural human-machine communication, yet remains challenging as it requires robust turn-taking detection to decide when the system should spe…

eess.AS2023

Haha-Pod: An Attempt for Laughter-based Non-Verbal Speaker Verification

Yuke Lin, Xiaoyi Qin, Ning Jiang +2

It is widely acknowledged that discriminative representation for speaker verification can be extracted from verbal speech. However, how much speaker information that non-verbal voc…

eess.AS2024

KunquDB: An Attempt for Speaker Verification in the Chinese Opera Scenario

Huali Zhou, Yuke Lin, Dong Liu +1

This work aims to promote Chinese opera research in both musical and speech domains, with a primary focus on overcoming the data limitations. We introduce KunquDB, a relatively lar…

eess.AS2024

The Database and Benchmark for the Source Speaker Tracing Challenge 2024

Ze Li, Yuke Lin, Tian Yao +6

Voice conversion (VC) systems can transform audio to mimic another speaker's voice, thereby attacking speaker verification (SV) systems. However, ongoing studies on source speaker…