12 citations · 18 across the 8 of their papers we have counts for
8 papers
End-to-end Online Speaker Diarization with Target Speaker Tracking
Weiqing Wang, Ming Li
This paper proposes an online target speaker voice activity detection system for speaker diarization tasks, which does not require a priori knowledge from the clustering-based diar…
ChatRadio-Valuer: A Chat Large Language Model for Generalizable Radiology Report Generation Based on Multi-institution and Multi-system Data
Tianyang Zhong, Wei Zhao, Yutong Zhang +39
Radiology report generation, as a key step in medical image analysis, is critical to the quantitative analysis of clinically informed decision-making levels. However, complex and d…
Haha-Pod: An Attempt for Laughter-based Non-Verbal Speaker Verification
Yuke Lin, Xiaoyi Qin, Ning Jiang +2
It is widely acknowledged that discriminative representation for speaker verification can be extracted from verbal speech. However, how much speaker information that non-verbal voc…
Multi-objective Progressive Clustering for Semi-supervised Domain Adaptation in Speaker Verification
Ze Li, Yuke Lin, Ning Jiang +4
Utilizing the pseudo-labeling algorithm with large-scale unlabeled data becomes crucial for semi-supervised domain adaptation in speaker verification tasks. In this paper, we propo…
VoiceLens: Controllable Speaker Generation and Editing with Flow
Yao Shi, Ming Li
Currently, many multi-speaker speech synthesis and voice conversion systems address speaker variations with an embedding vector. Modeling it directly allows new voices outside of t…
The DKU-MSXF Speaker Verification System for the VoxCeleb Speaker Recognition Challenge 2023
Ze Li, Yuke Lin, Xiaoyi Qin +3
This paper is the system description of the DKU-MSXF System for the track1, track2 and track3 of the VoxCeleb Speaker Recognition Challenge 2023 (VoxSRC-23). For Track 1, we utiliz…