3 papers
eess.AS2022
openFEAT: Improving Speaker Identification by Open-set Few-shot Embedding Adaptation with Transformer
Kishan K C, Zhenning Tan, Long Chen +4
Household speaker identification with few enrollment utterances is an important yet challenging problem, especially when household members share similar voice characteristics and r…
cs.CL2021
End-to-end Neural Diarization: From Transformer to Conformer
Yi Chieh Liu, Eunjung Han, Chul Lee +1
We propose a new end-to-end neural diarization (EEND) system that is based on Conformer, a recently proposed neural architecture that combines convolutional mappings and Transforme…
cs.SD2020
BW-EDA-EEND: Streaming End-to-End Neural Speaker Diarization for a Variable Number of Speakers
Eunjung Han, Chul Lee, Andreas Stolcke
We present a novel online end-to-end neural diarization system, BW-EDA-EEND, that processes data incrementally for a variable number of speakers. The system is based on the Encoder…