13 citations · 75 across the 46 of their papers we have counts for
5 papers · 1 filter
CNVSRC 2023: The First Chinese Continuous Visual Speech Recognition Challenge
Chen Chen, Zehua Liu, Xiaolou Li +2
The first Chinese Continuous Visual Speech Recognition Challenge aimed to probe the performance of Large Vocabulary Continuous Visual Speech Recognition (LVC-VSR) on two tasks: (1)…
A Glance is Enough: Extract Target Sentence By Looking at A keyword
Ying Shi, Dong Wang, Lantian Li +1
This paper investigates the possibility of extracting a target sentence from multi-talker speech using only a keyword as input. For example, in social security applications, the ke…
Phone-aware Neural Language Identification
Zhiyuan Tang, Dong Wang, Yixiang Chen +2
Pure acoustic neural models, particularly the LSTM-RNN model, have shown great potential in language identification (LID). However, the phonetic information has been largely overlo…
Improved Deep Speaker Feature Learning for Text-Dependent Speaker Recognition
Lantian Li, Yiye Lin, Zhiyong Zhang +1
A deep learning approach has been proposed recently to derive speaker identifies (d-vector) by a deep neural network (DNN). This approach has been applied to text-dependent speaker…
Deep Speaker Vectors for Semi Text-independent Speaker Verification
Lantian Li, Dong Wang, Zhiyong Zhang +1
Recent research shows that deep neural networks (DNNs) can be used to extract deep speaker vectors (d-vectors) that preserve speaker characteristics and can be used in speaker veri…