9 citations · 12 across the 3 of their papers we have counts for
5 papers
Automatic recognition of suprasegmentals in speech
Jiahong Yuan, Neville Ryant, Xingyu Cai +2
This study reports our efforts to improve automatic recognition of suprasegmentals by fine-tuning wav2vec 2.0 with CTC, a method that has been successful in automatic speech recogn…
The Role of Phonetic Units in Speech Emotion Recognition
Jiahong Yuan, Xingyu Cai, Renjie Zheng +2
We propose a method for emotion recognition through emotiondependent speech recognition using Wav2vec 2.0. Our method achieved a significant improvement over most previously report…
Decoupling recognition and transcription in Mandarin ASR
Jiahong Yuan, Xingyu Cai, Dongji Gao +3
Much of the recent literature on automatic speech recognition (ASR) is taking an end-to-end approach. Unlike English where the writing system is closely related to sound, Chinese c…
Fluent and Low-latency Simultaneous Speech-to-Speech Translation with Self-adaptive Training
Renjie Zheng, Mingbo Ma, Baigong Zheng +4
Simultaneous speech-to-speech translation is widely useful but extremely challenging, since it needs to generate target-language speech concurrently with the source-language speech…
On the Role of Style in Parsing Speech with Neural Models
Trang Tran, Jiahong Yuan, Yang Liu +1
The differences in written text and conversational speech are substantial; previous parsers trained on treebanked text have given very poor results on spontaneous speech. For spoke…