7 citations · 10 across the 4 of their papers we have counts for
5 papers · 1 filter
An Improved Single Step Non-autoregressive Transformer for Automatic Speech Recognition
Ruchao Fan, Wei Chu, Peng Chang +2
Non-autoregressive mechanisms can significantly decrease inference time for speech transformers, especially when the single step variant is applied. Previous work on CTC alignment-…
Low Resource German ASR with Untranscribed Data Spoken by Non-native Children -- INTERSPEECH 2021 Shared Task SPAPL System
Jinhan Wang, Yunzheng Zhu, Ruchao Fan +2
This paper describes the SPAPL system for the INTERSPEECH 2021 Challenge: Shared Task on Automatic Speech Recognition for Non-Native Children's Speech in German. ~ 5 hours of trans…
CASS-NAT: CTC Alignment-based Single Step Non-autoregressive Transformer for Speech Recognition
Ruchao Fan, Wei Chu, Peng Chang +1
We propose a CTC alignment-based single step non-autoregressive transformer (CASS-NAT) for speech recognition. Specifically, the CTC alignment contains the information of (a) the n…
Singing voice conversion with non-parallel data
Xin Chen, Wei Chu, Jinxi Guo +1
Singing voice conversion is a task to convert a song sang by a source singer to the voice of a target singer. In this paper, we propose using a parallel data free, many-to-one voic…
Hybrid CTC-Attention based End-to-End Speech Recognition using Subword Units
Zhangyu Xiao, Zhijian Ou, Wei Chu +1
In this paper, we present an end-to-end automatic speech recognition system, which successfully employs subword units in a hybrid CTC-Attention based system. The subword units are…