7 citations · 19 across the 8 of their papers we have counts for
9 papers
Adding Connectionist Temporal Summarization into Conformer to Improve Its Decoder Efficiency For Speech Recognition
Nick J. C. Wang, Zongfeng Quan, Shaojun Wang +1
The Conformer model is an excellent architecture for speech recognition modeling that effectively utilizes the hybrid losses of connectionist temporal classification (CTC) and atte…
A Study of Different Ways to Use The Conformer Model For Spoken Language Understanding
Nick J. C. Wang, Shaojun Wang, Jing Xiao
SLU combines ASR and NLU capabilities to accomplish speech-to-intent understanding. In this paper, we compare different ways to combine ASR and NLU, in particular using a single Co…
EfficientTTS: An Efficient and High-Quality Text-to-Speech Architecture
Chenfeng Miao, Shuang Liang, Zhencheng Liu +4
In this work, we address the Text-to-Speech (TTS) task by proposing a non-autoregressive architecture called EfficientTTS. Unlike the dominant non-autoregressive TTS models, which…
BS-NAS: Broadening-and-Shrinking One-Shot NAS with Searchable Numbers of Channels
Zan Shen, Jiang Qian, Bojin Zhuang +2
One-Shot methods have evolved into one of the most popular methods in Neural Architecture Search (NAS) due to weight sharing and single training of a supernet. However, existing me…
An Iterative Polishing Framework based on Quality Aware Masked Language Model for Chinese Poetry Generation
Liming Deng, Jie Wang, Hangming Liang +5
Owing to its unique literal and aesthetical characteristics, automatic generation of Chinese poetry is still challenging in Artificial Intelligence, which can hardly be straightfor…
Audio-Based Music Classification with DenseNet And Data Augmentation
Wenhao Bian, Jie Wang, Bojin Zhuang +3
In recent years, deep learning technique has received intense attention owing to its great success in image recognition. A tendency of adaption of deep learning in various informat…