9 citations · 16 across the 6 of their papers we have counts for
4 papers · 1 filter
Cloud-based Automatic Speech Recognition Systems for Southeast Asian Languages
Lei Wang, Rong Tong, Cheung Chi Leung +3
This paper provides an overall introduction of our Automatic Speech Recognition (ASR) systems for Southeast Asian languages. As not much existing work has been carried out on such…
Independent language modeling architecture for end-to-end ASR
Van Tung Pham, Haihua Xu, Yerbolat Khassanov +5
The attention-based end-to-end (E2E) automatic speech recognition (ASR) architecture allows for joint optimization of acoustic and language models within a single network. However,…
Constrained Output Embeddings for End-to-End Code-Switching Speech Recognition with Only Monolingual Data
Yerbolat Khassanov, Haihua Xu, Van Tung Pham +4
The lack of code-switch training data is one of the major concerns in the development of end-to-end code-switching automatic speech recognition (ASR) models. In this work, we propo…
Learning Acoustic Word Embeddings with Temporal Context for Query-by-Example Speech Search
Yougen Yuan, Cheung-Chi Leung, Lei Xie +3
We propose to learn acoustic word embeddings with temporal context for query-by-example (QbE) speech search. The temporal context includes the leading and trailing word sequences o…