9 papers
COALA: Robust Contextualized Speech-augmented Language Modeling for ASR via Contrastive Regularizer and Biasing Score Estimation
Jhih-Rong Guo, Bi-Cheng Yan, Tien-Hong Lo +1
Contextual biasing seeks to integrate external knowledge into automatic speech recognition (ASR) systems to accurately recognize domain-specific entities. In this paper, we propose…
LOPA: Enhancing Spoken Language Assessment via Latent Ordinal Prototype Alignment
Hong-Yun Lin, Fu-An Chao, Bi-Cheng Yan +1
Fueled by increasing model scale and multimodal inputs, Multimodal Large Language Models (MLLMs) have emerged as a promising paradigm for Spoken Language Assessment (SLA). While ef…
Probing the Hidden Talent of ASR Foundation Models for L2 English Oral Assessment
Fu-An Chao, Bi-Cheng Yan, Berlin Chen
In this paper, we explore the untapped potential of Whisper, a well-established automatic speech recognition (ASR) foundation model, in the context of L2 spoken language assessment…
HiPPO: Exploring A Novel Hierarchical Pronunciation Assessment Approach for Spoken Languages
Bi-Cheng Yan, Hsin-Wei Wang, Fu-An Chao +3
Automatic pronunciation assessment (APA) seeks to quantify a second language (L2) learner's pronunciation proficiency in a target language by offering timely and fine-grained diagn…
Multi-task Pretraining for Enhancing Interpretable L2 Pronunciation Assessment
Jiun-Ting Li, Bi-Cheng Yan, Yi-Cheng Wang +1
Automatic pronunciation assessment (APA) analyzes second-language (L2) learners' speech by providing fine-grained pronunciation feedback at various linguistic levels. Most existing…
Zero-Shot Text-to-Speech as Golden Speech Generator: A Systematic Framework and its Applicability in Automatic Pronunciation Assessment
Tien-Hong Lo, Meng-Ting Tsai, Yao-Ting Sung +1
Second language (L2) learners can improve their pronunciation by imitating golden speech, especially when the speech that aligns with their respective speech characteristics. This…