activity
20242026
collaborators

9 papers

cs.CL2026

COALA: Robust Contextualized Speech-augmented Language Modeling for ASR via Contrastive Regularizer and Biasing Score Estimation

Jhih-Rong Guo, Bi-Cheng Yan, Tien-Hong Lo +1

Contextual biasing seeks to integrate external knowledge into automatic speech recognition (ASR) systems to accurately recognize domain-specific entities. In this paper, we propose…

cs.CL2026

LOPA: Enhancing Spoken Language Assessment via Latent Ordinal Prototype Alignment

Hong-Yun Lin, Fu-An Chao, Bi-Cheng Yan +1

Fueled by increasing model scale and multimodal inputs, Multimodal Large Language Models (MLLMs) have emerged as a promising paradigm for Spoken Language Assessment (SLA). While ef…

cs.CL2026

Probing the Hidden Talent of ASR Foundation Models for L2 English Oral Assessment

Fu-An Chao, Bi-Cheng Yan, Berlin Chen

In this paper, we explore the untapped potential of Whisper, a well-established automatic speech recognition (ASR) foundation model, in the context of L2 spoken language assessment…

eess.AS2025

HiPPO: Exploring A Novel Hierarchical Pronunciation Assessment Approach for Spoken Languages

Bi-Cheng Yan, Hsin-Wei Wang, Fu-An Chao +3

Automatic pronunciation assessment (APA) seeks to quantify a second language (L2) learner's pronunciation proficiency in a target language by offering timely and fine-grained diagn…

cs.CL2025

Multi-task Pretraining for Enhancing Interpretable L2 Pronunciation Assessment

Jiun-Ting Li, Bi-Cheng Yan, Yi-Cheng Wang +1

Automatic pronunciation assessment (APA) analyzes second-language (L2) learners' speech by providing fine-grained pronunciation feedback at various linguistic levels. Most existing…

eess.AS2025

Zero-Shot Text-to-Speech as Golden Speech Generator: A Systematic Framework and its Applicability in Automatic Pronunciation Assessment

Tien-Hong Lo, Meng-Ting Tsai, Yao-Ting Sung +1

Second language (L2) learners can improve their pronunciation by imitating golden speech, especially when the speech that aligns with their respective speech characteristics. This…