4 papers
Lightweight Prompt Biasing for Contextualized End-to-End ASR Systems
Bo Ren, Yu Shi, Jinyu Li
End-to-End Automatic Speech Recognition (ASR) has advanced significantly yet still struggles with rare and domain-specific entities. This paper introduces a simple yet efficient pr…
Universal Preference-Score-based Pairwise Speech Quality Assessment
Yu-Fei Shi, Yang Ai, Zhen-Hua Ling
To compare the performance of two speech generation systems, one of the most effective approaches is estimating the preference score between their generated speech. This paper prop…
Pitch-and-Spectrum-Aware Singing Quality Assessment with Bias Correction and Model Fusion
Yu-Fei Shi, Yang Ai, Ye-Xin Lu +2
We participated in track 2 of the VoiceMOS Challenge 2024, which aimed to predict the mean opinion score (MOS) of singing samples. Our submission secured the first place among all…
SAMOS: A Neural MOS Prediction Model Leveraging Semantic Representations and Acoustic Features
Yu-Fei Shi, Yang Ai, Ye-Xin Lu +2
Assessing the naturalness of speech using mean opinion score (MOS) prediction models has positive implications for the automatic evaluation of speech synthesis systems. Early MOS p…