most citedThe VoiceMOS Challenge 2023: Zero-shot Subjective Speech Quality Prediction for Multiple Domains

2 citations · 5 across the 7 of their papers we have counts for

collaborators

7 papers

eess.AS2023

Partial Rank Similarity Minimization Method for Quality MOS Prediction of Unseen Speech Synthesis Systems in Zero-Shot and Semi-supervised setting

Hemant Yadav, Erica Cooper, Junichi Yamagishi +2

This paper introduces a novel objective function for quality mean opinion score (MOS) prediction of unseen speech synthesis systems. The proposed function measures the similarity o…

eess.AS20232 cited

The VoiceMOS Challenge 2023: Zero-shot Subjective Speech Quality Prediction for Multiple Domains

Erica Cooper, Wen-Chin Huang, Yu Tsao +3

We present the second edition of the VoiceMOS Challenge, a scientific event that aims to promote the study of automatic prediction of the mean opinion score (MOS) of synthesized an…

cs.SD20231 cited

DDSP-based Neural Waveform Synthesis of Polyphonic Guitar Performance from String-wise MIDI Input

Nicolas Jonason, Xin Wang, Erica Cooper +3

We explore the use of neural synthesis for acoustic guitar from string-wise MIDI input. We propose four different systems and compare them with both objective metrics and subjectiv…

cs.SD2023

SynVox2: Towards a privacy-friendly VoxCeleb2 dataset

Xiaoxiao Miao, Xin Wang, Erica Cooper +5

The success of deep learning in speaker recognition relies heavily on the use of large datasets. However, the data-hungry nature of deep learning methods has already being question…

cs.SD20231 cited

Range-Based Equal Error Rate for Spoof Localization

Lin Zhang, Xin Wang, Erica Cooper +2

Spoof localization, also called segment-level detection, is a crucial task that aims to locate spoofs in partially spoofed audio. The equal error rate (EER) is widely used to measu…

eess.AS2023

Improving Generalization Ability of Countermeasures for New Mismatch Scenario by Combining Multiple Advanced Regularization Terms

Chang Zeng, Xin Wang, Xiaoxiao Miao +2

The ability of countermeasure models to generalize from seen speech synthesis methods to unseen ones has been investigated in the ASVspoof challenge. However, a new mismatch scenar…