activity
20212026
most citedThe Singing Voice Conversion Challenge 2023

4 citations · 8 across the 11 of their papers we have counts for

collaborators
Showing cs.SDShow all

10 papers · 1 filter

cs.SD2025

An Extensive Analysis of the Singing Voice Conversion Challenge 2025 Evaluation Results

Lester Phillip Violeta, Xueyao Zhang, Jiatong Shi +4

We present a thorough analysis of the findings of the latest iteration of the Singing Voice Conversion Challenge, a scientific event aiming to compare and understand different voic…

cs.SD2025

Serenade: A Singing Style Conversion Framework Based On Audio Infilling

Lester Phillip Violeta, Wen-Chin Huang, Tomoki Toda

We propose Serenade, a novel framework for the singing style conversion (SSC) task. Although singer identity conversion has made great strides in the previous years, converting the…

cs.SD2024

A Preliminary Investigation on Flexible Singing Voice Synthesis Through Decomposed Framework with Inferrable Features

Lester Phillip Violeta, Taketo Akama

We investigate the feasibility of a singing voice synthesis (SVS) system by using a decomposed framework to improve flexibility in generating singing voices. Due to data-driven app…

cs.SD2024

Quantifying the effect of speech pathology on automatic and human speaker verification

Bence Mark Halpern, Thomas Tienkamp, Wen-Chin Huang +7

This study investigates how surgical intervention for speech pathology (specifically, as a result of oral cancer surgery) impacts the performance of an automatic speaker verificati…

cs.SD2023

Improving severity preservation of healthy-to-pathological voice conversion with global style tokens

Bence Mark Halpern, Wen-Chin Huang, Lester Phillip Violeta +2

In healthy-to-pathological voice conversion (H2P-VC), healthy speech is converted into pathological while preserving the identity. The paper improves on previous two-stage approach…

cs.SD2023

Electrolaryngeal Speech Intelligibility Enhancement Through Robust Linguistic Encoders

Lester Phillip Violeta, Wen-Chin Huang, Ding Ma +3

We propose a novel framework for electrolaryngeal speech intelligibility enhancement through the use of robust linguistic encoders. Pretraining and fine-tuning approaches have prov…