Showing cs.CLShow all
3 papers · 1 filter
cs.CL2024
Just Because We Camp, Doesn't Mean We Should: The Ethics of Modelling Queer Voices
Atli Sigurgeirsson, Eddie L. Ungless
Modern voice cloning models claim to be able to capture a diverse range of voices. We test the ability of a typical pipeline to capture the style known colloquially as "gay voice"…
cs.CL2024
A Human-in-the-Loop Approach to Improving Cross-Text Prosody Transfer
Himanshu Maurya, Atli Sigurgeirsson
Text-To-Speech (TTS) prosody transfer models can generate varied prosodic renditions, for the same text, by conditioning on a reference utterance. These models are trained with a r…
cs.CL2023
Controllable Speaking Styles Using a Large Language Model
Atli Thor Sigurgeirsson, Simon King
Reference-based Text-to-Speech (TTS) models can generate multiple, prosodically-different renditions of the same target text. Such models jointly learn a latent acoustic space duri…