2 citations · 4 across the 3 of their papers we have counts for
6 papers
Analysis of Voice Conversion and Code-Switching Synthesis Using VQ-VAE
Shuvayanti Das, Jennifer Williams, Catherine Lai
This paper presents an analysis of speech synthesis quality achieved by simultaneously performing voice conversion and language code-switching using multilingual VQ-VAE speech synt…
A Cross-Domain Approach for Continuous Impression Recognition from Dyadic Audio-Visual-Physio Signals
Yuanchao Li, Catherine Lai
The impression we make on others depends not only on what we say, but also, to a large extent, on how we say it. As a sub-branch of affective computing and social signal processing…
Location, Location: Enhancing the Evaluation of Text-to-Speech Synthesis Using the Rapid Prosody Transcription Paradigm
Elijah Gutierrez, Pilar Oplustil-Gallegos, Catherine Lai
Text-to-Speech synthesis systems are generally evaluated using Mean Opinion Score (MOS) tests, where listeners score samples of synthetic speech on a Likert scale. A major drawback…
It's not what you said, it's how you said it: discriminative perception of speech as a multichannel communication system
Sarenne Wallbridge, Peter Bell, Catherine Lai
People convey information extremely effectively through spoken interaction using multiple channels of information transmission: the lexical channel of what is said, and the non-lex…
Perception of prosodic variation for speech synthesis using an unsupervised discrete representation of F0
Zack Hodari, Catherine Lai, Simon King
In English, prosody adds a broad range of information to segment sequences, from information structure (e.g. contrast) to stylistic variation (e.g. expression of emotion). However,…
Polarity and Intensity: the Two Aspects of Sentiment Analysis
Leimin Tian, Catherine Lai, Johanna D. Moore
Current multimodal sentiment analysis frames sentiment score prediction as a general Machine Learning task. However, what the sentiment score actually represents has often been ove…