4 papers
Synthetic Speech, Real Signal: Paralinguistic Preservation and Cross-Lingual Augmentation via Voice Cloning
Roseline Polle, Owen Parsons, George Fairs +5
Synthetic data augmentation in speech is common practice for linguistic tasks like ASR, but has seen far less work for paralinguistic ones, especially clinical tasks where labelled…
A multimodal Bayesian Network for symptom-level depression and anxiety prediction from voice and speech data
Agnes Norbury, George Fairs, Alexandra L. Georgescu +3
During psychiatric assessment, clinicians observe not only what patients report, but important nonverbal signs such as tone, speech rate, fluency, responsiveness, and body language…
Meta-Learning Approaches for Speaker-Dependent Voice Fatigue Models
Roseline Polle, Agnes Norbury, Alexandra Livia Georgescu +2
Speaker-dependent modelling can substantially improve performance in speech-based health monitoring applications. While mixed-effect models are commonly used for such speaker adapt…
Path Signature Representation of Patient-Clinician Interactions as a Predictor for Neuropsychological Tests Outcomes in Children: A Proof of Concept
Giulio Falcioni, Alexandra Georgescu, Emilia Molimpakis +3
This research report presents a proof-of-concept study on the application of machine learning techniques to video and speech data collected during diagnostic cognitive assessments…