3 papers
eess.AS2026
An Evaluation Framework for Text-to-Speech Voice Reconstruction
Ariadna Sanchez, Christoph Minixhofer, Korin Richmond +3
Voice reconstruction using Text-to-Speech (TTS) offers a communication method for people with speech disorders, which aims to retain their speaker identity while improving intellig…
eess.AS2025
Exploring Acoustic Similarity in Emotional Speech and Music via Self-Supervised Representations
Yujia Sun, Zeyu Zhao, Korin Richmond +1
Emotion recognition from speech and music shares similarities due to their acoustic overlap, which has led to interest in transferring knowledge between these domains. However, the…
eess.AS2025
Cross-Lingual Speech Emotion Recognition: Humans vs. Self-Supervised Models
Zhichen Han, Tianqi Geng, Hui Feng +3
Utilizing Self-Supervised Learning (SSL) models for Speech Emotion Recognition (SER) has proven effective, yet limited research has explored cross-lingual scenarios. This study pre…