5 papers
Speech Quality Embeddings for Improved Detection and Classification of Degradations in Speech Signals
Michael Kuhlmann, Tobias Cord-Landwehr, Reinhold Haeb-Umbach
Automatic subjective speech quality assessment (SSQA) traditionally estimates speech quality on an utterance or system level. While this resolution was adequate for older transmiss…
Speech Quality-Based Localization of Low-Quality Speech and Text-to-Speech Synthesis Artefacts
Michael Kuhlmann, Alexander Werning, Thilo von Neumann +1
A large number of works view the automatic assessment of speech from an utterance- or system-level perspective. While such approaches are good in judging overall quality, they cann…
Synthesizing speech with selected perceptual voice qualities - A case study with creaky voice
Frederik Rautenberg, Fritz Seebauer, Jana Wiechmann +3
The control of perceptual voice qualities in a text-to-speech (TTS) system is of interest for applications where unmanipu- lated and manipulated speech probes can serve to illustra…
Towards Frame-level Quality Predictions of Synthetic Speech
Michael Kuhlmann, Fritz Seebauer, Petra Wagner +1
While automatic subjective speech quality assessment has witnessed much progress, an open question is whether an automatic quality assessment at frame resolution is possible. This…
Speech Synthesis along Perceptual Voice Quality Dimensions
Frederik Rautenberg, Michael Kuhlmann, Fritz Seebauer +3
While expressive speech synthesis or voice conversion systems mainly focus on controlling or manipulating abstract prosodic characteristics of speech, such as emotion or accent, we…