105 citations · 109 across the 10 of their papers we have counts for
6 papers · 1 filter
TIPAA-SSL: Text Independent Phone-to-Audio Alignment based on Self-Supervised Learning and Knowledge Transfer
Noé Tits, Prernna Bhatnagar, Thierry Dutoit
In this paper, we present a novel approach for text independent phone-to-audio alignment based on phoneme recognition, representation learning and knowledge transfer. Our method le…
ICE-Talk: an Interface for a Controllable Expressive Talking Machine
Noé Tits, Kevin El Haddad, Thierry Dutoit
ICE-Talk is an open source web-based GUI that allows the use of a TTS system with controllable parameters via a text field and a clickable 2D plot. It enables the study of latent s…
Laughter Synthesis: Combining Seq2seq modeling with Transfer Learning
Noé Tits, Kevin El Haddad, Thierry Dutoit
Despite the growing interest for expressive speech synthesis, synthesis of nonverbal expressions is an under-explored area. In this paper we propose an audio laughter synthesis sys…
The Theory behind Controllable Expressive Speech Synthesis: a Cross-disciplinary Approach
Noé Tits, Kevin El Haddad, Thierry Dutoit
As part of the Human-Computer Interaction field, Expressive speech synthesis is a very rich domain as it requires knowledge in areas such as machine learning, signal processing, so…
A Methodology for Controlling the Emotional Expressiveness in Synthetic Speech -- a Deep Learning approach
Noé Tits
In this project, we aim to build a Text-to-Speech system able to produce speech with a controllable emotional expressiveness. We propose a methodology for solving this problem in t…
ASR-based Features for Emotion Recognition: A Transfer Learning Approach
Noé Tits, Kevin El Haddad, Thierry Dutoit
During the last decade, the applications of signal processing have drastically improved with deep learning. However areas of affecting computing such as emotional speech synthesis…