3 papers
cs.LG2018
Representation Mixing for TTS Synthesis
Kyle Kastner, João Felipe Santos, Yoshua Bengio +1
Recent character and phoneme-based parametric TTS systems using deep learning have shown strong performance in natural speech generation. However, the choice between character or p…
cs.SD2017
Speech Dereverberation with Context-aware Recurrent Neural Networks
Joao Felipe Santos, Tiago H. Falk
In this paper, we propose a model to perform speech dereverberation by estimating its spectral magnitude from the reverberant counterpart. Our models are capable of extracting feat…
cs.SD2016
Music transcription modelling and composition using deep learning
Bob L. Sturm, João Felipe Santos, Oded Ben-Tal +1
We apply deep learning methods, specifically long short-term memory (LSTM) networks, to music transcription modelling and composition. We build and train LSTM networks using approx…