11 citations · 11 across the 1 of their papers we have counts for
1 paper
Jacob J Webber, Cassia Valentini-Botinhao, Evelyn Williams +2
Most state-of-the-art Text-to-Speech systems use the mel-spectrogram as an intermediate representation, to decompose the task into acoustic modelling and waveform generation. A mel…