4 papers
CVSS-X: A Multilingual Speech-to-Speech Translation Corpus for 28 Languages
Lucas Rafael Stefanel Gris, Alef Iury Siqueira Ferreira, Frederico Santos de Oliveira +4
We introduce CVSS-X, a large-scale synthetic speech-to-speech translation corpus that extends CVSS by reversing the translation direction. While CVSS translates from 21 languages i…
Yin Yang Convolutional Nets: Image Manifold Extraction by the Analysis of Opposites
Augusto Seben da Rosa, Frederico Santos de Oliveira, Anderson da Silva Soares +1
Computer vision in general presented several advances such as training optimizations, new architectures (pure attention, efficient block, vision language models, generative models,…
CML-TTS A Multilingual Dataset for Speech Synthesis in Low-Resource Languages
Frederico S. Oliveira, Edresson Casanova, Arnaldo Cândido Júnior +2
In this paper, we present CML-TTS, a recursive acronym for CML-Multi-Lingual-TTS, a new Text-to-Speech (TTS) dataset developed at the Center of Excellence in Artificial Intelligenc…
Evaluation of Speech Representations for MOS prediction
Frederico S. Oliveira, Edresson Casanova, Arnaldo Cândido Júnior +3
In this paper, we evaluate feature extraction models for predicting speech quality. We also propose a model architecture to compare embeddings of supervised learning and self-super…