1 citations · 1 across the 3 of their papers we have counts for
3 papers
F5-TTS-RO: Extending F5-TTS to Romanian TTS via Lightweight Input Adaptation
Radu-Gabriel Chivereanu, Tiberiu Boros
This work introduces a lightweight input-level adapter for the F5-TTS model that enables Romanian Language support. To preserve the existing capabilities of the model (voice clonin…
Aligning Actions and Walking to LLM-Generated Textual Descriptions
Radu Chivereanu, Adrian Cosma, Andy Catruna +2
Large Language Models (LLMs) have demonstrated remarkable capabilities in various domains, including data augmentation and synthetic data generation. This work explores the use of…
Generative Adversarial Training for Text-to-Speech Synthesis Based on Raw Phonetic Input and Explicit Prosody Modelling
Tiberiu Boros, Stefan Daniel Dumitrescu, Ionut Mironica +1
We describe an end-to-end speech synthesis system that uses generative adversarial training. We train our Vocoder for raw phoneme-to-audio conversion, using explicit phonetic, pitc…