3 papers
cs.SD2024
Zero-Shot vs. Few-Shot Multi-Speaker TTS Using Pre-trained Czech SpeechT5 Model
Jan LeheÄka, ZdenÄk HanzlÃÄek, JindÅich MatouÅ¡ek +1
In this paper, we experimented with the SpeechT5 model pre-trained on large-scale datasets. We pre-trained the foundation model from scratch and fine-tuned it on a large-scale robu…
cs.CL2024
A Comparative Analysis of Bilingual and Trilingual Wav2Vec Models for Automatic Speech Recognition in Multilingual Oral History Archives
Jan LeheÄka, Josef V. Psutka, LuboÅ¡ Å mÃdl +2
In this paper, we are comparing monolingual Wav2Vec 2.0 models with various multilingual models to see whether we could improve speech recognition performance on a unique oral hist…
cs.SD2024
Speech Technology Services for Oral History Research
Christoph Draxler, Henk van den Heuvel, Arjan van Hessen +2
Oral history is about oral sources of witnesses and commentors on historical events. Speech technology is an important instrument to process such recordings in order to obtain tran…