3 papers
cs.CL2025
Comparison of End-to-end Speech Assessment Models for the NOCASA 2025 Challenge
Aleksei Žavoronkov, Tanel Alumäe
This paper presents an analysis of three end-to-end models developed for the NOCASA 2025 Challenge, aimed at automatic word-level pronunciation assessment for children learning Nor…
cs.CL2025
TalTech Systems for the Interspeech 2025 ML-SUPERB 2.0 Challenge
Tanel Alumäe, Artem Fedorchenko
This paper describes the language identification and multilingual speech recognition system developed at Tallinn University of Technology for the Interspeech 2025 ML-SUPERB 2.0 Cha…
cs.CL2025
Optimizing Estonian TV Subtitles with Semi-supervised Learning and LLMs
Artem Fedorchenko, Tanel Alumäe
This paper presents an approach for generating high-quality, same-language subtitles for Estonian TV content. We fine-tune the Whisper model on human-generated Estonian subtitles a…