most citedThe Edinburgh International Accents of English Corpus: Towards the Democratization of English ASR

1 citations · 1 across the 5 of their papers we have counts for

collaborators

5 papers

cs.SD2023

Speech collage: code-switched audio generation by collaging monolingual corpora

Amir Hussein, Dorsa Zeinali, Ondřej Klejch +6

Designing effective automatic speech recognition (ASR) systems for Code-Switching (CS) often depends on the availability of the transcribed CS resources. To address data scarcity,…

cs.CL2023

Acoustic Word Embeddings for Untranscribed Target Languages with Continued Pretraining and Learned Pooling

Ramon Sanabria, Ondrej Klejch, Hao Tang +1

Acoustic word embeddings are typically created by training a pooling function using pairs of word-like units. For unsupervised systems, these are mined using k-nearest neighbor (KN…

eess.AS2023

ASR and Emotional Speech: A Word-Level Investigation of the Mutual Impact of Speech and Emotion Recognition

Yuanchao Li, Zeyu Zhao, Ondrej Klejch +2

In Speech Emotion Recognition (SER), textual data is often used alongside audio signals to address their inherent variability. However, the reliance on human annotated text in most…

cs.CL20231 cited

The Edinburgh International Accents of English Corpus: Towards the Democratization of English ASR

Ramon Sanabria, Nikolay Bogoychev, Nina Markl +3

English is the most widely spoken language in the world, used daily by millions of people as a first or second language in many different contexts. As a result, there are many vari…

cs.CL2021

Mask-combine Decoding and Classification Approach for Punctuation Prediction with real-time Inference Constraints

Christoph Minixhofer, Ondřej Klejch, Peter Bell

In this work, we unify several existing decoding strategies for punctuation prediction in one framework and introduce a novel strategy which utilises multiple predictions at each w…