55 citations · 70 across the 15 of their papers we have counts for
6 papers · 1 filter
Improving child speech recognition with augmented child-like speech
Yuanyuan Zhang, Zhengjun Yue, Tanvina Patel +1
State-of-the-art ASRs show suboptimal performance for child speech. The scarcity of child speech limits the development of child speech recognition (CSR). Therefore, we studied chi…
Using Data Augmentations and VTLN to Reduce Bias in Dutch End-to-End Speech Recognition Systems
Tanvina Patel, Odette Scharenborg
Speech technology has improved greatly for norm speakers, i.e., adult native speakers of a language without speech impediments or strong accents. However, non-norm or diverse speak…
Modelling word learning and recognition using visually grounded speech
Danny Merkx, Sebastiaan Scholten, Stefan L. Frank +2
Background: Computational models of speech recognition often assume that the set of target words is already given. This implies that these models do not learn to recognise speech f…
Evaluating Automatically Generated Phoneme Captions for Images
Justin van der Hout, Zoltán D'Haese, Mark Hasegawa-Johnson +1
Image2Speech is the relatively new task of generating a spoken description of an image. This paper presents an investigation into the evaluation of this task. For this, first an Im…
Bayesian Models for Unit Discovery on a Very Low Resource Language
Lucas Ondel, Pierre Godard, Laurent Besacier +7
Developing speech technologies for low-resource languages has become a very active research field over the last decade. Among others, Bayesian models have shown some promising resu…
Linguistic unit discovery from multi-modal inputs in unwritten languages: Summary of the "Speaking Rosetta" JSALT 2017 Workshop
Odette Scharenborg, Laurent Besacier, Alan Black +16
We summarize the accomplishments of a multi-disciplinary workshop exploring the computational and scientific issues surrounding the discovery of linguistic units (subwords and word…