4 papers · 1 filter
A corpus-based investigation of pitch contours of monosyllabic words in conversational Taiwan Mandarin
Xiaoyun Jin, Mirjam Ernestus, R. Harald Baayen
In Mandarin, the tonal contours of monosyllabic words produced in isolation or in careful speech are characterized by four lexical tones: a high-level tone (T1), a rising tone (T2)…
Modelling word learning and recognition using visually grounded speech
Danny Merkx, Sebastiaan Scholten, Stefan L. Frank +2
Background: Computational models of speech recognition often assume that the set of target words is already given. This implies that these models do not learn to recognise speech f…
Seeing the advantage: visually grounding word embeddings to better capture human semantic knowledge
Danny Merkx, Stefan L. Frank, Mirjam Ernestus
Distributional semantic models capture word-level meaning that is useful in many natural language processing tasks and have even been shown to capture cognitive aspects of word mea…
Language learning using Speech to Image retrieval
Danny Merkx, Stefan L. Frank, Mirjam Ernestus
Humans learn language by interaction with their environment and listening to other humans. It should also be possible for computational models to learn language directly from speec…