1 citations · 1 across the 2 of their papers we have counts for
3 papers · 1 filter
Real-Time MRI Video synthesis from time aligned phonemes with sequence-to-sequence networks
Sathvik Udupa, Prasanta Kumar Ghosh
Real-Time Magnetic resonance imaging (rtMRI) of the midsagittal plane of the mouth is of interest for speech production research. In this work, we focus on estimating utterance lev…
Improved acoustic-to-articulatory inversion using representations from pretrained self-supervised learning models
Sathvik Udupa, Siddarth C, Prasanta Kumar Ghosh
In this work, we investigate the effectiveness of pretrained Self-Supervised Learning (SSL) features for learning the mapping for acoustic to articulatory inversion (AAI). Signal p…
Estimating articulatory movements in speech production with transformer networks
Sathvik Udupa, Anwesha Roy, Abhayjeet Singh +2
We estimate articulatory movements in speech production from different modalities - acoustics and phonemes. Acoustic-to articulatory inversion (AAI) is a sequence-to-sequence task.…