3 citations · 12 across the 12 of their papers we have counts for
19 papers
Scaling Neural Face Synthesis to High FPS and Low Latency by Neural Caching
Frank Yu, Sid Fels, Helge Rhodin
Recent neural rendering approaches greatly improve image quality, reaching near photorealism. However, the underlying neural networks have high runtime, precluding telepresence and…
A comparative study of two-dimensional vocal tract acoustic modeling based on Finite-Difference Time-Domain methods
Debasish Ray Mohapatra, Victor Zappi, Sidney Fels
The two-dimensional (2D) numerical approaches for vocal tract (VT) modelling can afford a better balance between the low computational cost and accurate rendering of acoustic wave…
SPEAK WITH YOUR HANDS Using Continuous Hand Gestures to control Articulatory Speech Synthesizer
Pramit Saha, Debasish Ray Mohapatra, Sidney Fels
This work presents our advancements in controlling an articulatory speech synthesis engine, \textit{viz.}, Pink Trombone, with hand gestures. Our interface translates continuous fi…
Ultra2Speech -- A Deep Learning Framework for Formant Frequency Estimation and Tracking from Ultrasound Tongue Images
Pramit Saha, Yadong Liu, Bryan Gick +1
Thousands of individuals need surgical removal of their larynx due to critical diseases every year and therefore, require an alternative form of communication to articulate speech…
Learning Joint Articulatory-Acoustic Representations with Normalizing Flows
Pramit Saha, Sidney Fels
The articulatory geometric configurations of the vocal tract and the acoustic properties of the resultant speech sound are considered to have a strong causal relationship. This pap…
Variational Learning with Disentanglement-PyTorch
Amir H. Abdi, Purang Abolmaesumi, Sidney Fels
Unsupervised learning of disentangled representations is an open problem in machine learning. The Disentanglement-PyTorch library is developed to facilitate research, implementatio…