1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2024
Multimodal Segmentation for Vocal Tract Modeling
Rishi Jain, Bohan Yu, Peter Wu +2
Accurate modeling of the vocal tract is necessary to construct articulatory representations for interpretable speech processing and linguistics. However, vocal tract modeling is ch…
cs.SD2023
Towards Streaming Speech-to-Avatar Synthesis
Tejas S. Prabhune, Peter Wu, Bohan Yu +1
Streaming speech-to-avatar synthesis creates real-time animations for a virtual character from audio data. Accurate avatar representations of speech are important for the visualiza…
eess.IV2022★ 1 cited
Contrastive learning-based pretraining improves representation and transferability of diabetic retinopathy classification models
Minhaj Nur Alam, Rikiya Yamashita, Vignav Ramesh +6
Self supervised contrastive learning based pretraining allows development of robust and generalized deep learning models with small, labeled datasets, reducing the burden of label…