3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.SD2024
Prior-agnostic Multi-scale Contrastive Text-Audio Pre-training for Parallelized TTS Frontend Modeling
Quanxiu Wang, Hui Huang, Mingjie Wang +3
Over the past decade, a series of unflagging efforts have been dedicated to developing highly expressive and controllable text-to-speech (TTS) systems. In general, the holistic TTS…
cs.CV2023★ 3 cited
Fine-grained Text and Image Guided Point Cloud Completion with CLIP Model
Wei Song, Jun Zhou, Mingjie Wang +3
This paper focuses on the recently popular task of point cloud completion guided by multimodal information. Although existing methods have achieved excellent performance by fusing…