2 papers
cs.SD2025
Comprehend and Talk: Text to Speech Synthesis via Dual Language Modeling
Junjie Cao, Yichen Han, Ruonan Zhang +5
Existing Large Language Model (LLM) based autoregressive (AR) text-to-speech (TTS) systems, while achieving state-of-the-art quality, still face critical challenges. The foundation…
cs.CV2025
VideoGuard: Protecting Video Content from Unauthorized Editing
Junjie Cao, Kaizhou Li, Xinchun Yu +2
With the rapid development of generative technology, current generative models can generate high-fidelity digital content and edit it in a controlled manner. However, there is a ri…