1 citations · 1 across the 2 of their papers we have counts for
2 papers
eess.AS2023
ViT-TTS: Visual Text-to-Speech with Scalable Diffusion Transformer
Huadai Liu, Rongjie Huang, Xuan Lin +5
Text-to-speech(TTS) has undergone remarkable improvements in performance, particularly with the advent of Denoising Diffusion Probabilistic Models (DDPMs). However, the perceived q…
cs.MM2022★ 1 cited
AntPivot: Livestream Highlight Detection via Hierarchical Attention Mechanism
Yang Zhao, Xuan Lin, Wenqiang Xu +3
In recent days, streaming technology has greatly promoted the development in the field of livestream. Due to the excessive length of livestream records, it's quite essential to ext…