2 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.CV2023
Learning to Compress Unmanned Aerial Vehicle (UAV) Captured Video: Benchmark and Analysis
Chuanmin Jia, Feng Ye, Huifang Sun +2
During the past decade, the Unmanned-Aerial-Vehicles (UAVs) have attracted increasing attention due to their flexible, extensive, and dynamic space-sensing capabilities. The volume…
cs.CL2022★ 2 cited
Multilevel Transformer For Multimodal Emotion Recognition
Junyi He, Meimei Wu, Meng Li +2
Multimodal emotion recognition has attracted much attention recently. Fusing multiple modalities effectively with limited labeled data is a challenging task. Considering the succes…
cs.SD2022
MHTTS: Fast multi-head text-to-speech for spontaneous speech with imperfect transcription
Dabiao Ma, Yitong Zhang, Meng Li +1
Neural network based end-to-end Text-to-Speech (TTS) has greatly improved the quality of synthesized speech. While how to use massive spontaneous speech without transcription effic…