1 citations · 3 across the 3 of their papers we have counts for
3 papers
cs.CV2024★ 1 cited
Leveraging Retrieval Augment Approach for Multimodal Emotion Recognition Under Missing Modalities
Qi Fan, Hongyu Yuan, Haolin Zuo +2
Multimodal emotion recognition utilizes complete multimodal information and robust multimodal joint representation to gain high performance. However, the ideal condition of full mo…
cs.MM2024★ 1 cited
MCDubber: Multimodal Context-Aware Expressive Video Dubbing
Yuan Zhao, Zhenqi Jia, Rui Liu +3
Automatic Video Dubbing (AVD) aims to take the given script and generate speech that aligns with lip motion and prosody expressiveness. Current AVD models mainly utilize visual inf…
cs.SD2023★ 1 cited
Betray Oneself: A Novel Audio DeepFake Detection Model via Mono-to-Stereo Conversion
Rui Liu, Jinhua Zhang, Guanglai Gao +1
Audio Deepfake Detection (ADD) aims to detect the fake audio generated by text-to-speech (TTS), voice conversion (VC) and replay, etc., which is an emerging topic. Traditionally we…