1 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.CV2024★ 1 cited
Enriching Multimodal Sentiment Analysis through Textual Emotional Descriptions of Visual-Audio Content
Sheng Wu, Xiaobao Wang, Longbiao Wang +2
Multimodal Sentiment Analysis (MSA) stands as a critical research frontier, seeking to comprehensively unravel human emotions by amalgamating text, audio, and visual data. Yet, dis…
cs.CL2024
Error Correction by Paying Attention to Both Acoustic and Confidence References for Automatic Speech Recognition
Yuchun Shu, Bo Hu, Yifeng He +3
Accurately finding the wrong words in the automatic speech recognition (ASR) hypothesis and recovering them well-founded is the goal of speech error correction. In this paper, we p…
cs.SD2024★ 1 cited
ICMC-ASR: The ICASSP 2024 In-Car Multi-Channel Automatic Speech Recognition Challenge
He Wang, Pengcheng Guo, Yue Li +13
To promote speech processing and recognition research in driving scenarios, we build on the success of the Intelligent Cockpit Speech Recognition Challenge (ICSRC) held at ISCSLP 2…