9 citations · 9 across the 4 of their papers we have counts for
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2022★ 9 cited
Interactive Audio-text Representation for Automated Audio Captioning with Contrastive Learning
Chen Chen, Nana Hou, Yuchen Hu +3
Automated Audio captioning (AAC) is a cross-modal task that generates natural language to describe the content of input audio. Most prior works usually extract single-modality acou…
cs.SD2022
Speech Emotion Recognition with Co-Attention based Multi-level Acoustic Information
Heqing Zou, Yuke Si, Chen Chen +2
Speech Emotion Recognition (SER) aims to help the machine to understand human's subjective emotion from only audio information. However, extracting and utilizing comprehensive in-d…
cs.SD2022
Noise-robust Speech Recognition with 10 Minutes Unparalleled In-domain Data
Chen Chen, Nana Hou, Yuchen Hu +2
Noise-robust speech recognition systems require large amounts of training data including noisy speech data and corresponding transcripts to achieve state-of-the-art performances in…