5 citations · 10 across the 7 of their papers we have counts for
Showing 2024 · eess.ASShow all
3 papers · 2 filters
eess.AS2024
AudioSetCaps: An Enriched Audio-Caption Dataset using Automated Generation Pipeline with Large Audio and Language Models
Jisheng Bai, Haohe Liu, Mou Wang +5
With the emergence of audio-language models, constructing large-scale paired audio-language datasets has become essential yet challenging for model development, primarily due to th…
eess.AS2024★ 5 cited
Description on IEEE ICME 2024 Grand Challenge: Semi-supervised Acoustic Scene Classification under Domain Shift
Jisheng Bai, Mou Wang, Haohe Liu +11
Acoustic scene classification (ASC) is a crucial research problem in computational auditory scene analysis, and it aims to recognize the unique acoustic characteristics of an envir…
eess.AS2024
Sub-band and Full-band Interactive U-Net with DPRNN for Demixing Cross-talk Stereo Music
Han Yin, Mou Wang, Jisheng Bai +3
This paper presents a detailed description of our proposed methods for the ICASSP 2024 Cadenza Challenge. Experimental results show that the proposed system can achieve better perf…