5 citations · 10 across the 7 of their papers we have counts for
Showing 2023 · eess.ASShow all
2 papers · 2 filters
eess.AS2023★ 1 cited
Interactive Dual-Conformer with Scene-Inspired Mask for Soft Sound Event Detection
Han Yin, Jisheng Bai, Mou Wang +3
Traditional binary hard labels for sound event detection (SED) lack details about the complexity and variability of sound event distributions. Recently, a novel annotation workflow…
eess.AS2023★ 1 cited
AudioLog: LLMs-Powered Long Audio Logging with Hybrid Token-Semantic Contrastive Learning
Jisheng Bai, Han Yin, Mou Wang +4
Previous studies in automated audio captioning have faced difficulties in accurately capturing the complete temporal details of acoustic scenes and events within long audio sequenc…