3 papers
cs.SD2026
Multi-Domain Audio Question Answering Benchmark Toward Acoustic Content Reasoning
Chao-Han Huck Yang, Sreyan Ghosh, Qing Wang +14
We present Task 5 of the DCASE 2025 Challenge: an Audio Question Answering (AQA) benchmark spanning multiple domains of sound understanding. This task defines three QA subsets (Bio…
cs.SD2025
An Experimental Study on Joint Modeling for Sound Event Localization and Detection with Source Distance Estimation
Yuxuan Dong, Qing Wang, Hengyi Hong +2
In traditional sound event localization and detection (SELD) tasks, the focus is typically on sound event detection (SED) and direction-of-arrival (DOA) estimation, but they fall s…
eess.AS2024
MVANet: Multi-Stage Video Attention Network for Sound Event Localization and Detection with Source Distance Estimation
Hengyi Hong, Qing Wang, Jun Du +3
Sound event localization and detection with source distance estimation (3D SELD) involves not only identifying the sound category and its direction-of-arrival (DOA) but also predic…