Showing cs.SDShow all
3 papers · 1 filter
cs.SD2026
Multi-Domain Audio Question Answering Benchmark Toward Acoustic Content Reasoning
Chao-Han Huck Yang, Sreyan Ghosh, Qing Wang +14
We present Task 5 of the DCASE 2025 Challenge: an Audio Question Answering (AQA) benchmark spanning multiple domains of sound understanding. This task defines three QA subsets (Bio…
cs.SD2025
Improving Anomalous Sound Detection with Attribute-aware Representation from Domain-adaptive Pre-training
Xin Fang, Guirui Zhong, Qing Wang +7
Anomalous Sound Detection (ASD) is often formulated as a machine attribute classification task, a strategy necessitated by the common scenario where only normal data is available f…
cs.SD2025
An Enhanced Audio Feature Tailored for Anomalous Sound Detection Based on Pre-trained Models
Guirui Zhong, Qing Wang, Jun Du +3
Anomalous Sound Detection (ASD) aims at identifying anomalous sounds from machines and has gained extensive research interests from both academia and industry. However, the uncerta…