2 papers
cs.SD2026
AuditoryHuM: Auditory Scene Label Generation and Clustering using Human-MLLM Collaboration
Henry Zhong, Jörg M. Buchholz, Julian Maclaren +2
Manual annotation of audio datasets is labour intensive, and it is challenging to balance label granularity with acoustic separability. We introduce AuditoryHuM, a novel framework…
cs.SD2026
A dataset and model for auditory scene recognition for hearing devices: AHEAD-DS and OpenYAMNet
Henry Zhong, Jörg M. Buchholz, Julian Maclaren +2
Scene recognition is important for hearing devices, however; this is challenging, in part because of the limitations of existing datasets. Datasets often lack public accessibility,…