29 citations · 67 across the 48 of their papers we have counts for
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2026
SpotSound: Enhancing Large Audio-Language Models with Fine-Grained Temporal Grounding
Luoyi Sun, Xiao Zhou, Zeqian Li +3
Large Audio-Language Models (ALMs) have recently demonstrated remarkable capabilities in holistic audio understanding, yet they remain unreliable for temporal grounding, i.e., the…
cs.SD2024
A Generalist Audio Foundation Model for Comprehensive Body Sound Auscultation
Pingjie Wang, Liudan Zhao, Zihan Zhao +6
Accurate and efficient auscultation-based diagnostics are vital for early disease detection, especially in resource-limited settings where specialized clinical expertise is scarce.…
cs.SD2024★ 1 cited
HSDreport: Heart Sound Diagnosis with Echocardiography Reports
Zihan Zhao, Pingjie Wang, Liudan Zhao +7
Heart sound auscultation holds significant importance in the diagnosis of congenital heart disease. However, existing methods for Heart Sound Diagnosis (HSD) tasks are predominantl…