1 citations · 1 across the 11 of their papers we have counts for
1 paper · 1 filter
Saurabh Kataria, Xiao Hu
Audio-Language Models (ALMs) are making strides in understanding speech and non-speech audio. However, domain-specialist Foundation Models (FMs) remain the best for closed-ended sp…