1 paper
Qixuan Huang, Khalid Zaman, Masashi Unoki
Auditory large language models (ALLMs) have demonstrated strong general capabilities in audio understanding and reasoning tasks. However, their reliability is still undermined by h…