1 citations · 1 across the 9 of their papers we have counts for
1 paper · 1 filter
Yanxi Chen, Wenhui Zhu, Xiwen Chen +9
Although Large Audio-Language Models (LALMs) deliver state-of-the-art (SOTA) performance, they frequently suffer from hallucinations, e.g. generating text not grounded in the audio…