1 paper
Hongjin Song, Jiasheng Kuang, Xinyu Yang +4
Large audio-language models may mention acoustic events that are absent from the input. A separate audio event detector can verify these mentions, but doing so requires a second au…