1 paper
Masao Someki, Chien-yu Huang, Siddhant Arora +7
Long-form audio understanding poses significant challenges for large audio language models (LALMs) due to the extreme length of audio sequences and the need to reason over heteroge…