1 paper · 1 filter
Shivam Mehta, Nebojsa Jojic, Hannes Gamper
Integrating audio comprehension and generation into large language models (LLMs) remains challenging due to the continuous nature of audio and the resulting high sampling rates. He…