1 paper
Jing Peng, Zichao Nie, Zhisheng Zhang +2
Large Audio Language Models (LALMs) utilize either continuous features or discrete tokens, yet the optimal representation paradigm for general audio understanding remains debated.…