1 paper
Xinyuan Xie, Shunian Chen, Zhiheng Liu +4
Large Audio Language Models (LALMs) still struggle in complex acoustic scenes because they often fail to preserve task-relevant acoustic evidence before reasoning begins. We identi…