1 paper · 1 filter
Quanwei Tang, Zhiyu Tang, Xu Li +3
Although Speech Large Language Models (SpeechLLMs) excel at speech understanding and generation, their capacity for fine-grained, temporally aligned outputs remains underexplored.…