1 paper
Xinlu He, Swayambhu Nath Ray, Harish Mallidi +5
Unified architectures in multimodal large language models (MLLM) have shown promise in handling diverse tasks within a single framework. In the text-to-speech (TTS) task, current M…