1 citations · 1 across the 9 of their papers we have counts for
1 paper · 1 filter
Yiwen Guan, Viet Anh Trinh, Vivek Voleti +1
Recent advances in multi-modal large language models (MLLMs) have opened new possibilities for unified modeling of speech, text, images, and other modalities. Building on our prior…