8 citations · 11 across the 5 of their papers we have counts for
1 paper · 1 filter
Shuyang Li, Talha Azfar, Ruimin Ke
Large Language Models (LLMs), capable of handling multi-modal input and outputs such as text, voice, images, and video, are transforming the way we process information. Beyond just…