1 paper
Guiyao Tie, Xueyang Zhou, Tianhe Gu +7
Recent advances in Multi-Modal Large Language Models (MLLMs) have enabled unified processing of language, vision, and structured inputs, opening the door to complex tasks such as l…