1 paper
Xiaokun Sun, Yubo Wang, Haoyu Cao +1
Recently, Multimodal Large Language Models (MLLMs) have demonstrated significant potential in complex visual tasks through the integration of Chain-of-Thought (CoT) reasoning. Howe…