1 paper
Zefeng Wang, Zhen Han, Shuo Chen +6
Multimodal LLMs (MLLMs) with a great ability of text and image understanding have received great attention. To achieve better reasoning with MLLMs, Chain-of-Thought (CoT) reasoning…