1 paper · 1 filter
Yuliang Zhan, Xinyu Tang, Han Wan +3
Recently, Chain-of-Thought (CoT) reasoning has significantly enhanced the capabilities of large language models (LLMs), but Vision-Language Models (VLMs) still struggle with multi-…