1 paper
Wenfang Sun, Hao Chen, Yingjun Du +2
Large vision-language models have achieved remarkable progress in visual reasoning, yet most existing systems rely on single-step or text-only reasoning, limiting their ability to…