1 paper · 1 filter
Songyan Zhang, Wenhui Huang, Zihui Gao +2
The emergence of general human knowledge and impressive logical reasoning capacity in rapidly progressed vision-language models (VLMs) have driven increasing interest in applying V…