1 paper
Tuo Zhang, Tiantian Feng, Yibin Ni +7
Large vision-language models (VLMs) have demonstrated remarkable abilities in understanding everyday content. However, their performance in the domain of art, particularly cultural…