1 paper · 1 filter
Haoyu Zhang, Yangyang Guo, Mohan Kankanhalli
Large Vision-Language Models (LVLMs) exhibit remarkable vision-language capabilities and are increasingly deployed in real-world applications such as personal assistants, document…