1 paper
Seyed Amir Kasaei, Arash Marioriyad, Mahbod Khaleti +3
Large Vision-Language Models (LVLMs) have achieved remarkable proficiency in explicit visual recognition, effectively describing what is directly visible in an image. However, a cr…