1 paper
Ming-Chang Chiu, Shicheng Wen, Pin-Yu Chen +1
In vision-language models (VLMs), the ability to perceive and interpret color and physical environment is crucial for achieving contextually accurate understanding and interaction.…