1 paper
Jeongwoo Lee, Baek Duhyeong, Eungyeol Han +5
Recent advances in Vision-Language Models (VLMs) have demonstrated impressive multimodal understanding in general domains. However, their applicability to decision-oriented domains…