1 paper · 1 filter
Jiaxuan Li, Junwen Mo, MinhDuc Vo +2
Multimodal Large Language Models (MLLMs) have made notable advances in visual understanding, yet their abilities to recognize objects modified by specific attributes remain an open…