1 paper · 1 filter
Anju Gopinath, Nikhil Krishnaswamy, Bruce Draper
Multimodal Large Language Models (MLLMs) struggle with tasks that require reasoning about 2D object orientation in images, as documented in prior work. Tong et al. and Nichols et a…