1 paper
Qi Zhi Lim, Chin Poo Lee, Kian Ming Lim +1
The increasing availability of multimodal data across text, tables, and images presents new challenges for developing models capable of complex cross-modal reasoning. Existing meth…