1 paper
Yuki Toi, Tao Xiao, Kazushi Tomoto +2
In recent years, Large Language Models (LLMs) have made significant strides, leading to the emergence of multimodal LLMs capable of processing diverse inputs such as images and aud…