3 citations · 3 across the 5 of their papers we have counts for
1 paper · 1 filter
Yuqi Yang, Peng-Tao Jiang, Jing Wang +4
Multi-modal large language models (MLLMs) can understand image-language prompts and demonstrate impressive reasoning ability. In this paper, we extend MLLMs' output by empowering M…