4 citations · 4 across the 1 of their papers we have counts for
1 paper · 1 filter
Shiyin Lu, Yang Li, Qing-Guo Chen +4
Current Multimodal Large Language Models (MLLMs) typically integrate a pre-trained LLM with another pre-trained vision transformer through a connector, such as an MLP, endowing the…