1 paper · 1 filter
Shusen Liu, Haichao Miao, Zhimin Li +3
With recent advances in multi-modal foundation models, the previously text-only large language models (LLM) have evolved to incorporate visual input, opening up unprecedented oppor…