6 citations · 7 across the 2 of their papers we have counts for
1 paper · 1 filter
Diji Yang, Kezhen Chen, Jinmeng Rao +4
Visual language tasks require AI models to comprehend and reason with both visual and textual content. Driven by the power of Large Language Models (LLMs), two prominent methods ha…