3 citations · 3 across the 5 of their papers we have counts for
1 paper · 1 filter
Yuanchun Shen, Ruotong Liao, Zhen Han +2
While multi-modal models have successfully integrated information from image, video, and audio modalities, integrating graph modality into large language models (LLMs) remains unex…