7 citations · 9 across the 5 of their papers we have counts for
1 paper · 1 filter
Yikai Gao, Ding Xia, Xi Yang
In domain-specific multimodal long documents, images and text jointly convey complex knowledge that cannot be fully captured by plain text alone. However, existing paradigms like D…