7 citations · 9 across the 6 of their papers we have counts for
1 paper · 2 filters
Yikai Gao, Ding Xia, Xi Yang
In domain-specific multimodal long documents, images and text jointly convey complex knowledge that cannot be fully captured by plain text alone. However, existing paradigms like D…