2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CV2026
Visual Attention Faithfulness in Vision-Language Models is Heterogeneous
Xurui Song, Weishi Wang, Zhongqi Yue +5
Whether attention weights faithfully reflect model reasoning has been actively debated in NLP, yet this question remains largely unexplored for the visual modality in Vision-Langua…
cs.CL2025★ 2 cited
Document Intelligence in the Era of Large Language Models: A Survey
Weishi Wang, Hengchang Hu, Zhijie Zhang +3
Document AI (DAI) has emerged as a vital application area, and is significantly transformed by the advent of large language models (LLMs). While earlier approaches relied on encode…