2 papers
cs.CL2026
TableVista: Benchmarking Multimodal Table Reasoning under Visual and Structural Complexity
Zheyuan Yang, Liqiang Shang, Junjie Chen +6
We introduce TableVista, a comprehensive benchmark for evaluating foundation models in multimodal table reasoning under visual and structural complexity. TableVista consists of 3,0…
cs.LG2026
VisMMOE: Exploiting Visual-Expert Affinity for Efficient Visual-Language MoE Offloading
Cheng Xu, Xiaofeng Hou, Jiacheng Liu +1
Large-scale vision-language mixture-of-experts (VL-MoE) models provide strong multimodal capability, but efficient deployment on memory-constrained platforms remains difficult. Exi…