1 paper · 1 filter
Junliang Liu, Jingyu Xiao, Wenxin Tang +5
Multimodal large language models (MLLMs) are increasingly deployed as the core reasoning engine for web-facing systems, powering GUI agents and front-end automation that must inter…