1 paper
Yilin Zhang, Yingkai Hua, Chunyu Wei +2
Vision-language model (VLM) based web agents demonstrate impressive autonomous GUI interaction but remain vulnerable to deceptive interface elements. Existing approaches either det…