1 paper · 1 filter
Tanqiu Jiang, Min Bai, Nikolaos Pappas +2
Vision-language model (VLM)-based web agents increasingly power high-stakes selection tasks like content recommendation or product ranking by combining multimodal perception with p…