3 papers
cs.CV2025
Resolving Ambiguity in Gaze-Facilitated Visual Assistant Interaction Paradigm
Zeyu Wang, Baiyu Chen, Kun Yan +5
With the rise in popularity of smart glasses, users' attention has been integrated into Vision-Language Models (VLMs) to streamline multi-modal querying in daily scenarios. However…
cs.HC2025
When Ads Become Profiles: Uncovering the Invisible Risk of Web Advertising at Scale with LLMs
Baiyu Chen, Benjamin Tag, Hao Xue +2
Regulatory limits on explicit targeting have not eliminated algorithmic profiling on the Web, as optimisation systems still adapt ad delivery to users' private attributes. The wide…
cs.CL2025
Multi-Stage Verification-Centric Framework for Mitigating Hallucination in Multi-Modal RAG
Baiyu Chen, Wilson Wongso, Xiaoqian Hu +2
This paper presents the technical solution developed by team CRUISE for the KDD Cup 2025 Meta Comprehensive RAG Benchmark for Multi-modal, Multi-turn (CRAG-MM) challenge. The chall…