2 papers
cs.AI2025
WebSynthesis: World-Model-Guided MCTS for Efficient WebUI-Trajectory Synthesis
Yifei Gao, Junhong Ye, Jiaqi Wang +1
Recent advancements in large language models (LLMs) have significantly improved the capabilities of web agents. However, effectively navigating complex and dynamic web environments…
cs.CV2025
VaLiD: Mitigating the Hallucination of Large Vision Language Models by Visual Layer Fusion Contrastive Decoding
Jiaqi Wang, Yifei Gao, Jitao Sang
Large Vision-Language Models (LVLMs) have demonstrated remarkable capabilities in multimodal task reasoning. However, they often generate responses that appear plausible yet do not…