2 papers
cs.CV2025
VPN: Visual Prompt Navigation
Shuo Feng, Zihan Wang, Yuchen Li +6
While natural language is commonly used to guide embodied agents, the inherent ambiguity and verbosity of language often hinder the effectiveness of language-guided navigation in c…
cs.CV2025
Improving Brain-to-Image Reconstruction via Fine-Grained Text Bridging
Runze Xia, Shuo Feng, Renzhi Wang +3
Brain-to-Image reconstruction aims to recover visual stimuli perceived by humans from brain activity. However, the reconstructed visual stimuli often missing details and semantic i…