2 papers
cs.RO2026
VLingNav: Embodied Navigation with Adaptive Reasoning and Visual-Assisted Linguistic Memory
Shaoan Wang, Yuanfei Luo, Xingyu Chen +6
VLA models have shown promising potential in embodied navigation by unifying perception and planning while inheriting the strong generalization abilities of large VLMs. However, mo…
cs.CV2025
LET-US: Long Event-Text Understanding of Scenes
Rui Chen, Xingyu Chen, Shaoan Wang +2
Event cameras output event streams as sparse, asynchronous data with microsecond-level temporal resolution, enabling visual perception with low latency and a high dynamic range. Wh…