3 papers
cs.CV2025
Vision-Language Memory for Spatial Reasoning
Zuntao Liu, Yi Du, Taimeng Fu +3
Spatial reasoning is a critical capability for intelligent robots, yet current vision-language models (VLMs) still fall short of human-level performance in video-based spatial reas…
cs.CV2025
Nonlinear Motion-Guided and Spatio-Temporal Aware Network for Unsupervised Event-Based Optical Flow
Zuntao Liu, Hao Zhuang, Junjie Jiang +2
Event cameras have the potential to capture continuous motion information over time and space, making them well-suited for optical flow estimation. However, most existing learning-…
cs.CV2024
Spike-EVPR: Deep Spiking Residual Networks with SNN-Tailored Representations for Event-Based Visual Place Recognition
Zuntao Liu, Yaohui Li, Chenming Hu +3
Event cameras are ideal for visual place recognition (VPR) in challenging environments due to their high temporal resolution and high dynamic range. However, existing methods conve…