2 papers
cs.CV2026
NavOne: One-Step Global Planning for Vision-Language Navigation on Top-Down Maps
Dijia Zhan, Jinyi Li, Chenxi Zheng +4
Existing Vision-Language Navigation (VLN) methods typically adopt an egocentric, step-by-step paradigm, which struggles with error accumulation and limits efficiency. While recent…
cs.CV2025
StarPose: 3D Human Pose Estimation via Spatial-Temporal Autoregressive Diffusion
Haoxin Yang, Weihong Chen, Xuemiao Xu +5
Monocular 3D human pose estimation remains a challenging task due to inherent depth ambiguities and occlusions. Compared to traditional methods based on Transformers or Convolution…