2 papers
cs.CV2024
SURDS: Benchmarking Spatial Understanding and Reasoning in Driving Scenarios with Vision Language Models
Xianda Guo, Ruijun Zhang, Yiqun Duan +7
Accurate spatial reasoning in outdoor environments - covering geometry, object pose, and inter-object relationships - is fundamental to downstream tasks such as mapping, motion for…
cs.CV2022
Learning 3D Semantics from Pose-Noisy 2D Images with Hierarchical Full Attention Network
Yuhang He, Lin Chen, Junkun Xie +1
We propose a novel framework to learn 3D point cloud semantics from 2D multi-view image observations containing pose error. On the one hand, directly learning from the massive, uns…