Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
AffordBot: 3D Fine-grained Embodied Reasoning via Multimodal Large Language Models
Xinyi Wang, Xun Yang, Yanlong Xu +3
Effective human-agent collaboration in physical environments requires understanding not only what to act upon, but also where the actionable elements are and how to interact with t…
cs.CV2025
CMMLoc: Advancing Text-to-PointCloud Localization with Cauchy-Mixture-Model Based Framework
Yanlong Xu, Haoxuan Qu, Jun Liu +2
The goal of point cloud localization based on linguistic description is to identify a 3D position using textual description in large urban environments, which has potential applica…