2 papers
cs.RO2026
3DWay: Generalizing Robot Manipulation via 3D Consistent Waypoints
Ziqin Huang, Yingyue Li, Chenyangguang Zhang +6
Intermediate representations are key to bridging the modality gap between generalizable manipulation policies and large-scale pretrained vision-language models (VLMs). Among these,…
cs.CV2024
GFreeDet: Exploiting Gaussian Splatting and Foundation Models for Model-free Unseen Object Detection in the BOP Challenge 2024
Xingyu Liu, Gu Wang, Chengxi Li +4
We present GFreeDet, an unseen object detection approach that leverages Gaussian splatting and vision Foundation models under model-free setting. Unlike existing methods that rely…