activity
20232026
most citedMOHO: Learning Single-view Hand-held Object Reconstruction with Multi-view Occlusion-Aware Supervision

1 citations · 1 across the 5 of their papers we have counts for

collaborators

5 papers

cs.RO2026

3DWay: Generalizing Robot Manipulation via 3D Consistent Waypoints

Ziqin Huang, Yingyue Li, Chenyangguang Zhang +6

Intermediate representations are key to bridging the modality gap between generalizable manipulation policies and large-scale pretrained vision-language models (VLMs). Among these,…

cs.CV2024

GFreeDet: Exploiting Gaussian Splatting and Foundation Models for Model-free Unseen Object Detection in the BOP Challenge 2024

Xingyu Liu, Gu Wang, Chengxi Li +4

We present GFreeDet, an unseen object detection approach that leverages Gaussian splatting and vision Foundation models under model-free setting. Unlike existing methods that rely…

cs.CV2024

LaPose: Laplacian Mixture Shape Modeling for RGB-Based Category-Level Object Pose Estimation

Ruida Zhang, Ziqin Huang, Gu Wang +5

While RGBD-based methods for category-level object pose estimation hold promise, their reliance on depth data limits their applicability in diverse scenarios. In response, recent e…

cs.CV2023

D-SCo: Dual-Stream Conditional Diffusion for Monocular Hand-Held Object Reconstruction

Bowen Fu, Gu Wang, Chenyangguang Zhang +6

Reconstructing hand-held objects from a single RGB image is a challenging task in computer vision. In contrast to prior works that utilize deterministic modeling paradigms, we empl…

cs.CV20231 cited

MOHO: Learning Single-view Hand-held Object Reconstruction with Multi-view Occlusion-Aware Supervision

Chenyangguang Zhang, Guanlong Jiao, Yan Di +7

Previous works concerning single-view hand-held object reconstruction typically rely on supervision from 3D ground-truth models, which are hard to collect in real world. In contras…