activity
20242026
collaborators

17 papers

cs.CV2026

Re-M3Dr: Rebalanced MultiModal Mean Deviation Regression

Haojie Yin, Chengcheng Feng, Tianyi Liu +2

Mean Deviation (MD) is a critical metric for assessing visual field loss in ophthalmology. While previous work has focused solely on predicting MD from Optical Coherence Tomography…

cs.RO2025

3D-CDRGP: Towards Cross-Device Robotic Grasping Policy in 3D Open World

Weiguang Zhao, Chenru Jiang, Chengrui Zhang +4

Given the diversity of devices and the product upgrades, cross-device research has become an urgent issue that needs to be tackled. To this end, we pioneer in probing the cross-dev…

cs.CV2025

Towards Training-Free Open-World Classification with 3D Generative Models

Xinzhe Xia, Weiguang Zhao, Yuyao Yan +4

3D open-world classification is a challenging yet essential task in dynamic and unstructured real-world scenarios, requiring both open-category and open-pose recognition. To addres…

cs.CV2025

HOMER: Homography-Based Efficient Multi-view 3D Object Removal

Jingcheng Ni, Weiguang Zhao, Daniel Wang +4

3D object removal is an important sub-task in 3D scene editing, with broad applications in scene understanding, augmented reality, and robotics. However, existing methods struggle…

cs.CV2025

Unlock Pose Diversity: Accurate and Efficient Implicit Keypoint-based Spatiotemporal Diffusion for Audio-driven Talking Portrait

Chaolong Yang, Kai Yao, Yuyao Yan +7

Audio-driven single-image talking portrait generation plays a crucial role in virtual reality, digital human creation, and filmmaking. Existing approaches are generally categorized…

cs.CV2025

BFANet: Revisiting 3D Semantic Segmentation with Boundary Feature Analysis

Weiguang Zhao, Rui Zhang, Qiufeng Wang +2

3D semantic segmentation plays a fundamental and crucial role to understand 3D scenes. While contemporary state-of-the-art techniques predominantly concentrate on elevating the ove…