2 papers
cs.CV2025
Duoduo CLIP: Efficient 3D Understanding with Multi-View Images
Han-Hung Lee, Yiming Zhang, Angel X. Chang
We introduce Duoduo CLIP, a model for 3D representation learning that learns shape encodings from multi-view images instead of point clouds. The choice of multi-view images allows…
cs.CV2025
S2O: Static to Openable Enhancement for Articulated 3D Objects
Denys Iliash, Hanxiao Jiang, Yiming Zhang +2
Despite much progress in large 3D datasets there are currently few interactive 3D object datasets, and their scale is limited due to the manual effort required in their constructio…