7 papers
FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization
Quanjian Song, Yefeng Shen, Mengting Chen +5
Human-centric video customization, particularly at the garment level, has shown significant commercial value. However, existing approaches cannot support low-latency and interactiv…
IVGT: Implicit Visual Geometry Transformer for Neural Scene Representation
Yuqi Wu, Tianyu Hu, Wenzhao Zheng +4
Reconstructing coherent 3D geometry and appearance from unposed multi-view images is a fundamental yet challenging problem in computer vision. Most existing visual geometry foundat…
Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment
Jerry Jiang, Haowen Sun, Denis Gudovskiy +4
Spatial intelligence in vision-language models (VLMs) attracts research interest with the practical demand to reason in the 3D world.Despite promising results, most existing method…
PointVDP: Learning View-Dependent Projection by Fireworks Rays for 3D Point Cloud Segmentation
Yang Chen, Yueqi Duan, Haowen Sun +3
In this paper, we propose view-dependent projection (VDP) to facilitate point cloud segmentation, designing efficient 3D-to-2D mapping that dynamically adapts to the spatial geomet…
Ambiguity-aware Point Cloud Segmentation by Adaptive Margin Contrastive Learning
Yang Chen, Yueqi Duan, Haowen Sun +2
This paper proposes an adaptive margin contrastive learning method for 3D semantic segmentation on point clouds. Most existing methods use equally penalized objectives, which ignor…
DreamCinema: Cinematic Transfer with Free Camera and 3D Character
Weiliang Chen, Fangfu Liu, Diankun Wu +3
We are living in a flourishing era of digital media, where everyone has the potential to become a personal filmmaker. Current research on video generation suggests a promising aven…