5 papers
PartMat: Material-Aware 3D Part Decomposition with a Single Global Latent
Guangming Fu, Jin Song, Yiyun Fei +3
Part-level 3D generation has recently attracted increasing attention for producing structured and editable 3D assets. However, existing methods typically decompose objects accordin…
HomeDiffusion: Zero-Shot Object Customization with Multi-View Representation Learning for Indoor Scenes
Guoqiu Li, Jin Song, Yiyun Fei
Recently, zero-shot object customization generation methods have rapidly developed and shown tremendous potential for applications. For instance, in the e-commerce domain, consumer…
Home3D 1.0: A High-Fidelity Image-to-3D Asset Generation System for Interior Design
Yiyun Fei, Guoqiu Li, Jin Song +11
We present Home3D 1.0, a modular image-to-3D generation system that produces high-quality 3D assets from a single reference image, targeting interior design and e-commerce applicat…
Mantis: A Versatile Vision-Language-Action Model with Disentangled Visual Foresight
Yi Yang, Xueqi Li, Yiyang Chen +7
Recent advances in Vision-Language-Action (VLA) models demonstrate that visual signals can effectively complement sparse action supervisions. However, letting VLA directly predict…
Beyond Pixels: Benchmarking and Reward-Based Assessing Framework for Visual Spatial Aesthetics
Yuan Gao, Jin Song, Yiyun Fei +2
In recent years, Image Quality Assessment (IQA) for AI-generated images (AIGI) has advanced rapidly; however, existing methods primarily target portraits and artistic images, lacki…