3 papers
cs.CV2026
HomeDiffusion: Zero-Shot Object Customization with Multi-View Representation Learning for Indoor Scenes
Guoqiu Li, Jin Song, Yiyun Fei
Recently, zero-shot object customization generation methods have rapidly developed and shown tremendous potential for applications. For instance, in the e-commerce domain, consumer…
cs.CV2026
Home3D 1.0: A High-Fidelity Image-to-3D Asset Generation System for Interior Design
Yiyun Fei, Guoqiu Li, Jin Song +11
We present Home3D 1.0, a modular image-to-3D generation system that produces high-quality 3D assets from a single reference image, targeting interior design and e-commerce applicat…
cs.CV2025
Mantis: A Versatile Vision-Language-Action Model with Disentangled Visual Foresight
Yi Yang, Xueqi Li, Yiyang Chen +7
Recent advances in Vision-Language-Action (VLA) models demonstrate that visual signals can effectively complement sparse action supervisions. However, letting VLA directly predict…