5 papers
RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation
Lexi Pang, Liheng Zhang, Hang Ye +2
Text-to-image (T2I) diffusion models have shown remarkable success in generating high-quality images from text prompts. Recent efforts extend these models to incorporate conditiona…
Visually-grounded Humanoid Agents
Hang Ye, Xiaoxuan Ma, Fan Lu +3
Digital human generation has been studied for decades and supports a wide range of real-world applications. However, most existing systems are passively animated, relying on privil…
GeneMAN: Generalizable Single-Image 3D Human Reconstruction from Multi-Source Human Data
Wentao Wang, Hang Ye, Fangzhou Hong +5
Given a single in-the-wild human photo, it remains a challenging task to reconstruct a high-fidelity 3D human model. Existing methods face difficulties including a) the varying bod…
PinPoint3D: Fine-Grained 3D Part Segmentation from a Few Clicks
Bojun Zhang, Hangjian Ye, Hao Zheng +4
Fine-grained 3D part segmentation is crucial for enabling embodied AI systems to perform complex manipulation tasks, such as interacting with specific functional components of an o…
FreeCloth: Free-form Generation Enhances Challenging Clothed Human Modeling
Hang Ye, Xiaoxuan Ma, Hai Ci +2
Achieving realistic animated human avatars requires accurate modeling of pose-dependent clothing deformations. Existing learning-based methods heavily rely on the Linear Blend Skin…