2 papers
cs.CV2024
Human-VDM: Learning Single-Image 3D Human Gaussian Splatting from Video Diffusion Models
Zhibin Liu, Haoye Dong, Aviral Chharia +1
Generating lifelike 3D humans from a single RGB image remains a challenging task in computer vision, as it requires accurate modeling of geometry, high-quality texture, and plausib…
cs.CV2023
Contrastive Transformer Learning with Proximity Data Generation for Text-Based Person Search
Hefeng Wu, Weifeng Chen, Zhibin Liu +3
Given a descriptive text query, text-based person search (TBPS) aims to retrieve the best-matched target person from an image gallery. Such a cross-modal retrieval task is quite ch…