3 papers
cs.CV2026
UniGP: Taming Diffusion Transformer for Prior-Preserved Unified Generation and Perception
Qin Guo, Hao Luo, Dongxu Yue +4
Recent advances in diffusion models have shown impressive performance in controllable image generation and dense prediction tasks. However, existing approaches typically treat diff…
cs.CV2025
JCo-MVTON: Jointly Controllable Multi-Modal Diffusion Transformer for Mask-Free Virtual Try-on
Aowen Wang, Wei Li, Hao Luo +4
Virtual try-on systems have long been hindered by heavy reliance on human body masks, limited fine-grained control over garment attributes, and poor generalization to real-world, i…
cs.CV2025
Coherent Video Inpainting Using Optical Flow-Guided Efficient Diffusion
Bohai Gu, Hao Luo, Song Guo +2
The text-guided video inpainting technique has significantly improved the performance of content generation applications. A recent family for these improvements uses diffusion mode…